Site Reliability Engineer

CARE AND RESPITE ENTERPRISES, INC.
Portsmouth, NH, United States
20 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$105,000.0 - $115,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Systems Engineering Microsoft Azure Bash Shell Cloud Computing Continuous Integration DevOps Disaster Recovery Github Python (Programming Language) Performance Tuning Reliability Engineering
+19 more
Prometheus Systems Architecture Datadog Data Logging Scripting Google Cloud Cloud Platform System Delivery Pipeline Grafana Reliability of Systems Cloudformation Gitlab-ci Kubernetes Infrastructure Automation Frameworks Information Technology Terraform Splunk Docker Jenkins

Job description

Careerscape is supporting a client opening for a Hybrid Site Reliability Engineer. This role focuses on maintaining highly available systems, improving infrastructure reliability, automating operational processes, monitoring production environments, and ensuring scalable cloud-based services operate efficiently in a hybrid work environment. This is an excellent opportunity for someone with strong technical skills who is passionate about automation, cloud technologies, system reliability, and continuous improvement. The Site Reliability Engineer will work closely with software engineers, DevOps teams, security professionals, and infrastructure teams to optimize system performance, resolve production issues, improve deployment pipelines, and build reliable, scalable platforms. This role is ideal for candidates interested in cloud infrastructure, DevOps engineering, platform engineering, or production operations. Responsibilities

  • Monitor production systems and cloud infrastructure
  • Improve application availability, reliability, and scalability
  • Build and maintain infrastructure automation solutions
  • Manage CI/CD pipelines and deployment processes
  • Respond to incidents and perform root cause analysis
  • Optimize system performance and resource utilization
  • Develop monitoring, logging, and alerting solutions
  • Collaborate with engineering teams to improve system architecture
  • Maintain infrastructure as code using automation tools
  • Support disaster recovery and business continuity initiatives
  • Implement security and reliability best practices
  • Perform additional infrastructure and platform engineering duties as assigned

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field preferred
  • 2-5 years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Infrastructure, or related roles
  • Strong knowledge of Linux system administration
  • Experience with AWS, Microsoft Azure, or Google Cloud Platform
  • Experience with Kubernetes, Docker, and container orchestration
  • Proficiency with Infrastructure as Code tools such as Terraform or CloudFormation
  • Experience with CI/CD tools including GitHub Actions, Jenkins, or GitLab CI
  • Knowledge of monitoring platforms such as Prometheus, Grafana, Datadog, or Splunk
  • Strong scripting skills using Python, Bash, or Go
  • Excellent troubleshooting, communication, and collaboration skills

Benefits & conditions

  • Hybrid work flexibility within the United States
  • Competitive compensation package
  • Medical, dental, and vision insurance
  • Paid time off, holidays, and sick leave
  • 401(k) retirement savings plan with company match
  • Annual performance bonuses
  • Professional development and certification reimbursement
  • Paid training and technical conference opportunities
  • Career growth into Senior Site Reliability Engineer, Platform Engineer, DevOps Architect, Infrastructure Manager, or Cloud Engineering leadership roles
  • Collaborative and innovative engineering environment

About the company

Planet Fitness

  • Hampton, NH About Us: Founded in 1992 in Dover, NH, Planet Fitness is one of the largest and fastest-growing franchisors and operators of fitness centers in the world by number of members an…, About Us Founded in 1992 in Dover, NH, Planet Fitness is one of the largest and fastest-growing franchisors and operators of fitness centers in the world by number of members and…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all