Senior Cloud Reliability Engineer (SRE)

Horizontal Talent
Richmond, VA, United States
10 days ago
Apply on www.jofdav.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Cloud Computing DevOps Python (Programming Language) Cloud Services Software Engineering Software Modules System Availability
+7 more
Grafana Amazon Virtual Private Cloud (VPC) Functional Programming Cloudwatch Terraform Golang Programming Languages

Job description

Join our team as a Senior Cloud Reliability Engineer, where you’ll play a pivotal role in developing scalable cloud solutions and enhancing platform reliability. This position offers the chance to work with cutting-edge technologies and collaborate with diverse teams in a dynamic environment. Responsibilities

  • Design and develop cloud solutions to improve platform reliability and reduce operational toil.
  • Build and optimize Infrastructure as Code using Terraform for efficient AWS resource management.
  • Develop CI/CD pipelines and automated testing frameworks to ensure high-quality code delivery.
  • Establish and implement SRE standards and metrics, including SLI and SLOs.
  • Participate in incident management and provide technical support for SRE tools.
  • Collaborate with cross-functional teams within Agile frameworks to deliver integrated cloud automation solutions.
  • Stay updated with emerging AWS services and drive the adoption of innovative solutions.

Requirements

  • Extensive experience in software development with a focus on reliability and platform engineering.
  • Advanced proficiency in Python for building enterprise-grade tools and APIs.
  • Strong expertise in AWS environments and core services like EC2, VPC, S3, and Lambda.
  • Proficiency in Infrastructure as Code using Terraform, including module development.
  • Experience with CI/CD pipelines and DevOps practices.
  • Knowledge of observability tools and practices, including Grafana and AWS CloudWatch.

Preferred Skills

  • Experience with GoLang or other programming languages.
  • Familiarity with Agile and Scaled Agile environments.
  • Understanding of ITSM processes and resilience testing practices.

About the company

At Horizontal, we are dedicated to fostering a diverse and inclusive workplace where all individuals are valued and respected. We believe that diversity drives innovation and success, and we welcome candidates from all backgrounds to apply.

By applying for this position, you acknowledge and agree that Horizontal Talent may contact you regarding your application using automated technology, including phone calls, SMS/text messages, or email, which may be delivered by our virtual AI recruiter, Alex. Other about 4 hours ago

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jofdav.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

6:16 min

Event-driven Golang backend architecture and cloud deployment

Irina Branovic Irina Branovic · World Congress 2026 Europe

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

Videos

See all

Related articles

See all