Senior Site Reliability Engineer in Mountain View

Energy Jobline
Mountain View, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours

Tech stack

Distributed Systems Python (Programming Language) Reliability Engineering Message Oriented Middleware Containerization Information Technology Terraform

Requirements

  • +4 years of relevant experience bringing software to production at high scale
  • Participation in on-call rotation, triaging and addressing production issues
  • Obsession with automation and instrumentation
  • Understanding of complex systems and failure scenarios
  • Excellent communication skills
  • Knowledge of AWS services, containers and container management frameworks
  • Familiarity with Message Bus based systems and distributed architectures
  • Proficiency in Terraform , Python and/or Go

What we’d like to see

  • BS or MS degree in the Computer Science field, or equivalent hands-on experience.
  • Experience in product oriented environments
  • Scalable distributed applications experience

Benefits & conditions

  • Competitive compensation with stock options
  • Comprehensive medical, vision, and dental insurance
  • 401k matching
  • Fitness and wellness stipend
  • Mobile phone reimbursement
  • Mental well-being benefits
  • Professional learning and development stipend
  • Parental leave, including adoptive and foster parents
  • 3 weeks paid time off (increases with tenure) and unlimited sick leave

About the company

Job DescriptionJob DescriptionAt ASAPP, our mission is simple: deliver the best AI-powered customer experience-faster than anyone else. To achieve that, we’re guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed, ownership, and a relentless focus on outcomes. We work in tight, skilled teams, prioritize clarity over complexity, and continuously evolve through curiosity, data, and craftsmanship. We’re seeking technologists and problem solvers who thrive in fast-paced environments, love collaborating with great talent, and approach every day like it’s Day 1. We’re a globally diverse team with hubs in New York City, Mountain View, Latin America, and India-embracing both hybrid and remote work to bring the best minds together, wherever they are. If you’re driven by continuous learning, rapid pivots, and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job-it’s a journey.

Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP’s infrastructure and products. The team owns the entire infrastructure stacks. SREs design and implement the tools that automate building reliable and performant systems. We emphasize building tools over manual processes. We implement, not administer. We’re obsessed with automation, not repetition. Our job is to focus on building reliable infrastructure and tools for our product teams so that they can solve customer problems and deliver new features, not reinvent platforms.

What you’ll do

  • Work with product engineering teams on service architecture and implementation
  • Deliver Infrastructure configuration as code and automate everything
  • Direct and implement monitoring and alerting systems to support rapid problem diagnosis
  • Perform Root Cause Analysis and design and deliver resolutions
  • Work on our Kubernetes / AWS infrastructure to support our product engineers
  • Design secure and performant networking solutions in our production systems

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.energyjobline.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

2:41 min

Pushing and hosting containerized applications in Azure

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

1:29 min

Overcoming challenges in AI-assisted distributed system development

Przemysław Ładyński Przemysław Ładyński · WWC Europe 2026

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all