Senior Site Reliability Engineer (SRE)

LeoLabs, Inc.
San Francisco, CA, United States
29 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$192,000.0
Working hours
Shift work
Job source

Tech stack

Amazon Web Services Microsoft Azure Databases Continuous Integration DevOps Distributed Systems Github Monitoring of Systems Python (Programming Language) PostgreSQL Reliability Engineering Cloud Services
+16 more
Software Deployment Systems Architecture Datadog Circleci Scripting Delivery Pipeline Grafana Reliability of Systems Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Terraform Docker Programming Languages Microservices

Job description

LeoLabs is seeking a skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will bridge the gap between development and operations, ensuring that our systems are scalable, reliable, and efficient. You will be responsible for automating processes, monitoring system performance, and resolving incidents to enhance our service reliability., * System Reliability: Design, implement, and maintain scalable and reliable systems.

  • Monitoring and Incident Response: Set up monitoring tools and create incident response plans to quickly identify and resolve issues, as well as implementing preventative measures.
  • Automation: Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks.
  • Capacity Planning: Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions.
  • Collaboration: Work closely with development teams to enhance product reliability and streamline the deployment process.
  • Documentation: Create and maintain documentation for system architecture, processes, and incident reports.
  • On-Call Support: Participate in on-call rotations to provide 24/7 support for critical systems.

Security: Implement and enforce security best practices across all systems, ensuring compliance with industry standards., * Complete onboarding to understand our business, vision, and team structure.

  • Get familiar with LeoLabs’ engineering stack, security posture, and key initiatives.
  • Gain an understanding about how your role fits into LeoLabs broader organization.

Within 3 months, you’ll:

  • Independently deploy infrastructure changes using Infrastructure as Code.
  • Identify key reliability risks and recommend improvements.
  • Improve dashboards, alerts, and operational runbooks.

Within 6 months, you’ll:

  • Improve deployment pipelines, infrastructure provisioning, or self-service capabilities.
  • Optimize infrastructure utilization and cloud costs without compromising reliability.
  • Drive automation that reduces operational toil and improves deployment reliability.

Within 12 months, you’ll:

  • Lead cross-functional initiatives to improve availability, scalability, and operational efficiency.
  • Be a key advisor for site reliability in new product developments and platform evolution.
  • Mentor junior engineers and foster the development culture.

Requirements

  • Education: Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience.
  • Experience: 5+ years of experience in a Site Reliability Engineering, DevOps, or related role.
  • Technical Skills:
  • Proficiency in scripting or programming language (e.g., Python, Go)
  • Experience with cloud services (AWS, Azure)
  • Proficiency with containerization (Docker, Kubernetes, ECS)
  • Proficiency in configuration management tools (Terraform, Atlantis, Terragrunt)
  • Familiarity with CI/CD tools (GitHub Actions, AWS CodeBuild, CircleCI)
  • Experience with monitoring tools (Grafana, Datadog)
  • Familiarity with database technologies (RDS, Aurora, PostgreSQL)
  • Experience with large-scale distributed systems and microservices architecture.
  • Problem-Solving: Strong analytical and problem-solving skills with the ability to troubleshoot complex systems.
  • Communication: Excellent verbal and written communication skills, with the ability to collaborate effectively across teams.
  • Ability to obtain a U.S. personnel security clearance.

Preferred qualifications

  • Active TS/SCI clearance

Benefits & conditions

Pulled from the full job description Health insurance Vision insurance Dental insurance Unlimited paid time off, * Global workforce: flexible remote/hybrid opportunities

  • Work on complex, meaningful missions with real-world impact
  • Unlimited paid time off for most roles
  • Competitive salary and equity packages
  • Comprehensive health, dental, and vision coverage
  • Access to the forefront of commercial space operations and defense innovation

Compensation for this role is based on the San Francisco Bay Area market and may be adjusted based on the candidate’s final location. The estimated base salary is $192,000, with additional compensation opportunities to include bonus and equity.

About the company

At LeoLabs, we’re building the living map of activity in space. Through our proprietary global radar network and AI-enabled analytics platform, we collect millions of measurements daily on more than 25,000 objects in low Earth orbit (LEO). Our radar-powered intelligence protects billions in assets, monitors adversarial behavior, and ensures safe operations for commercial and government missions.

We’re not just building technology, we are redefining global security, safety, and transparency in space. As orbital activity accelerates and threats grow more complex, LeoLabs is a trusted partner for Space Domain Awareness, Space Traffic Management, and Satellite Operations for top-tier space operators and allied defense organizations.

If you’re looking to work on mission-critical challenges at the forefront of aerospace, national security, and AI, your impact starts here.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all