> Markdown version of [/jobs/ext/3053147-site-reliability-engineering-lead](https://www.wearedevelopers.com/jobs/ext/3053147-site-reliability-engineering-lead). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineering Lead - **Company:** Sonar, Inc. - **Location:** Austin, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Amazon Web Services, Cloud Computing, Software Quality, Code Review, DevOps, Disaster Recovery, Identity and Access Management, Python (Programming Language), Reliability Engineering, Cloud Services, Software Engineering, Data Logging, Reliability of Systems, Cloud Optimization, Terraform - **Published:** September 24, 2026 - **Apply:** https://www.thejobnetwork.com/job/a447b718-6502-48cc-9e8e-6bbe9af50c23/site-reliability-engineering-lead ## About the Role * 10+ years of experience in software engineering, with a significant focus on Site Reliability Engineering (SRE), Cloud Operations, or Infrastructure Engineering. * Operational Mindset: Deep understanding of the DevOps/SRE mindset, including experience managing mission-critical shared services (e.g., Aurora DBs, OpenSearch, Control Planes). * Cloud Expertise: Advanced knowledge of AWS (or similar providers) and experience managing organizational-scale infrastructure (IAM, OUs, Account Vending). * Infrastructure as Code: Experience with the coding lifecycle in an infrastructure context (e.g., Python, CDK, Terraform) and the ability to perform rigorous code reviews for infrastructure components. * Observability & Resiliency: Proven experience defining observability patterns (logging, tracing, metrics) and designing Disaster Recovery/Business Continuity strategies. * Agile & FinOps: Experience with Agile methodologies and a strong understanding of cloud cost optimization (Rightsizing, Spot instances, Reserved Instances). ## Description Position description As a member of one of our engineering teams, you'll be a key player in making SonarQube Cloud and SonarQube Server the best tools for Code Quality and Security, providing new features to deliver high-quality and powerful products and services that help our users write better software. You will have the opportunity to see your features come to life in Production with short iteration loops. While keeping our roadmap and business priorities in mind, you will be able to have a high impact on the software that we own and develop. By joining us you will bring your experience and expertise to help push our product to its next stage of evolution and fulfill the needs of our large user and customer base. What you will do * Lead and manage a team of Engineers (Cloud Engineers and Site Reliability Engineers), providing guidance, support, and mentorship to help individuals grow in autonomy and master the complexities of cloud operations. * Hold yourself and your team accountable to high engineering standards, specifically focusing on system reliability, performance, and security. * Manage the team's operational workload, including on-call health, incident response, and the reduction of manual toil through automation. * Foster a safe culture of feedback and continuous improvement, encouraging the team to conduct blameless post-mortems and share architectural insights. * Collaborate with other value stream squads to ensure the production platform meets the needs of our developers while maintaining strict governance. * Communicate a clear vision for the squad that aligns with the platform engineering roadmap, focusing on resiliency, business continuity, and cost optimization. * Partner with the Hiring team to recruit talented Engineers for the team. Participate in improving the hiring process for the team and ensuring we recruit enough to reach our goals. * Lead by example, modeling behaviors of servant leadership and high-stakes decision-making. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Are Code Reviews Worth It? Insights from 16 Years of Review Data](https://www.wearedevelopers.com/videos/1135-are-code-reviews-worth-it-insights-from-16-years-of-review-data) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers)