> Markdown version of [/jobs/ext/2298774-lead-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2298774-lead-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Site Reliability Engineer - **Company:** Federal Reserve Bank of Richmond - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $146,700.0 - $190,500.0 - **Contract:** Temporary contract - **Skills:** Java (Programming Language), Amazon Web Services, Amazon Cloudfront, Amazon Elastic Compute Cloud, Amazon S3, Cloud Computing, Code Review, Databases, Continuous Integration, Data Structures, Software Design Patterns, DevOps, Disaster Recovery, Distributed Systems, Amazon DynamoDB, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Key Management, Node.Js, Open Web Application Security, Scrum Methodology, Reliability Engineering, Runbook, Software Engineering, Software Vulnerability Management, Datadog, Large Language Models, Grafana, Multi-Cloud, Reliability of Systems, HybridCloud, Amazon Virtual Private Cloud (VPC), Gitlab, Event Driven Architecture, Containerization, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Deployment Automation, AWS Fargate, Route53, Cloudwatch, Api Gateway, Terraform, Splunk, New Relic (SaaS), Dynatrace, Serverless Computing, Docker, Static Application Security Testing, Microservices, Dynamic Application Security Testing - **Published:** August 29, 2026 - **Apply:** https://rb.wd5.myworkdayjobs.com/FRS/job/San-Francisco-CA/Lead-Site-Reliability-Engineer_R-0000033169-1 ## About the Role * Strong proficiency in Java, Python, and Node.js * Experience with microservices architecture and distributed systems * Solid understanding of data structures, algorithms, and design patterns * Proficiency in writing clean, maintainable, and testable code Cloud Infrastructure (AWS): * Extensive experience with AWS services including: + Compute: Lambda, ECS, EC2, Fargate + Storage: S3, EBS, EFS + Database: RDS, DynamoDB, Aurora + Networking: VPC, Route53, CloudFront, API Gateway + Monitoring: CloudWatch, X-Ray * AWS certifications (Solutions Architect, DevOps Engineer) preferred DevOps & CI/CD: * Expert-level knowledge of GitLab (CI/CD pipelines, runners, GitOps) * Advanced Terraform skills for infrastructure provisioning and management * Experience with containerization (Docker) and orchestration (Kubernetes/ECS) * Proficiency with configuration management tools Security: * Hands-on experience with SAST (Static Application Security Testing) tools * Knowledge of DAST (Dynamic Application Security Testing) methodologies * Understanding of security best practices, OWASP Top 10, and compliance frameworks * Experience with secrets management and identity access management (IAM) Monitoring & Observability: * Experience with monitoring tools (Grafana, Datadog, New Relic, or similar) * Log aggregation and analysis (CloudWatch Logs, Splunk) * Distributed tracing with aws X-Ray, * Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience * 7+ years of experience in Site Reliability Engineering, DevOps, or related roles * 3+ years in a lead or senior technical position * Proven track record of managing large-scale production systems * Experience with on-call rotations and incident management * GenAI based Applications: Working knowledge of LLMs and agentic applications a plus * Experience with serverless architectures and event-driven systems * Familiarity with chaos engineering principles and practices * Background in Agile/Scrum methodologies * Experience with multi-cloud or hybrid cloud environments The selected candidate will reside within a reasonable commuting distance, as defined by the employing Reserve Bank, and will work full-time onsite. * Eligible Locations for Hire: Richmond, VA, San Francisco, CA * The following Reserve Bank locations are preferred due to the concentration of System IT team members in these locations: San Francisco, and Richmond, VA Base Salary Range: Min: $146,700 Mid: $190,500 Max: $234,300 (Location: San Francisco) ## Description We are seeking an experienced Lead Site Reliability Engineer to join our engineering team and drive the reliability, scalability, and performance of our critical systems. This role combines deep technical expertise in software engineering, cloud infrastructure, and DevOps practices to ensure our services meet the highest standards of availability and operational excellence. Responsibilities System Reliability & Performance * Design, implement, and maintain highly available, scalable, and resilient systems across cloud infrastructure * Establish and monitor SLIs, SLOs, and SLAs to ensure optimal system performance * Lead incident response, conduct root cause analysis, and implement preventive measures * Develop and maintain disaster recovery and business continuity plans Infrastructure & Automation * Architect and manage cloud infrastructure on AWS using Infrastructure as Code (Terraform) * Automate deployment pipelines, monitoring, and operational workflows * Optimize cloud resource utilization and cost management Engineering & Development * Build and maintain internal tools and services to improve operational efficiency * Collaborate with development teams to implement reliability best practices * Conduct code reviews and provide technical guidance on system design * Develop monitoring solutions, alerting systems, and observability frameworks Security & Compliance * Integrate security practices into CI/CD pipelines (SAST/DAST) * Implement and maintain security controls across infrastructure and applications * Ensure compliance with industry standards and regulatory requirements * Conduct security assessments and vulnerability management Leadership & Collaboration * Mentor junior SRE team members and promote SRE culture across the organization * Partner with software engineering teams to improve system reliability * Drive technical initiatives and contribute to architectural decisions * Document processes, runbooks, and technical specifications ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Software Developer Salary in Switzerland [2023]](https://www.wearedevelopers.com/magazine/215-software-developer-salary-in-switzerland-2023) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025)