> Markdown version of [/jobs/ext/1978865-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1978865-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Circle (nyse: Crcl) - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $152,500.0 - $205,000.0 - **Contract:** Permanent contract - **Skills:** JavaScript (Programming Language), Artificial Intelligence, Cloud Computing, Code Review, Continuous Integration, DevOps, Distributed Systems, Domain Name System (DNS), Identity and Access Management, Python (Programming Language), Routing, Reliability Engineering, Software Engineering, TypeScript, Load Balancing, Cloud Platform System, Computer Network Technologies, Backend, Containerization, Kubernetes, Deployment Automation, Terraform - **Published:** August 7, 2026 - **Apply:** https://www.dice.com/job-detail/e392cc59-a8a6-4688-bfbb-948aec94aae3 ## About the Role * 5+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a closely related software engineering role supporting production systems. * Deep, hands-on Kubernetes expertise: designing, operating, securing, and troubleshooting production clusters and containerized workloads at scale. * Strong Terraform experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes through reviewable, automated workflows. * Production software-development experience in Go, Python, or JavaScript/TypeScript, with the ability to build maintainable backend services, tooling, and automation-not only scripts. * Demonstrated success improving the reliability, performance, scalability, or cost efficiency of distributed systems in production. * Experience with cloud infrastructure and core networking concepts, including IAM, DNS, load balancing, routing, service networking, and secure connectivity. * Strong observability and troubleshooting skills using metrics, logs, traces, alerting, and incident data to diagnose complex systems. * Experience defining and operating against SLIs, SLOs, error budgets, incident-management processes, postmortems, and disaster-recovery practices. * Familiarity with CI/CD, GitOps or deployment automation, and safe rollout strategies such as canary or blue-green deployments. * A security-minded approach to infrastructure and a track record of partnering effectively with Security and engineering teams in regulated or high-availability environments. * Clear written and verbal communication, strong ownership, and the judgment to balance speed, risk, and operational excellence. * Experience applying AI-assisted tooling to engineering or operations workflows is a plus. Circle is on a mission to create an inclusive financial future, with transparency at our core. We consider a wide variety of elements when crafting our compensation ranges and total compensation packages. Starting pay is determined by various factors, including but not limited to: relevant experience, skill set, qualifications, and other business and organizational needs. Please note that compensation ranges may differ for candidates in other locations. ## Description As a Senior Site Reliability Engineer on Circle's platform team, you'll design, build, and operate the secure, scalable platform infrastructure behind critical digital-assets, AI, and application workloads. You will bring an engineering mindset to production operations: writing and maintaining services and automation, developing reliable Kubernetes platforms, and using Terraform to make infrastructure repeatable, auditable, and easy to evolve. You'll work closely with platform, product, and application engineering teams to translate workload requirements into resilient technical designs across hybrid and public-cloud environments. This role is for an experienced SRE or infrastructure engineer who enjoys solving hard distributed-systems problems, taking ownership of production outcomes, and raising the reliability, performance, security, and cost-effectiveness of the systems our customers depend on. What you'll work on: * Design, build, and operate Kubernetes platforms that provide secure, highly available, and scalable foundations for critical production services across hybrid and public-cloud environments. * Build infrastructure as code with Terraform, creating reusable modules, safe delivery workflows, and well-governed infrastructure changes. * Develop backend services, internal tools, and operational automation in Go, Python, or JavaScript/TypeScript to eliminate manual work and improve the developer experience. * Partner with engineering and product teams to understand workload requirements and design pragmatic solutions for reliability, performance, capacity, security, and cost. * Improve the production lifecycle through reliable CI/CD, deployment automation, progressive delivery, and clear operational ownership. * Define and evolve observability practices across metrics, logs, traces, alerting, and dashboards so teams can detect issues early and troubleshoot effectively. * Own production reliability by participating in on-call, leading incident response, performing root-cause analysis, and driving blameless postmortems and durable corrective actions. * Establish and maintain reliability targets through meaningful SLIs, SLOs, error budgets, capacity planning, disaster-recovery testing, and resilience improvements. * Embed security and compliance into platform operations, partnering with Security to protect infrastructure, workloads, and data while meeting applicable regulatory requirements. * Apply AI-assisted and data-driven operational techniques to improve signal detection, reduce alert noise, accelerate root-cause analysis, and surface opportunities for automation. * Raise the bar for the team through thoughtful code reviews, documentation, knowledge sharing, and mentorship. * Mentor and support team growth, fostering collaboration and scalability. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Navigating Growth, Scaling Challenges, and Office Expansions with David Singleton, CTO at Stripe](https://www.wearedevelopers.com/videos/100362-navigating-growth-scaling-challenges-and-office-expansions-with-david-singleton-cto-at-stripe) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [React Developer Salary [2023]](https://www.wearedevelopers.com/magazine/198-react-developer-salary-2023) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Fullstack Developer Salary UK](https://www.wearedevelopers.com/magazine/251-fullstack-developer-salary-uk)