> Markdown version of [/jobs/ext/1399441-staff-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1399441-staff-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer - **Company:** Okta, Inc. - **Location:** Bellevue, WA, United States - **Experience:** Expert - **Salary:** $194,000.0 - $267,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Apache HTTP Server, Apache Tomcat, Bash Shell, Cloud Computing, System Configuration, Continuous Integration, Software Debugging, Linux, Domain Name System (DNS), Hypertext Transfer Protocols (HTTP), Apache Hypertext Transfer Protocol Server, Python (Programming Language), Network Protocols, Nginx, Public Key Infrastructure, Reliability Engineering, Ansible, TCP/IP, Tcpdump, Wireshark, SSL Certificate Management, Load Balancing, Okta, System Availability, Git, Puppet, Terraform, Docker, Golang - **Published:** July 23, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=936486a5bf1cc41a ## About the Role * 8+ years of operations experience configuring, deploying, monitoring and troubleshooting applications and infrastructure in the cloud * 8+ years administering or operating within a Linux environment, strong experience using Linux based tooling, and ability to debug systems level problems * In-depth understanding of TCP/IP, HTTP, Load Balancing, DNS and other networking protocols * Solid understanding and experience with Apache httpd, nginx, Apache Tomcat, or similar * Strong problem solving and debugging skills coupled with a desire to take on ownership and responsibility * Proficiency in Bash, Python, Golang, or similar. Experienced with git * Experience working with Terraform, Ansible, Chef, Puppet or similar automation tools * Excellent written and verbal communication skills * Willingness to work on-call And extra credit if you have experience in any of the following! * Experience working in a security-oriented cloud environment * Experience debugging software using gdb, strace, ltrace, tcpdump, Wireshark, etc. * Experience working with Docker and Kubernetes * Experience with PKI / certificate management ## Description Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Team The Site Reliability team is dedicated to architecting and owning the foundational infrastructure tooling and CI/CD platforms that support Okta's SRE ecosystem. In this development-focused role, you will leverage a modern tech-stack to build durable, automated systems that maximize platform reliability and engineering velocity. The ideal candidate is someone who enjoys analyzing systems and identifying areas of opportunity to improve system performance, availability and capacity. They are part systems administrator, part network administrator, and part developer. What you'll be doing * Maintain a highly available cloud infrastructure edge for the Okta identity platform * Automate AWS infrastructure with Terraform and/or Chef * Evolve the system by introducing changes to improve efficiency, scalability, and velocity ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Post-Quantum Cryptography: Preparing for Q-Day](https://www.wearedevelopers.com/videos/100179-post-quantum-cryptography-preparing-for-q-day) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)