> Markdown version of [/jobs/ext/2163525-senior-manager-site-reliability-engineering-infrastructure-platform](https://www.wearedevelopers.com/jobs/ext/2163525-senior-manager-site-reliability-engineering-infrastructure-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Manager, Site Reliability Engineering - Infrastructure Platform - **Company:** Okta, Inc. - **Location:** Chicago, IL, United States - **Experience:** Expert - **Salary:** $232,000.0 - $319,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Apache HTTP Server, Software as a Service, Cloud Computing, Cloud Engineering, Computer Networks, Continuous Integration, Monitoring of Systems, Nginx, Systems Development Life Cycle, Release Management, Reliability Engineering, Cloud Services, Grafana, Multi-Cloud, Kubernetes, Infrastructure Automation Frameworks, Information Technology, BIG-IP Access Policy Manager (APM), Terraform, Splunk - **Published:** August 21, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/18027405?backUrl=%2Fcareer%2F18027405%2FSenior-Manager-Site-Reliability-Engineering-Infrastructure-Platform-Illinois-Chicago ## About the Role * 6+ years of experience in technical leadership & people management * 3+ years of experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS. Experience supporting a multi-Cloud environment will be a plus. * Strong expertise in cloud-native architectures, Edge infrastructure (WAF, ALB, NLB, Apache, Nginx), IaC (Terraform), Splunk, Grafana and CI/CD pipelines. * Strong background and hands-on experience in SRE automation& tooling * Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment. * Demonstrated ability to lead cross-functional teams and manage large-scale programs * Effective verbal, written communication and interpersonal skills * Computer Science Degree or related degree or equivalent experience, * This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing the U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire. ## Description Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Infrastructure Platform and Shared Services Team Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. As the Sr. Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, Observability, automation platform & tooling. What you'll be doing * Lead the Infra platform and shared services org and various initiatives across SRE & Infrastructure organization. * Build a world-class observability platform and monitoring capabilities enabled with self-service * Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and intuitive self-service capabilities. * Own the design and operation of scalable, self-service Cloud infrastructure platforms (e.g. Observability Platform, SRE Productivity, deployments, and Edge Infrastructure) * Lead, mentor, and grow a high-performing team of engineers and managers across SRE and infrastructure shared services domains. * Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints. * Improve SDLC processes for Cloud infrastructure as a code, including the maturity of product deployements, change and release management * Manage service and business expectations and prioritize resource allocation * Maintain a deep knowledge of industry best practices, evolving trends, and technologies ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Post-Quantum Cryptography: Preparing for Q-Day](https://www.wearedevelopers.com/videos/100179-post-quantum-cryptography-preparing-for-q-day) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) - [Let developers develop again](https://www.wearedevelopers.com/videos/463-let-developers-develop-again) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025)