> Markdown version of [/jobs/ext/2359826-staff-platform-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2359826-staff-platform-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Platform Site Reliability Engineer - **Company:** Index Exchange - **Location:** London, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Computing Platforms, Systems Engineering, Big Data, Cloud Computing, Protocol Stack, DevOps, Distributed Data Store, Distributed Systems, Domain Name System (DNS), Apache Hadoop, Monitoring of Systems, Apache HBase, Python (Programming Language), Key Management, Linux Kernel, Networking Basics, Performance Tuning, Software Architecture, Role-Based Access Control, Reliability Engineering, Ansible, Prometheus, Service Discovery, Software Engineering, Ceph (Software), SSL Certificate Management, Load Balancing, Cloud Platform System, Delivery Pipeline, Grafana, Apache Spark, HybridCloud, Core Api, Containerization, Git Flow, Kubernetes, Bare Metal, Apache Kafka, Terraform - **Published:** August 27, 2026 - **Apply:** https://uk.indeed.com/viewjob?jk=9c5894d76cdd1c33 ## About the Role * 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps. * Deep experience with Linux internals: kernel tuning, network stack, system observability, security. * Strong Kubernetes expertise: cluster lifecycle, networking, storage, RBAC, multi-cluster-across bare-metal and cloud (EKS, GKE). * Infrastructure-as-code at scale: Terraform, Ansible, GitOps (ArgoCD or similar). * Proficiency in Go, Python, or both-for building libraries, SDKs, and platform APIs, not just scripts. * Solid networking fundamentals (L2-L7), load balancing, DNS, service discovery. * A track record of driving technical strategy across teams-not just executing within one., * Distributed storage systems (e.g. Ceph) * Big data infrastructure: Hadoop, Spark, HBase, Kafka. * Observability stack design: Prometheus, Grafana, ELK, Mimir, Loki, Tempo. * Secrets management (Vault), certificate management, access control at scale. * Hybrid cloud architectures: federating public cloud (AWS, GCP) with on-premises environments. * Experience with bare-metal infrastructure in globally distributed data centers ## Description Cloud Platform Engineering builds the platform that makes all of this possible. We architect and deliver the foundational infrastructure-Kubernetes clusters, deployment pipelines, infrastructure-as-code frameworks, container platforms, storage systems-and the software layer on top of it: standard libraries, internal SDKs, platform APIs, and the interfaces that every engineering team at Index Exchange builds on top of. Platform engineering here is software engineering. We don't operate day-to-day; our operational counterpart, Systems Engineering, carries that load. We stay focused on what moves the needle: designing and building systems at scale. As a Staff Platform Engineer, you will own architectural decisions, deliver large-scale infrastructure projects, and shape the technical direction of a platform that processes more real-time transactions than almost any system on earth. The vision: We build, scale, and maintain Index Cloud-a multi-tenant, globally distributed compute platform that serves both our exchange workloads and the workloads of every partner in our ecosystem. This is not incremental improvement. It's building a platform-as-a-product at internet scale. What You'll Work On * Build the platform, not run it. You'll design and deliver multi-tenant Kubernetes infrastructure across bare-metal and public cloud, build infrastructure-as-code frameworks that push changes across thousands of servers, and write the standard libraries, SDKs, and platform APIs that every engineering team at Index Exchange depends on. This is the work that takes Index Cloud to the next level. * Solve hard distributed systems problems. Real-time bidding with sub-millisecond overhead. Multi-datacenter consistency. Deploying safely to a fleet of thousands. Load balancing at a scale where the easy approaches don't work. If the phrase "globally distributed auction system" makes you lean in, keep reading. * Own the architecture. You'll drive technical direction through RFCs and design reviews, set standards for shared platform domains, and make the calls on tooling, security posture, and system design. This is a Staff-level role-you'll influence engineering direction across multiple divisions. * Make engineers faster. Build golden paths, self-service tooling, and platform APIs that let engineering teams ship without friction. The best platform work is invisible to its users-it just works. * Mentor and raise the bar. Coach engineers, foster a culture of engineering excellence, and collaborate across Cloud Platform Operations, SRE, Network, Security, and Software Engineering teams., If you require an accommodation, please share the details of your request and any information how we can assist you with the hiring recruiter when they contact you. Index Exchange will make reasonable efforts to ensure accommodation requests are met throughout the recruitment process. Index everywhere, Index anywhere Our corporate headquarters are in Toronto, with major offices in New York, Montreal, Kitchener, London, San Francisco, and many other global cities. As a major global advertising exchange, we are committed to operating as a tightly knit global team and embracing and empowering talent wherever our colleagues may be. AI Disclaimer: Automated tools including artificial intelligence may be used to assist in screening and assessing applicants for this role. All hiring decisions involve human review. Vacancy Status: This posting is for an existing vacancy. ## Related Videos - [Platform Engineering vs. DevOps Why not both?](https://www.wearedevelopers.com/videos/885-platform-engineering-vs-devops-why-not-both) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Eclipse Che for Infrastructure Automation](https://www.wearedevelopers.com/videos/1611-eclipse-che-for-infrastructure-automation) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)