> Markdown version of [/jobs/ext/2689437-site-reliability-engineer-apple-data-platform-sre-apple-services-engineering](https://www.wearedevelopers.com/jobs/ext/2689437-site-reliability-engineer-apple-data-platform-sre-apple-services-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering - **Company:** Apple Inc. - **Location:** Cupertino, CA, United States - **Experience:** Expert - **Salary:** $184,700.0 - **Contract:** Permanent contract - **Skills:** Airflow, Amazon S3, Code Review, Computer Programming, Data Infrastructure, Data Security, Linux, Disaster Recovery, Distributed Data Store, Distributed Systems, Apache Hadoop, Hadoop Distributed File System, Apache HBase, Python (Programming Language), Reliability Engineering, Software Engineering, Ceph (Software), Automated Data Processing (ADP), Apache Yarn, Apache Spark, Generative AI, Siri, Data Lakes, Kubernetes, Information Technology, Golang - **Published:** September 3, 2026 - **Apply:** https://www.themuse.com/jobs/apple/senior-site-reliability-engineer-apple-data-platform-sre-apple-services-engineering ## About the Role 15+ years of experience in SRE or related work managing infrastructure at scale Experience with Ceph object storage operations Kubernetes cluster operations experience, particularly running stateful data workloads Experience with scale testing, disaster recovery, and capacity planning across distributed data systems Experience driving multi-year platform migrations or large-scale architectural transitions Ability to define the technical roadmap for a data platform organization and drive cross-functional alignment on architectural standards and best practices Background in data security, access control, or compliance-sensitive data environments Minimum Qualifications BS/MS in Computer Science or equivalent 12+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale 5+ years of experience in technical leadership roles, with demonstrated ability to lead horizontally across teams without direct authority Broad expertise across the data platform stack: Hadoop (HDFS, YARN), HBase, Apache Spark, Data Lake architectures, S3-compatible storage solutions, and Apache Airflow History of defining and driving SLO/error budget frameworks and reliability practices across multiple teams or services Demonstrable programming skills to develop shared tooling, lead code reviews, and set engineering standards, Strong written and verbal communication skills - able to present technical strategy to both engineers and leadership Advanced knowledge of Linux, networking, and distributed systems fundamentals ## Description As a principal contributor and technical lead in our Apple Data Platform (ADP) SRE organization, you will apply SRE principles as you mentor and partner with our engineers and partner teams, ensuring large-scale analytics infrastructure runs reliably and efficiently. This role focuses on driving reliability standards, architectural consistency, and engineering excellence across peer SRE teams and partner engineering organizations - spanning Hadoop, HBase, Spark, Data Lakes, and Airflow ecosystems - through technical leadership, cross-functional alignment, and the development of platform-wide tooling, observability, and operational practices that raise the reliability bar for all of ADP. This role includes production on-call responsibilities., Apple Service Engineering (ASE) teams build and scale the platforms and infrastructure behind many of Apple's services - including iCloud, iTunes, Siri, and Maps. We are the foundation on which Apple's software developers build the products that our customers love. We are looking for a passionate and dedicated Technical Lead to drive SRE standards and engineering excellence across the entire Apple Data Platform organization. The Apple Data Platform (ADP) SRE Technical Lead partners with multiple SRE and engineering teams across the data platform - including teams responsible for Hadoop and HBase infrastructure, Spark, S3-compatible storage, and Airflow-orchestrated pipelines. Rather than owning a single vertical, this role sets the technical direction for how reliability is practiced across ADP: defining SLOs, establishing architectural review processes, developing shared tooling and automation, and ensuring that SRE principles are applied consistently as the platform scales. You will be a force multiplier - making every team around you more effective. Responsibilities: Serve as the SRE Technical Lead across ADP, partnering with vertical SRE teams and software engineering organizations to ensure reliability standards are consistently applied across the full data platform Define and drive adoption of SLO frameworks, error budget policies, and incident management practices across ADP services Provide architectural review and reliability guidance for new services and major platform changes, identifying risks and influencing design before they reach production Lead the development of shared observability, automation, and infrastructure-as-code tooling that benefits multiple ADP teams simultaneously Identify and eliminate systemic sources of toil and instability across the platform; advocate for and deliver platform-wide reliability improvements Mentor and grow SRE engineers across teams, establishing a culture of engineering excellence and continuous improvement Represent ADP SRE in cross-organizational forums, communicating technical strategy and reliability posture to ASE and Apple leadership Programming in Python and Golang, supported by Generative AI tooling, to accelerate development of mission-critical shared automation and tools Production on-call and incident management responsibilities, including leading response for high-severity cross-platform incidents ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Lessons from Steve Jobs - Learnings from the Past for the Future](https://www.wearedevelopers.com/videos/1021-lessons-from-steve-jobs-learnings-from-the-past-for-the-future) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025)