> Markdown version of [/jobs/ext/2005775-sr-site-reliability-engineer-core-platform-embedded-reliability-hybrid](https://www.wearedevelopers.com/jobs/ext/2005775-sr-site-reliability-engineer-core-platform-embedded-reliability-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid) - **Company:** CrowdStrike - **Location:** Sunnyvale, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Acceptance Test-Driven Development, Software as a Service, Code Review, Shard (Database Architecture), Distributed Systems, Python (Programming Language), Node.Js, Open Source Technology, Reliability Engineering, Scala (Programming Language), Multithreading, Cloud Platform System, Concurrency, Parallel Computation, Backend, Kotlin, Falcon Platform, Information Technology, Production Code, Microservices - **Published:** August 9, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/sr-site-reliability-engineer-core-platform-and-embedded-reliability-hybrid-sunnyvale-ca-usa-58863169 ## About the Role _ years building microservices for SaaS in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js) * Expert-level proficiency in at least one language, especially Go * Deep understanding of distributed systems concepts and failure modes * Experience scaling backend systems with sharding, partitioning, and capacity planning * Strong multi-threading, concurrency, and parallel processing knowledge * Track record of architectural decisions with production impact * Influence without direct authority across org boundaries * Solid engineering practices: testing, code review, resilient architecture * Thrives in fast-paced, test-driven, collaborative environments * Desire to ship production code and see it run in production * Degree in Computer Science, or commensurate experience * Experience applying AI to improve decisions and workflows Key requirements * market-leading compensation * comprehensive wellness programs * vacation and holidays * paid parental and aaaaa org leaves * professional development opportunities * employee networks and volunteer opportunities ## Description Experteer Overview As a Principal SRE, you build foundational reliability for CrowdStrike Falcon by shaping core libraries and services and embedding with product teams to scale reliability. You own architectural decisions and drive multi-cloud, high-scale systems with hands-on engineering. You influence across the Falcon Platform, champion observability, and automate toil to improve performance and resilience. This is a high-impact, engineering-heavy role with autonomy and opportunities to define standards across the organization. Compensation / Benefits * Define and drive multi-year reliability roadmaps with engineering leadership * Architecturally improve services, libraries, and platforms impacting multiple product groups * Develop and maintain scalable, reliable services * Extend libraries for cross-cutting cloud platform concerns * Lead reliability, scalability, performance, and cost-efficiency initiatives in large distributed systems * Establish observability practices and drive automation including CD * Define service-level objectives and error budgets to guide prioritization * Lead optimization efforts (profiling, bottlenecks, capacity, cloud efficiency) * Conduct resilience engineering (chaos testing, failure modeling) * Implement automation and infrastructure-as-code to reduce manual toil * Provide technical leadership during incidents and post-incident retrospectives * Identify opportunities to extract common patterns into shared libraries/tools * Continuously re-evaluate architecture for performance, stability, and developer experience * Drive strategic technical decisions and cross-organizational improvements * Mentor engineers and uplift architectural standards * Contribute to open-source practices and Go-centric engineering conventions * Collaborate across teams to own deliverables and build shared components * Build with a self-starter mindset and accountability Tasks * 10+ years of distributed systems and backend service experience at scale * 5+ years building microservices for SaaS in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js) * Expert-level proficiency in at least one language, especially Go * Deep understanding of distributed systems concepts and failure modes * Experience scaling backend systems with sharding, partitioning, and capacity planning * Strong multi-threading, concurrency, and parallel processing knowledge * Track record of architectural decisions with production impact * Influence without direct authority across org boundaries * Solid engineering practices: testing, code review, resilient architecture * Thrives in fast-paced, test-driven, collaborative environments * Desire to ship production code and see it run in production * Degree in Computer Science, or commensurate experience * Experience applying AI to improve decisions and workflows Key requirements * market-leading compensation * comprehensive wellness programs * vacation and holidays * paid parental and adoption leaves * professional development opportunities * employee networks and volunteer opportunities ## Related Videos - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Why Kotlin is the better Java and how you can start using it](https://www.wearedevelopers.com/videos/661-why-kotlin-is-the-better-java-and-how-you-can-start-using-it) - [Inside Bitpanda's Tech Stack: Scaling a European Fintech Leader - Markus Dorner](https://www.wearedevelopers.com/videos/1979-inside-bitpanda-s-tech-stack-scaling-a-european-fintech-leader-markus-dorner) ## Related Articles - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Dev Digest 131 - AI'm not sure about OSS](https://www.wearedevelopers.com/magazine/472-dev-digest-131-ai-m-not-sure-about-oss)