> Markdown version of [/jobs/ext/2004485-sr-site-reliability-engineer-core-platform-embedded-reliability-hybrid](https://www.wearedevelopers.com/jobs/ext/2004485-sr-site-reliability-engineer-core-platform-embedded-reliability-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid) - **Company:** CrowdStrike - **Location:** Redmond, WA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Acceptance Test-Driven Development, Software as a Service, Continuous Delivery, Data Structures, Shard (Database Architecture), Distributed Systems, Python (Programming Language), Node.Js, Open Source Technology, Performance Tuning, Reliability Engineering, Scala (Programming Language), Multithreading, Cloud Platform System, Concurrency, Parallel Computation, Backend, Kotlin, Falcon Platform, Information Technology, Programming Languages, Microservices - **Published:** August 9, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/sr-site-reliability-engineer-core-platform-and-embedded-reliability-hybrid-redmond-wa-usa-58863133 ## About the Role Experteer Overview As a Principal SRE, you drive reliability and scalability across CrowdStrike Falcon's core platform, partnering with product groups to define multi-year roadmaps and architect foundational services. You will hands-on engineer, re-architect critical systems, and lead initiatives that reduce toil and improve observability at scale. Your work shapes shared infrastructure and guides product teams, enabling rapid, reliable delivery of AI-enabled threat detection. This is a high-impact role with autonomy to influence architecture and delivery across the Falcon Platform. Compensation / Benefits * Define and drive multi-year reliability roadmaps with engineering leadership * Design and implement architectural improvements to services, libraries, and platforms * Develop and maintain scalable, reliable services * Extend libraries for cross-cutting cloud platform concerns * Lead reliability, scalability, performance, and cost-efficiency initiatives in large-scale systems * a, distributed systems and service-oriented backends at scale * 5+ years developing microservices for SaaS in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js) * Expert-level proficiency in at least one programming language, with expert-level Go or willingness to reach expert in Go * Deep understanding of distributed systems, consensus, replication, and scalability patterns * Experience scaling backend systems: sharding, partitioning, horizontal scaling, capacity planning, performance optimization * Strong grasp of multi-threading, concurrency, and parallel processing * Proven track record of architectural decisions at organizational scope * Strong systems thinking and ability to influence across boundaries * Engineering best practices: testing, peer review, resilient architectures * Thrives in fast-paced, test-driven, collaborative environment; team-oriented * Desire to ship code and see it run in production * Degree in Computer Science, or commensurate aaaaa aaaz_ in data structures and algorithms * Experience utilizing AI to enhance decision-making and efficiency Key requirements * Market leader in compensation and equity awards * Wellness programs * Paid parental and adoption leaves * Professional development opportunities * Employee Networks and volunteer opportunities * Great Place to Work Certified ## Description Experteer Overview As a Principal SRE, you drive reliability and scalability across CrowdStrike Falcon's core platform, partnering with product groups to define multi-year roadmaps and architect foundational services. You will hands-on engineer, re-architect critical systems, and lead initiatives that reduce toil and improve observability at scale. Your work shapes shared infrastructure and guides product teams, enabling rapid, reliable delivery of AI-enabled threat detection. This is a high-impact role with autonomy to influence architecture and delivery across the Falcon Platform. Compensation / Benefits * Define and drive multi-year reliability roadmaps with engineering leadership * Design and implement architectural improvements to services, libraries, and platforms * Develop and maintain scalable, reliable services * Extend libraries for cross-cutting cloud platform concerns * Lead reliability, scalability, performance, and cost-efficiency initiatives in large-scale systems * Establish observability practices and drive automation including continuous delivery * Define and implement SLOs and error budgets to guide prioritization * Lead performance and cost optimization through profiling and capacity planning * Conduct resilience engineering including chaos experiments and failure modelling * Implement automation and infrastructure-as-code to reduce manual toil * Provide technical leadership during incidents and post-incident retrospectives * Identify opportunities to extract common patterns into shared libraries/tools * Continuously re-evaluate architectures for improvement in performance, stability, user experience * Drive strategic technical decisions and influence infrastructure improvements * Mentor engineers and raise technical IQ across the org * Contribute to open source and advocate software engineering best practices (Go) * Collaborate across teams to own deliverables and foster accountable, energetic execution Tasks * 10+ years building and operating distributed systems and service-oriented backends at scale * 5+ years developing microservices for SaaS in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js) * Expert-level proficiency in at least one programming language, with expert-level Go or willingness to reach expert in Go * Deep understanding of distributed systems, consensus, replication, and scalability patterns * Experience scaling backend systems: sharding, partitioning, horizontal scaling, capacity planning, performance optimization * Strong grasp of multi-threading, concurrency, and parallel processing * Proven track record of architectural decisions at organizational scope * Strong systems thinking and ability to influence across boundaries * Engineering best practices: testing, peer review, resilient architectures * Thrives in fast-paced, test-driven, collaborative environment; team-oriented * Desire to ship code and see it run in production * Degree in Computer Science, or commensurate experience in data structures and algorithms * Experience utilizing AI to enhance decision-making and efficiency Key requirements * Market leader in compensation and equity awards * Wellness programs * Paid parental and adoption leaves * Professional development opportunities * Employee Networks and volunteer opportunities * Great Place to Work Certified ## Related Videos - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) - [Why Kotlin is the better Java and how you can start using it](https://www.wearedevelopers.com/videos/661-why-kotlin-is-the-better-java-and-how-you-can-start-using-it) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)