> Markdown version of [/jobs/ext/1991792-staff-site-reliability-engineer-sre-hybrid](https://www.wearedevelopers.com/jobs/ext/1991792-staff-site-reliability-engineer-sre-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer (SRE) (Hybrid) - **Company:** Cisco Systems, Inc. - **Location:** New York, NY, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Computing, Databases, Continuous Integration, Distributed Systems, Reliability Engineering, System Availability, Kubernetes, Deployment Automation, Splunk - **Published:** August 8, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/staff-site-reliability-engineer-sre-hybrid-new-york-ny-usa-58847810 ## About the Role with and resiliency reviews * Drive major reliability and scalability initiatives across Kubernetes, deployment infra, databases, and networking * Automate to reduce toil and boost engineering productivity * Build internal platforms and tooling for reliable operations at scale * Lead incident response and postmortems with long-term remediation * Partner with engineering leadership on platform architecture and deployment strategy * Mentor engineers via design reviews and operational guidelines * Collaborate with customers and teams to design secure, scalable deployment architectures for cloud and on-prem environments Tasks * 8+ years' experience with a Bachelor's degree or 6+ yrs with Masters or 3+ years with a PhD, or equivalent related experience * At least 6 years in Site Reliability Engineering, Platform/Cloud/Infrastructure Engineering, or related fields * 5+ years operating large-scale Kubernetes platforms in production * Experience designing highly available, scalable distributed systems * Experience with AWS, GCP, or other public clouds * Strong experience designing CI/CD platforms and deployment automation at scale Key requirements * medical, dental and vision insurance * 401(k) with matching contribution * paid parental leave * paid holidays and vacation policies * sick time and personal wellness days * volunteer days (optional) ## Description Experteer Overview As Staff Site Reliability Engineer, you will lead reliability, scalability, and operational architecture for Splunk Agent Observability's platform. You will set long-term reliability strategy, drive major infrastructure initiatives, and raise engineering standards across deployment automation, production operations, and platform resiliency. You'll influence platform architecture and guide cloud and on-prem deployments, collaborating with cross-functional teams and customers. This role offers impact at scale, shaping how AI-enabled deployments are observed, controlled, and trusted. You'll partner with leadership to elevate incident response and drive long-term improvements. Compensation / Benefits * Define and drive the technical roadmap for platform reliability, scalability and operational excellence * Lead architecture and evolution of deployment platforms for cloud and air-gapped environments * Establish reliability standards including SLOs, readiness, capacity planning, and resiliency reviews * Drive major reliability and scalability initiatives across Kubernetes, deployment infra, databases, and networking * Automate to reduce toil and boost engineering productivity * Build internal platforms and tooling for reliable operations at scale * Lead incident response and postmortems with long-term remediation * Partner with engineering leadership on platform architecture and deployment strategy * Mentor engineers via design reviews and operational guidelines * Collaborate with customers and teams to design secure, scalable deployment architectures for cloud and on-prem environments Tasks * 8+ years' experience with a Bachelor's degree or 6+ yrs with Masters or 3+ years with a PhD, or equivalent related experience * At least 6 years in Site Reliability Engineering, Platform/Cloud/Infrastructure Engineering, or related fields * 5+ years operating large-scale Kubernetes platforms in production * Experience designing highly available, scalable distributed systems * Experience with AWS, GCP, or other public clouds * Strong experience designing CI/CD platforms and deployment automation at scale Key requirements * medical, dental and vision insurance * 401(k) with matching contribution * paid parental leave * paid holidays and vacation policies * sick time and personal wellness days * volunteer days (optional) ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)