> Markdown version of [/jobs/ext/3518688-sr-staff-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3518688-sr-staff-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr Staff Site Reliability Engineer - **Company:** Palo Alto Networks - **Location:** Madrid, Spain - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Continuous Integration, DevOps, Distributed Systems, Github, Monitoring of Systems, Python (Programming Language), Reliability Engineering, Prometheus, Systems Architecture, Google Cloud, Cloud Platform System, Grafana, Multi-Cloud, Reliability of Systems, Gitlab-ci, Git Flow, Kubernetes, Terraform, Pagerduty, Jenkins - **Published:** October 1, 2026 - **Apply:** https://dejobs.org/x/x/2481784171CB4A6CA2087E143A252800/job/ ## About the Role * 5+ years of experience in SRE roles in production environments at scale * Strong hands-on experience with Kubernetes and Terraform * Strong hands-on experience with at least one major cloud platform (GCP or AWS required) * Experience building and configuring monitoring systems (e.g., Prometheus, Grafana) * Familiarity with CI/CD and GitOps tools (GitLab CI, GitHub Actions, Jenkins, Flux) * Proficiency in Python for scripting and automation * Proven success in a fully remote or distributed team environment, demonstrating strong self-management and time organization. * Strong troubleshooting and problem-solving skills with a passion for incident handling * Ability to work in fast-paced environments with high context switching * Highly responsive, proactive, and ownership-driven * Strong collaboration and communication skills * Curious mindset and eagerness to learn ## Description Join a team of senior engineers operating in a large-scale, multi-cloud production environment supporting tens of thousands of enterprise customers worldwide. This is not a typical SRE role - you'll work at the core of a complex, high-impact system alongside experienced DevOps professionals in a fast-paced, cybersecurity-focused organization., * Own and operate large-scale, global production environments across multiple cloud providers (GCP, AWS, Azure) * Actively monitor, investigate, and resolve incidents triggered by automated alerting systems (PagerDuty / Incident Response) * Drive end-to-end troubleshooting across complex, distributed systems with high context switching * Design, deploy, and improve monitoring and observability systems (e.g., Prometheus, Grafana) - not just react to alerts * Collaborate closely with internal teams (CX, CS, Engineering) to ensure system reliability and performance * Work hands-on with modern DevOps and infrastructure tools including Kubernetes, Terraform, CI/CD pipelines, and GitOps workflows * Develop and maintain automation and tooling (primarily in Python) * Gain deep understanding of system architecture and interconnected services * Contribute to a culture of operational excellence in a high-scale, high-availability environment * Champion asynchronous communication, documentation, and tooling standards to ensure seamless collaboration across time zones and fully distributed teams. * On call responsibilities:Daytime hours (12:00-20:00 CET/CEST, based on candidate location and team coverage needs) Occasional weekends and holidays (rotation-based) ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [GitLab CI pipelines for a whole company](https://www.wearedevelopers.com/videos/143-gitlab-ci-pipelines-for-a-whole-company) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [Mobile vs. Backend DevOps](https://www.wearedevelopers.com/videos/1662-mobile-vs-backend-devops) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Mastering Remote Work: Tips for Developers](https://www.wearedevelopers.com/magazine/558-mastering-remote-work-tips-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)