> Markdown version of [/jobs/ext/1687490-site-reliability-engineer-observability-hybrid](https://www.wearedevelopers.com/jobs/ext/1687490-site-reliability-engineer-observability-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (Observability) - (Hybrid) - **Company:** Edreams - **Location:** Barcelona, Spain (Remote available) - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Continuous Integration, Software Debugging, Distributed Systems, Python (Programming Language), Reliability Engineering, Site Reliability Engineering Practices, YAML, Scripting, Mttr, Kubernetes, Deployment Automation, Terraform, Golang - **Published:** July 21, 2026 - **Apply:** https://www.buscojobs.com.es/site-reliability-engineer-observability-hybrid-en-barcelona-ID-363412644 ## About the Role * Solid incident management experience * Advanced debugging in distributed systems * Experience with Infrastructure as Code (Terraform, Terragrunt) * Scripting skills (Python, YAML, Go) * Automation tools familiarity (Terraform, ArgoCD, Crossplane) * Kubernetes expertise * Experience with GCP (and cloud providers like AWS/Azure) * OpenTelemetry/APM instrumentation knowledge * Proactive, data?driven reliability mindset * CAN DO attitude and learning agility ## Description Experteer Overview In this hybrid Barcelona role, you will drive the reliability of our global platform by applying SRE practices to maximize uptime and observability.You'll define and monitor SLIs/SLOs, reduce toil through automation, and lead incident response and post?mortem learning.You'll collaborate with cross?functional teams to instrument code, build dashboards, and scale resilient infrastructure that underpins a world?leading travel subscription platform.Compensaciones / Beneficios* Lead incident response, triage, and troubleshooting in complex distributed systems* Design and implement automated remediation to reduce operational toil* Manage comprehensive observability (monitoring, alerts, logs) for ecosystem health* Facilitate blameless post?mortems and drive improvements from incidents* Act as internal consultant/evangelist; train product teams on instrumentation (OpenTelemetry/APM)* Support GitOps and automation to ensure stability (CI/CD, IaC, containerization)* Define and monitor SLIs/SLOs to align performance with business needs* Optimize infrastructure via code for scalability and availability* Create/manage automated deployment lifecycles* Collaborate with other teams to apply best practices in the development lifecycle* Implement CI, CD, and deployment methodologies to improve MTTD/MTTRResponsabilidades* Solid incident management experience* Advanced debugging in distributed systems* Experience with Infrastructure as Code (Terraform, Terragrunt)* Scripting skills (Python, YAML, Go)* Automation tools familiarity (Terraform, ArgoCD, Crossplane)* Kubernetes expertise* Experience with GCP (and cloud providers like AWS/Azure)* OpenTelemetry/APM instrumentation knowledge* Proactive, data?driven reliability mindset* CAN DO attitude and learning agilityRequisitos principales* Hybrid home?office model* Relocation support* Premium equipment and role?based options* Coursera access and ongoing training* Career development programs (eVOLVE)* Flexible benefits and performance?based bonuses ## Related Videos - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [CI/CD with Github Actions](https://www.wearedevelopers.com/videos/856-ci-cd-with-github-actions) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025)