Senior Sre Engineer

Dempo
Álava, Spain
10 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Compensation
€45,000.0 - €55,000.0
Working hours
Regular working hours

Tech stack

Bash Shell Cloud Computing Continuous Integration Data Deduplication Domain Name System (DNS) Python (Programming Language) Networking Basics Open Source Technology Reliability Engineering Software Engineering Scripting Load Balancing
+5 more
Istio Multi-Cloud Git Kubernetes Terraform

Job description

Senior SRE EngineerJoin DempoOwn reliability at scale with us!We are looking for a Senior SRE Engineer to join our infrastructure team and take technical leadership over production resilience.This role sits at the Senior level on our Cloud/Platform/SRE career path - reliability engineering with a heavy focus on metrics and production systems.You’ll define SLIs and SLOs, lead incident response as commander, drive observability strategy end to end, and mentor cloud/platform engineers as you go.You’ll work closely with Product and Engineering, balancing speed, quality, and long-term reliability, while making the architectural calls that keep our systems resilient under load.ResponsibilitiesReliability & Incident Management· Lead incidents as commander: set and revise severity, and know when to mitigate first and diagnose later· Own the incident record and timeline standard, including the link between deployments and incidents· Communicate with stakeholders while an incident is active· Conduct blameless postmortems and drive toil identification and elimination as measured workObservability· Implement the three pillars of observability (logs, metrics, traces) end to end· Design metrics and query strategy - recording rules, dashboard design that separates on-call needs from analyst needs· Define SLIs and SLOs for critical services, choosing the indicator that reflects user experience over the one that’s easiest to measure· Design alerting systems - routing, escalation, deduplication, and alert fatigue reduction (multi-window burn-rate alerts)Platform & Production Systems· Design workload health signals - liveness, readiness, and startup probes - and reason about workload lifecycle (SIGTERM handling, termination grace periods, connection draining)· Build runbook automation and self-healing systems to reduce operational toil· Contribute to CI/CD framework improvements and cost optimization initiativesTechnical Leadership· Make architectural decisions for reliability-critical systems· Mentor cloud/platform engineers· Influence technical direction on infrastructure and platform decisionsRequirements· 5+ years of experience in SRE, Cloud, or Platform engineering roles· Profound knowledge of SRE principles and incident response practices· Profound knowledge of the observability stack: log aggregation and query design, metrics/dashboard design, and distributed tracing· Profound knowledge of Kubernetes cluster operations and workload objects (Deployments, StatefulSets, Jobs, DaemonSets) and their failure modes· Solid to profound knowledge of networking fundamentals: DNS as infrastructure, TLS certificate lifecycle, load balancing and reverse proxies· Hands-on experience with Infrastructure as Code (Terraform or equivalent)· Profound knowledge of process and OS-level architecture trade-offs as they apply to containers (immutable infrastructure, image strategy)· Scripting proficiency (Bash or Python) for tooling and automation· Strong Git and collaborative workflow experienceNice to Have· Experience with chaos engineering or failure injection programs· Exposure to multi-region or multi-cloud design trade-offs· Familiarity with service mesh implementations· Prior mentoring or technical leadership experience· Certifications such as Site Reliability Engineering (SRE) Foundation or Observability Foundation· Contributions to open-source observability or Kubernetes toolingWhat We OfferPermanent contract.Flexible working hours (core hours: 09:***:30).Remote-first culture, with the option to work from our Granada office.30 working days of annual leave, plus December 24th and December 31st as additional company days off that do not count against your holiday allowance.Private health insurance.Your choice of MacBook or Lenovo.Continuous professional development:· Unlimited access to learning platforms.· Learning time during working hours.· Budget for certifications and specialised training.Employee referral programme.Opportunity-based bonuses.Stable, long-term projects.Real opportunities for professional growth.The opportunity to play a key role in the growth of a modern Software Engineering company.Selection Process1.With pleasure, we receive your CV and we give you a call2.Now it’s when we put faces to names, we’d love to get to chat with you!3.Let’s deepen a bit more with a technical interview, a chance to meet your People Partner4.We get back to you with offer and feedSalary Range· €45,000 - €55,000Dempo is where technology, teamwork, and your professional growth come together.

Requirements

choosing the indicator that reflects user experience over the one that’s easiest to measure· Design alerting systems - routing, escalation, deduplication, and alert fatigue reduction (multi-window burn-rate alerts)Platform & Production Systems· Design workload health signals - liveness, readiness, and startup probes - and reason about workload lifecycle (SIGTERM handling, termination grace periods, connection draining)· Build runbook automation and self-healing systems to reduce operational toil· Contribute to CI/CD framework improvements and cost optimization initiativesTechnical Leadership· Make architectural decisions for reliability-critical systems· Mentor cloud/platform engineers· Influence technical direction on infrastructure and platform decisionsRequirements· 5+ years of experience in SRE, Cloud, or Platform engineering roles· Profound knowledge of SRE principles and incident response practices· Profound knowledge of the observability stack: log aggregation and query design, metrics/dashboard design, and distributed tracing· Profound knowledge of Kubernetes cluster operations and workload objects (Deployments, StatefulSets, Jobs, DaemonSets) and their failure modes· Solid to profound knowledge of networking fundamentals: DNS as infrastructure, TLS certificate lifecycle, load balancing and reverse proxies· Hands-on experience with Infrastructure as Code (Terraform or equivalent)· Profound knowledge of process and OS-level architecture trade-offs as they apply to containers (immutable infrastructure, image strategy)· Scripting proficiency (Bash or Python) for tooling and automation· Strong Git and collaborative workflow experienceNice to Have· Experience with chaos engineering or failure injection programs· Exposure to multi-region or multi-cloud design trade-offs· Familiarity with service mesh implementations· Prior mentoring or technical leadership experience· Certifications such as Site Reliability Engineering (SRE) Foundation or Observability Foundation· Contributions to open-source observability or Kubernetes toolingWhat We OfferPermanent contract.Flexible working hours (core hours: 09:***:30).

About the company

Remote-first culture, with the option to work from our Granada office.30 working days of annual leave, plus December 24th and December 31st as additional company days off that do not count against your holiday allowance.Private health insurance.Your choice of MacBook or Lenovo.Continuous professional development:· Unlimited access to learning platforms. · Learning time during working hours. · Budget for certifications and specialised training.Employee referral programme.Opportunity-based bonuses.Stable, long-term projects.Real opportunities for professional growth.The opportunity to play a key role in the growth of a modern Software Engineering company.Selection Process1. With pleasure, we receive your CV and we give you a call2. Now it’s when we put faces to names, we’d love to get to chat with you! 3. Let’s deepen a bit more with a technical interview, a chance to meet your People Partner4. We get back to you with offer and feedSalary Range· €45,000 - €55,000Dempo is where technology, teamwork, and your professional growth come together.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all