Senior Site Reliability Engineer/Devops (Hybrid)

Camlin Group
Vitoria-Gasteiz, Spain
21 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English

Tech stack

Amazon Web Services Cloud Computing DevOps Disaster Recovery Reliability Engineering

Job description

About CamlinCamlin is a global technology leader that operates with the vision of bringing revolutionary products to life for a wide range of industries, including power and rail, and also has interests in a number of R&D projects in a variety of scientific sectors.At Camlin we believe in high quality engineering and design, allowing us to develop market leading products and services.In short, we love creating value for our customers by solving difficult problems.As of now, Camlin operates in over 20 countries worldwide.About The RoleWe’re looking for a hands-on Senior SRE to ensure our production services run reliably, securely, and predictably at scale.You’ll focus on real-world system behaviour - detecting failures, leading incident response, improving resilience, and using SLI/SLOs and automation to make reliability, risk, cost, and delivery trade-offs explicitWhat You’ll DoOwn reliability and resilience for live production servicesDefine and evolve SLIs/SLOs aligned to customer impactImprove observability, alert quality, and operational signalsLead during incidents and contribute to blameless post-incident reviewsStrengthen disaster recovery and recoverability (RTO/RPO)Reduce operational toil through smart automationSupport cost-awareness and reliability trade-offs (FinOps mindset)Mentor engineers and contribute to SRE best practicesWhat We’re Looking ForStrong experience running live production systems in an SRE/reliability roleDeep understanding of SLI/SLO-driven modelsProven incident response and on-call experienceSolid observability and automation skillsCloud experience (AWS preferred)Calm, structured approach under pressureFluent EnglishOur ValuesWe work togetherWe believe in peopleWe won’t accept the ‘way it has always been done’We listen to learnWe’re trying to do the right thingEqual Employment Opportunity StatementIndividuals seeking employment at Camlin are considered without regards to race, colour, religion, national origin, age, sex, marital states, ancestry, physical or mental disability, gender identity or sexual orientation.

Requirements

Strong experience running live production systems in an SRE/reliability role Deep understanding of SLI/SLO-driven models Proven incident response and on-call experience Solid observability and automation skills Cloud experience (AWS preferred) Calm, structured approach under pressure Fluent English

Benefits & conditions

We work together We believe in people We won’t accept the ‘way it has always been done’ We listen to learn We’re trying to do the right thing Equal Employment Opportunity Statement Individuals seeking employment at Camlin are considered without regards to race, colour, religion, national origin, age, sex, marital states, ancestry, physical or mental disability, gender identity or sexual orientation.

About the company

Camlin is a global technology leader that operates with the vision of bringing revolutionary products to life for a wide range of industries, including power and rail, and also has interests in a number of R&D projects in a variety of scientific sectors. At Camlin we believe in high quality engineering and design, allowing us to develop market leading products and services. In short, we love creating value for our customers by solving difficult problems. As of now, Camlin operates in over 20 countries worldwide.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · World Congress 2024

2:14 min

Crafting an effective disaster recovery and communication plan

Mihaela-Roxana Ghidersa · LIVE

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

Videos

See all

Related articles

See all