Site Reliability Engineer

STAFIDE
Amsterdam, Netherlands
7 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Microsoft Azure Distributed Systems Performance Tuning Reliability Engineering Infrastructure Automation Frameworks Bicep Terraform Microservices

Job description

  • Manage, monitor, and improve the reliability, availability, and performance of production systems.
  • Take ownership of production environments and drive continuous operational improvements.
  • Work with Azure cloud services, including compute, storage, and networking.
  • Support distributed systems and microservices to ensure high system resilience.
  • Collaborate with cross-functional teams to resolve incidents and optimize system performance.
  • Implement automation using Infrastructure as Code (Terraform, ARM, Bicep) and CI/CD pipelines.
  • Contribute to Agile ways of working while ensuring operational excellence.

Requirements

  • 6+ years of experience in Site Reliability Engineering (SRE), Reliability Engineering, or similar roles.
  • Strong hands-on experience managing production systems and driving reliability improvements.
  • Experience working with Microsoft Azure cloud services.
  • Knowledge of distributed systems and microservices architecture.
  • Experience with Infrastructure as Code tools such as Terraform, ARM, or Bicep.
  • Familiarity with automation tools and CI/CD pipelines.
  • Strong stakeholder management, communication, and collaboration skills.
  • Experience working in Agile environments.

You should possess the ability to:

  • Ensure the reliability, scalability, and availability of production systems.
  • Troubleshoot and resolve complex production issues efficiently.
  • Automate infrastructure provisioning and deployment processes.
  • Collaborate effectively with technical and business stakeholders across teams.
  • Drive continuous improvement initiatives for system performance and operational efficiency.
  • Work independently while taking ownership of critical production environments.
  • Adapt to evolving technologies and operational requirements.

Benefits & conditions

  • An opportunity to work on large-scale, cloud-based production environments using modern SRE practices.
  • Exposure to Azure cloud technologies, automation, and Infrastructure as Code.
  • A collaborative Agile work environment with cross-functional teams.
  • Opportunities to contribute to highly available, scalable, and resilient systems.
  • A role that values ownership, innovation, and continuous improvement.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.nl

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

3:07 min

Transitioning architecture to microservices at Netflix

Steve Upton Steve Upton · WWC 2022

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

2:56 min

Provisioning a secure container infrastructure with Bicep

Matthias Falkenberg +1 · WWC 2022

3:38 min

Shifting from monolithic architectures to microservices and containers

Markus Kett Markus Kett · WWC 2022

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

Videos

See all

Related articles

See all