Site Reliability Engineer

Jobgether
Germany
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

PHP (Programming Language) Amazon Web Services Bash Shell Cloud Computing Configuration Management Computer Networks Continuous Integration Linux DevOps Disaster Recovery Domain Name System (DNS) Hypertext Transfer Protocols (HTTP)
+28 more
PostgreSQL Windows Servers MySQL Network Architecture Nginx Reliability Engineering Ansible Prometheus Software Engineering TCP/IP Virtualization Technology Wide Area Networks Scripting Transport Layer Security Grafana Reliability of Systems Git Gitlab-ci Kubernetes Infrastructure Automation Frameworks Information Technology Cloudflare Puppet Rundeck Terraform Oracle Cloud Infrastructure Docker Elk Stack

Job description

Join a technology-driven team responsible for maintaining the reliability, scalability, and security of mission-critical infrastructure supporting high-availability platforms. In this role, you will combine infrastructure engineering, automation, and operational excellence to ensure seamless system performance across cloud and on-premises environments. Working alongside cross-functional engineering teams, you will proactively optimize production systems, strengthen monitoring capabilities, and respond to complex operational challenges. This is an excellent opportunity for an experienced infrastructure professional who enjoys solving technical problems, improving reliability through automation, and contributing to resilient, enterprise-grade services in a fast-paced environment. Accountabilities

  • Design, implement, maintain, and optimize highly available infrastructure supporting mission-critical applications and services.
  • Monitor production environments, analyze system performance, and proactively identify opportunities to improve stability, scalability, and operational efficiency.
  • Respond to technical escalations, troubleshoot infrastructure, networking, hardware, and software issues, and lead resolution of critical incidents.
  • Develop and maintain monitoring, alerting, backup, recovery, and disaster recovery procedures to maximize uptime and minimize business impact.
  • Manage cloud platforms, virtualization technologies, network infrastructure, and remote monitoring systems to ensure secure and reliable operations.
  • Build and maintain infrastructure automation using configuration management, scripting, and Infrastructure-as-Code tools.
  • Participate in post-incident reviews, document operational improvements, and contribute to continuous reliability and security enhancements.
  • Collaborate with engineering and operations teams to strengthen CI/CD pipelines, system resilience, and infrastructure best practices.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, or a related field; relevant professional certifications are advantageous.
  • At least 3 years of experience as a Site Reliability Engineer, DevOps Engineer, Systems Administrator, or in a similar infrastructure role.
  • Strong experience with cloud platforms such as AWS or Oracle Cloud.
  • Proficiency with Linux and Windows server administration, virtualization technologies, and enterprise infrastructure management.
  • Experience with Docker, Kubernetes, Terraform, Git, GitLab CI/CD, ELK Stack, Prometheus, and Grafana.
  • Knowledge of MySQL, PostgreSQL, networking concepts (LAN/WAN, HTTP, TCP/IP), system security, and backup/recovery strategies.
  • Experience with automation and scripting tools such as Ansible, Bash, Rundeck, or Puppet.
  • Familiarity with Nginx, PHP-FPM, SSL, DNS, and Cloudflare is considered an advantage.
  • Strong analytical thinking, troubleshooting skills, attention to detail, and the ability to work independently and collaboratively.
  • Availability to respond to critical production incidents outside standard business hours when required.

Benefits & conditions

  • Competitive salary package.
  • Private health insurance.
  • Annual wellness allowance.
  • Birthday leave.
  • Company-sponsored team-building events and social activities.
  • Relocation support, where applicable.
  • Opportunity to work with modern cloud technologies, automation tools, and large-scale infrastructure.
  • Professional development opportunities within a collaborative engineering environment.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.de

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

7:28 min

Constructing a new Docker layer from scratch

Oliver Seitz Oliver Seitz · WWC Europe 2026

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all