> Markdown version of [/jobs/ext/3524074-staff-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3524074-staff-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer - **Company:** IONOS SE - **Location:** Berlin, Germany - **Contract:** Permanent contract - **Skills:** Agile Methodology, Configuration Management, Computer Programming, Computer Networks, Linux, Python (Programming Language), Open Source Technology, Reliability Engineering, Prometheus, Scripting, Grafana, Kubernetes - **Published:** October 2, 2026 - **Apply:** https://www.adzuna.de/details/5902913309 ## About the Role * You have a keen interest in infrastructure topics, good network knowledge. * Hands-on experience with the administration of complex services on Linux/K8s. * Infrastructure as code and automation are important parts of your own way of working, like us, you are a fan of Open Source, especially Linux. * Good knowledge with system-level development of highly available software artifacts, written in programming or scripting languages like Go and Python. * Has worked with observability stacks (Loki, Prometheus, Mimir, Grafana). * The ability to communicate with an international team fluently in English, German is a benefit. * Willingness to undergo extended security vetting (SÜ2). ## Description In a culture with an emphasis on creative ideas, agile working methods, initiative, and participation you can expect the following tasks: Tasks * Provision, operate and migrate distributed, highly available services on Linux / K8s. * Improve and develop the full stack from hardware over OS to the application including configuration management and monitoring. * Development and maintain in-house built K8s operators. * Support our developers to automate operational tasks and rollouts, provide interfaces for integration for other teams. * Administration and troubleshooting of our highly available and complex infrastructure including participation in an on-call service rotation. * Support company goals through efficient operating architecture, active lifecycle management and lean processes in an agile DevOps environment, operate certified, security-relevant infrastructure.