> Markdown version of [/jobs/ext/130820-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/130820-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Flip GmbH - **Location:** Konigsbach-Stein, Germany (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Microsoft Azure, Software as a Service, Cloud Computing, Cloud Engineering, Continuous Integration, DevOps, Disaster Recovery, Python (Programming Language), PostgreSQL, Reliability Engineering, Prometheus, Software Engineering, Pulumi, Istio, Grafana, Backend, Kubernetes, Linkerd (Service Mesh), Terraform - **Published:** May 28, 2026 - **Apply:** https://de.indeed.com/viewjob?jk=0ebc753eb4252e07 ## About the Role Do you have experience in Terraform?, We're looking for a hands-on, SaaS-minded senior Site Reliability Engineer who treats scalability and reliability as a first-class product concern., * 5+ years of hands-on experience as a Site Reliability Engineer (SRE), Platform Engineer, DevOps Engineer, Infrastructure Engineer, Cloud Engineer, or Backend Engineer with a strong infrastructure focus. * Proven track record building and operating high-throughput, highly available systems in production. * Deep, production-level experience with Kubernetes on any Hyperscaler. * Strong experience with modern observability stacks (e.g. Prometheus, Mimir, VictoriaMetrics, Dash0, Loki, ELK) and a clear point of view on SLIs, SLOs and error budgets. * Solid software development skills in Go (strongly preferred, since our IaC runs on Pulumi in Go) or Python. * Hands-on experience with Infrastructure as Code (Pulumi, OpenTofu, Terraform) and GitOps (e.g. ArgoCD) + CI/CD pipeline design. * Demonstrated ability to lead complex infrastructure initiatives from design to production - including writing RFCs and driving architecture decisions within your team. * Experience mentoring engineers and raising the technical bar within a team. * Comfortable owning major incidents end-to-end and turning learnings into systemic change. * Strong communication skills and business-fluent English. * Willingness to participate in on-call rotations to ensure the reliability of our platform., * Rolled out production-ready API-Gateways with Gateway API (e.g. Envoy Gateway). * Operated multi-cluster service meshes (e.g. Cilium, Linkerd, Istio) * Deployed and maintained Kubernetes Operators (e.g. Strimzi, CNPG). * Operated highly available PostgreSQL in production. ## Description As a Senior Site Reliability Engineer in our Platform Squad, you'll own critical reliability domains end-to-end and drive the technical direction within the squad - leading architectural decisions on our platform, mentoring teammates, and continuously raising the reliability bar inside the team. This role is for an engineer with a proven track record of building and operating high-throughput, highly available systems, who wants senior-level technical ownership and real impact through deep engineering work inside a tight, well-scoped team. What awaits you with us * Co-own the architecture: Help drive the architecture and evolution of our cloud infrastructure on Azure and our Kubernetes clusters - designed for high throughput and highest availability - to support Flip's rapid growth across the globe. * Drive the resilience strategy: Define how we approach global scaling, zero-downtime deployments, rollback mechanisms and disaster recovery, and make sure the platform stays available around the clock. * Evolve our observability stack: Improve our LGTM stack (Loki, Grafana, Tempo, Mimir) into a foundation our engineers can trust. * Improve our IaC Platform: Eliminate toil at the source, and make our infrastructure truly self-service for engineering teams. * Lead in incidents: Take a leading role in platform-related major incidents, drive blameless post-mortems for the squad, and translate findings into systemic improvements. * Mentor within the squad: Coach teammates, run RFCs and design reviews inside the team, and help engineers grow into stronger SREs. * Shape our roadmap: Partner with your squad to define the platform's direction. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Empowering Thousands of Developers: Our Journey to an Internal Developer Platform](https://www.wearedevelopers.com/videos/1519-empowering-thousands-of-developers-our-journey-to-an-internal-developer-platform) - [Get started with securing your cloud-native Java microservices applications](https://www.wearedevelopers.com/videos/123-get-started-with-securing-your-cloud-native-java-microservices-applications) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) ## Related Articles - [The Biggest German Tech Companies](https://www.wearedevelopers.com/magazine/424-the-biggest-german-tech-companies) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Finding Jobs in Germany](https://www.wearedevelopers.com/magazine/375-finding-jobs-in-germany) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)