> Markdown version of [/jobs/ext/1981368-engineering-manager-site-reliability-observability](https://www.wearedevelopers.com/jobs/ext/1981368-engineering-manager-site-reliability-observability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Engineering Manager - Site Reliability & Observability - **Company:** DOCTOLIB SAS - **Location:** Paris, France - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Cloud Computing, Elasticsearch, Key Management, Reliability Engineering, Prometheus, Software Engineering, Datadog, Data Logging, Technical Debt, Infrastructure Automation Frameworks, Terraform - **Published:** August 8, 2026 - **Apply:** https://eu.experteer.com/career/view-jobs/engineering-manager-site-reliability-and-observability-x-f-m-paris-iledefrance-frankreich-58835350 ## About the Role and * Improve developer experience by reducing technical debt and enabling scalable infrastructure * Own on-call experience and contribute to incident response and postmortems * Collaborate with Product, Engineering, and architecture teams to align reliability with product goals * Represent the team in engineering leadership forums and architectural reviews * Foster partnerships with software engineering to embed reliability early in the lifecycle Tasks * 5+ years software engineering or SRE experience in cloud-native environments * 3+ years engineering management experience * Strong observability tooling knowledge (OpenTelemetry, Prometheus, Datadog, Elasticsearch) * Experience with infrastructure as code (Terraform) and secrets management * Ability to mentor engineers, review designs, and guide architecture * Fluent in English Key requirements * health insurance * 25 days vacation + RTT * mental health and coaching services * flexibility days (work from abroad) * lunch vouchers * transport subsidy reimbursement 50% of public transport subscription ## Description Experteer Overview As Engineering Manager for the SRE team, you will lead a group of Site Reliability Engineers to ensure Doctolib's platform is reliable, scalable, and resilient at European scale. You will set reliability and observability direction and collaborate with product and engineering teams to enable safe, fast delivery. You'll drive large-scale reliability initiatives, SLOs, and incident prevention across hundreds of applications. Your leadership will foster a culture of operational excellence and continuous improvement while aligning with the company mission to improve healthcare access. Pay / Benefits * Lead and grow a team of Site Reliability Engineers and shape their technical growth * Define and evolve reliability and observability strategy (infrastructure automation, logging, metrics, tracing, alerting) * Drive roadmap for large-scale reliability initiatives, including SLOs and error budgets * Own transversal services (secrets management, infrastructure as code tooling) * Improve developer experience by reducing technical debt and enabling scalable infrastructure * Own on-call experience and contribute to incident response and postmortems * Collaborate with Product, Engineering, and architecture teams to align reliability with product goals * Represent the team in engineering leadership forums and architectural reviews * Foster partnerships with software engineering to embed reliability early in the lifecycle Tasks * 5+ years software engineering or SRE experience in cloud-native environments * 3+ years engineering management experience * Strong observability tooling knowledge (OpenTelemetry, Prometheus, Datadog, Elasticsearch) * Experience with infrastructure as code (Terraform) and secrets management * Ability to mentor engineers, review designs, and guide architecture * Fluent in English Key requirements * health insurance * 25 days vacation + RTT * mental health and coaching services * flexibility days (work from abroad) * lunch vouchers * transport subsidy reimbursement 50% of public transport subscription ## Related Videos - [Handling incidents collaboratively is like solving a rubix cube](https://www.wearedevelopers.com/videos/680-handling-incidents-collaboratively-is-like-solving-a-rubix-cube) - [Data binning and understanding histograms](https://www.wearedevelopers.com/videos/2086-data-binning-and-understanding-histograms) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [The OpenTelemetry mistakes I keep seeing (and how to stop making them)](https://www.wearedevelopers.com/videos/100158-the-opentelemetry-mistakes-i-keep-seeing-and-how-to-stop-making-them) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Implementing Feature Environments with AWS and Terraform](https://www.wearedevelopers.com/videos/531-implementing-feature-environments-with-aws-and-terraform) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best Companies to Work For in Paris: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/190-best-companies-to-work-for-in-paris-top-25-companies-in-2023) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Does The Tech Industry Have The Best Work-life Balance?](https://www.wearedevelopers.com/magazine/427-does-the-tech-industry-have-the-best-work-life-balance)