> Markdown version of [/jobs/ext/122636-senior-site-reliability-engineer-observability](https://www.wearedevelopers.com/jobs/ext/122636-senior-site-reliability-engineer-observability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - Observability - **Company:** Doctolib - **Location:** Berlin, Germany - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Web Services, Computer Vision, Microsoft Azure, Elasticsearch, Monitoring of Systems, Mobile Application Software, Python (Programming Language), Open Source Technology, Performance Tuning, Reliability Engineering, Logstash, Prometheus, Ruby, TypeScript, Datadog, Data Logging, Google Cloud, Swift (Programming Language), Kotlin, Kubernetes, React Native, Docker, Golang, Programming Languages - **Published:** May 13, 2026 - **Apply:** https://de.indeed.com/viewjob?jk=261a4b19c3489f49 ## About the Role Do you have experience in TypeScript?, * Have a solid hands-on experience (3y+) on a large-scale production platform * Have proven experience with cloud platforms such as AWS, Azure or Google Cloud * Have solid understanding of containerization and orchestration technologies (Docker and Kubernetes) * Have a strong understanding of Helm for managing Kubernetes manifests and ArgoCD for GitOps workflows * Have deep expertise in observability tooling and architecture, such as: + Logging: Fluent Bit, OpenTelemetry, Loki, Elasticsearch, Logstash, Vector + Tracing: OpenTelemetry or proprietary APMs + Metrics: Prometheus, Thanos, Datadog, or equivalent * Have proficiency in at least one programming language (Ruby, Python, Go, Java, etc.) and a deep understanding of infrastructure as code principles * Have experience with monitoring and observability tools * Like troubleshooting performance issues in complex environments * Are fluent in English It would be fantastic if you: * Have experience contributing to open-source observability projects * Have worked in a high-growth tech environment * Are passionate about developer experience and platform engineering ## Description Your mission will be to shape Doctolib's observability strategy and ensure our platform remains reliable, debuggable, and scalable at a European scale. You will work in a feature team developing logging, metrics, tracing, and alerting capabilities, contributing directly to supporting 400,000 health professionals and 80 million patients in their daily healthcare journey. Working in the tech team at Doctolib means building innovative products and features to improve the daily lives of care teams and patients., * Lead the observability strategy across the platform, with an emphasis on building scalable, developer-friendly logging and tracing capabilities * Identify and lead large-scale cross-cutting reliability initiatives, including improvements to our incident detection, response, and postmortem analysis capabilities * Take part in the on-call rotation, and actively contribute to improving our on-call experience by refining alerting, reducing noise, and ensuring actionable telemetry, * Our solutions are built on a single fully cloud-native platform that supports web and mobile app interfaces, multiple languages, and is adapted to country and healthcare specialty requirements. * Our stack is composed of Rails, TypeScript, Java, Python, Kotlin, Swift, and React Native. * We leverage AI ethically across our products to empower patients and health professionals. Discover our AI vision here. ## Related Videos - [Navigating the Corporate Jungle: Life as a Developer in a large Company](https://www.wearedevelopers.com/videos/621-navigating-the-corporate-jungle-life-as-a-developer-in-a-large-company) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Planet-Scale Dashboards](https://www.wearedevelopers.com/videos/1618-planet-scale-dashboards) - [Fireside Chat with Werner Vogels, VP & CTO, Amazon.com & Daniel Gebler, CTO at Picnic](https://www.wearedevelopers.com/videos/1405-fireside-chat-with-werner-vogels-vp-cto-amazon-com-daniel-gebler-cto-at-picnic) ## Related Articles - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The Biggest German Tech Companies](https://www.wearedevelopers.com/magazine/424-the-biggest-german-tech-companies) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [7 Most Popular Web Developer Jobs in Europe](https://www.wearedevelopers.com/magazine/163-7-most-popular-web-developer-jobs-in-europe) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)