> Markdown version of [/jobs/ext/2726650-site-reliability-engineering-team-sre](https://www.wearedevelopers.com/jobs/ext/2726650-site-reliability-engineering-team-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineering team (SRE) - **Company:** BlaBlaCar - **Location:** Paris, France (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Github, Python (Programming Language), Nagios, Object-Oriented Software Development, Reliability Engineering, Software Engineering, Systems Integration, Datadog, Data Logging, Scripting, Google Cloud, Grafana, Saucelabs, Backend, Kubernetes, Playwright, Terraform, Jenkins, Programming Languages - **Published:** September 5, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-blablacar-3-8012139 ## About the Role * 3 to 7 years of experience in SRE, DevOps, or Software Engineering roles * Working in a multidisciplinary environment will request strong communication skills : you'll need to adapt your communication level to other teams expertise and be able to understand their needs * Strong knowledge of observability tools (e.g., Datadog) and understanding of metrics, logging, and tracing. * Troubleshooting/oncall experience in production environments, diagnosing and resolving technical issues effectively (experience with Kubernetes is a plus). * Full working proficiency in English * Fit with our BlaBlaPrinciples * Thriving in a collaborative, fast-growing and innovative environment * Ability to take ownership, aligned with business priorities and navigating in different contexts * Nice to have: + Familiarity with incident management platforms (e.g., Grafana IRM) is a bonus + Experience working with Service Level Objectives (SLOs) and Service Level Indicators (SLIs) + Exposure to programming in Go or a strong interest in learning it. + Experience in integrating Opentelemetry + Backend services are built using multiple programming languages: while development skills aren't required, familiarity with object-oriented programming and scripting languages is an advantage. + Familiarity with web/mobile testing tools or a strong curiosity to understand how software is tested at scale. ## Description By joining our Foundations department, you will be working alongside talented individuals grouped in small agile teams that each have strong ownership on their piece of these goals. Foundations is composed of seven teams which "provide consistent, easy to use, infrastructures, services, and expertise to support BlaBlaCar's growth and evolution". The Site Reliability Engineering team (SRE) is responsible to provide best in class Observability, Alerting and Incident management tools and processes to service teams. As an enabling team, we help BlaBlacar engineers to efficiently improve their service reliability. Empowering developers and bringing them our reliability expertise are at the core of our daily work. Technical stack: * Core Infrastructure: Kubernetes, Google Cloud Platform * GitOps/Delivery: GitHub, Terraform, Flux, Helm, Jenkins * Observability/Incident Management: Datadog, Opentelemetry, Grafana IRM, * In house Synthetic Tests platform: Playwright, Qualcium, SauceLabs * Languages: Go / Python for Tooling, Typescripts/JS for the testing platform, * Support software engineers by creating, maintaining, and improving observability and alerting tools and frameworks. You embrace the use of AI, leveraging agentic to eliminate toil and streamline your daily tasks * Own the Service Level Objectives (SLOs) framework, assist in the design and maintenance of indicators (SLI) and objectives to ensure service reliability. * Owning the incident management process by defining best practices, standards, and ensuring continuous improvement through post-mortems and chaos engineering. While developers handle incidents within their scope, you could step in as Incident Commander during high-severity incidents, leading coordination efforts . * Develop and maintain tools, such as Terraform modules or Go apps, to help automate and enhance reliability across services. * Build and promote reporting on operational metrics and incidents to drive distributed and continuous improvement., * Hybrid status for this role : 2-3 days at the Office * 4 additional weeks on top of legal maternity/paternity leaves * 50% healthcare coverage (Alan) * Financial support for home office equipment * Minimum 25 days holiday per year * Local meal plan policy (Swile card) * 50% transportation paid (Forfait Mobilité Durable) * Free unlimited carpooling & bus rides * Personal growth via trainings, mentorship, and internal mobility opportunities * Employee Stock ownership plan * Regular team building events * 1 day off per year to test our product, * a 45-min video-call with one of our Talent Acquisition Manager, to get to know you, understand your career expectations and answer your questions * a 60-min video-call with Damien Bertau, Hiring Manager, to discuss your experience and share more details about the team * a 90-min system design interview with 2 team members to discuss about your technical expertise * a 45-min video-call with Maxime Fouilleul, Head of Foundations, to get a wider vision of the department and its strategy ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs)