> Markdown version of [/jobs/ext/2009082-applied-ai-engineer-site-reliability-engineer-emea](https://www.wearedevelopers.com/jobs/ext/2009082-applied-ai-engineer-site-reliability-engineer-emea). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Applied AI Engineer, Site Reliability Engineer - EMEA - **Company:** Mistral AI - **Location:** Paris, France - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Cyber Security, Computer Networks, DevOps, Python (Programming Language), Linux Kernel, Systems Development Life Cycle, Role-Based Access Control, Reliability Engineering, Site Reliability Engineering Practices, Ansible, Prometheus, Google Cloud, Large Language Models, Grafana, Software Security, Multi-Cloud, Build Management, Kubernetes, Deployment Automation, CIS Benchmarks, Terraform, Golang - **Published:** August 10, 2026 - **Apply:** https://eu.experteer.com/career/view-jobs/applied-ai-engineer-site-reliability-engineer-emea-paris-iledefrance-frankreich-58867687 ## About the Role Collaborate with Technical Support as L3 escalation, perform blameless post-mortems * Automate provisioning and enforce secure-config baselines * Lead security and operational excellence across Mistral-hosted and customer-hosted fleets Tasks * Fluent in English * 5+ years in SRE, Production Engineering, or DevOps with tooling shipping track record * Strong multi-tenant Kubernetes fluency (namespaces, network policies, RBAC, admission control, scale) * On-call discipline, incident response, blameless post-mortem culture * Observability stack in production (Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz) * Infrastructure as code (Terraform, Ansible or equivalents) * Proficiency in Python and/or Golang for tooling and automation * Security mindset: secure-SDLC, CVE response, supply-chain integrity as reliability properties * Strong written communication (runbooks, post-mortems, customer incident comms) * Ability to operate autonomously in ambiguous, fast-paced environments * Solid Linux internals, networking, distributed-systems fundamentals * Nice to have: cloud/app security (AppSec, K8s security, SBOM, cosign, SLSA) * Nice to have: production experience with LLM/model-serving stacks * Nice to have: multi-cloud or on-prem hybrid environments (AWS, GCP, Azure, sovereign clouds) * Nice to have: open-source contributions in SRE/observability/security tooling Key requirements * healthcare coverage * parental leave * retirement plans * relocation support * wellness programs * meal and transportation allowances ## Description Experteer Overview As a founding member of the Applied AI SRE sub-team, you will design and operate a fleet-focused reliability framework for Mistral's AI solutions. You'll scale SRE practices across Mistral-hosted and customer-hosted deployments, driving observability, runbooks, and security guardrails. You'll own on-call incidents, post-mortems, and CVE response while partnering with product and security teams. This is a chance to shape how we deliver robust, scalable AI for enterprise clients at scale. You'll work across a fast-paced environment with a strong emphasis on impact, ownership, and collaboration. Pay / Benefits * Design and build a fleet-wide reliability framework (SLOs, observability, runbooks) * Run Tier-1 customer environments, maintain SLO compliance, on-call and incident response * Productize deployment, security baselines, and scale of Applied AI solutions * Own security operations for customer deployments, including CVE response and supply-chain integrity * Collaborate with Technical Support as L3 escalation, perform blameless post-mortems * Automate provisioning and enforce secure-config baselines * Lead security and operational excellence across Mistral-hosted and customer-hosted fleets Tasks * Fluent in English * 5+ years in SRE, Production Engineering, or DevOps with tooling shipping track record * Strong multi-tenant Kubernetes fluency (namespaces, network policies, RBAC, admission control, scale) * On-call discipline, incident response, blameless post-mortem culture * Observability stack in production (Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz) * Infrastructure as code (Terraform, Ansible or equivalents) * Proficiency in Python and/or Golang for tooling and automation * Security mindset: secure-SDLC, CVE response, supply-chain integrity as reliability properties * Strong written communication (runbooks, post-mortems, customer incident comms) * Ability to operate autonomously in ambiguous, fast-paced environments * Solid Linux internals, networking, distributed-systems fundamentals * Nice to have: cloud/app security (AppSec, K8s security, SBOM, cosign, SLSA) * Nice to have: production experience with LLM/model-serving stacks * Nice to have: multi-cloud or on-prem hybrid environments (AWS, GCP, Azure, sovereign clouds) * Nice to have: open-source contributions in SRE/observability/security tooling Key requirements * healthcare coverage * parental leave * retirement plans * relocation support * wellness programs * meal and transportation allowances ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)