Applied AI Engineer, Site Reliability Engineer - EMEA

Mistral Inc
Paris, TX, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Cyber Security Computer Networks Software Debugging DevOps Python (Programming Language) Linux Kernel Systems Development Life Cycle Role-Based Access Control Reliability Engineering Ansible Prometheus
+6 more
Grafana Kubernetes Deployment Automation CIS Benchmarks Terraform Golang

Job description

Experteer Overview You will join as a founding engineer in the Applied AI SRE sub-team, shaping the reliability framework for Mistral’s enterprise deployments. You’ll design and operate a fleet-wide platform with a focus on reliability, security, and scalable delivery across hosted and customer-hosted environments. The role blends hands-on SRE work with productized tooling to reduce toil and ensure consistent SLO compliance. You’ll work closely with cross-functional teams to translate AI solutions into robust, scalable customer outcomes. This position offers the chance to influence how Mistral delivers enterprise AI at scale and across accounts. Compensation / Benefits * BUILD: design and productize reliability for a fleet of Mistral platforms and apps; create SLO templates and runbooks; implement observability. * RUN: manage Tier-1 customer environments; ensure SLOs, on-call, incident response, drift management, and L3 escalation. * ENABLE: automate deployment, security baselines, and scale for Applied AI solutions; provide on-demand provisioning and guardrails. * SECURE: lead security operations for customer deployments; manage CVE responses and supply-chain integrity; coordinate with InfoSec. Tasks * 5+ years in SRE, Production Engineering, or DevOps with a track record shipping tooling * Strong multi-tenant Kubernetes fluency (namespaces, network policy, RBAC, admission control, large-scale operations) * On-call discipline with blameless post-mortems and runbook-first mindset * Observability stack experience (Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz) * Infrastructure as code (Terraform, Ansible or equivalents) * Proficient in Python and/or Golang for tooling * Security mindset: secure-SDLC, CVE response, and supply-chain integrity as reliability properties * Strong written communication for runbooks, post-mortems, and customer-facing incident communications * Comfort operating autonomously in a fast-paced, ambiguous environment; defend team scope when needed * Solid Linux internals, networking debugging, and distributed-systems fundamentals Key requirements * healthcare coverage * parental leave * retirement plans * relocation support * wellness programs * meal and transportation allowances

Requirements

  • scale for Applied AI solutions; provide on-demand provisioning and guardrails. * SECURE: lead security operations for customer deployments; manage CVE responses and supply-chain integrity; coordinate with InfoSec. Tasks * 5+ years in SRE, Production Engineering, or DevOps with a track record shipping tooling * Strong multi-tenant Kubernetes fluency (namespaces, network policy, RBAC, admission control, large-scale operations) * On-call discipline with blameless post-mortems and runbook-first mindset * Observability stack experience (Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz) * Infrastructure as code (Terraform, Ansible or equivalents) * Proficient in Python and/or Golang for tooling * Security mindset: secure-SDLC, CVE response, and supply-chain integrity as reliability properties * Strong written communication for runbooks, post-mortems, and customer-facing incident communications * Comfort operating autonomously in a fast-paced, ambiguous environment; defend team scope when needed * Solid Linux internals, networking debugging, and distributed-systems fundamentals Key requirements * healthcare coverage * parental leave * retirement plans * relocation support * wellness programs * meal and transportation allowances

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · WWC 2024

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:33 min

Case study on adopting Kubernetes and Golang effectively

Andrew Holway · LIVE

Videos

See all

Related articles

See all