> Markdown version of [/jobs/ext/3007885-site-reliability-engineer-ai-platform](https://www.wearedevelopers.com/jobs/ext/3007885-site-reliability-engineer-ai-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, AI Platform - **Company:** Algolia - **Location:** Paris, France (Remote available) - **Experience:** Expert - **Salary:** €69,768.0 - €96,900.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Engineering, Software Debugging, Distributed Systems, Python (Programming Language), Reliability Engineering, Graphics Processing Unit (GPU), Kubernetes, Machine Learning Operations - **Published:** September 20, 2026 - **Apply:** https://startup.jobs/senior-site-reliability-engineer-ai-platform-algolia-2-10127594 ## About the Role We are looking for a Senior Site Reliability Engineer who can independently own complex production systems, drive technical decisions across teams, and help shape reliable and efficient infrastructure at scale., * Strong hands-on production experience with at least one major cloud provider: GCP, AWS or Azure * Strong experience designing and operating Kubernetes and cloud-native production systems at scale * Strong understanding of distributed systems, networking and reliability engineering * Experience operating business-critical systems with strong availability, scalability and operational requirements * Ability to independently own ambiguous, cross-team technical problems and drive them to measurable outcomes * Strong automation mindset and ability to balance reliability, engineering velocity and cost * Excellent written and spoken English NICE TO HAVE: * Go and/or Python engineering experience * Experience with infrastructure supporting AI/ML workloads, model serving, GPUs or other compute-intensive systems * Comfortable working AI-first, using coding agents, agentic development workflows, AI-assisted debugging and automation to accelerate engineering and operations, * GRIT - Problem-solving and perseverance capability in an ever-changing and growing environment. * TRUST - Willingness to trust our co-workers and to take ownership. * CANDOR - Ability to receive and give constructive feedback. ## Description * Own and evolve production infrastructure supporting AI-related workloads and services at scale * Design and operate highly available Kubernetes-based platforms * Drive reliability through SLOs, observability, capacity planning and production guardrails * Lead complex production investigations and turn findings into durable architectural improvements * Improve shared infrastructure across networking, databases, service communication and compute * Build better CI/CD, progressive delivery, automation and developer experience * Drive cloud infrastructure efficiency and FinOps initiatives * Participate in and improve on-call and incident response * Mentor engineers and raise the technical bar for reliability and production engineering ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [The Cloud is Calling: Answer with In-Demand Skills](https://www.wearedevelopers.com/videos/945-the-cloud-is-calling-answer-with-in-demand-skills) - [Using AI Without Losing Your Skills](https://www.wearedevelopers.com/videos/2045-using-ai-without-losing-your-skills) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Navigating the AI Wave in DevOps](https://www.wearedevelopers.com/videos/853-navigating-the-ai-wave-in-devops) ## Related Articles - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)