> Markdown version of [/jobs/ext/2991743-site-reliability-engineer-ai-platform](https://www.wearedevelopers.com/jobs/ext/2991743-site-reliability-engineer-ai-platform). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, AI Platform - **Company:** Algolia - **Location:** Paris, France (Remote available) - **Salary:** €69,768.0 - €96,900.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Cloud Computing, Software Debugging, Distributed Systems, Python (Programming Language), Reliability Engineering, AI Platforms, Kubernetes, Deployment Automation, Machine Learning Operations - **Published:** September 19, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-ai-platform-algolia-2-10127596 ## About the Role We are looking for a Site Reliability Engineer with strong production fundamentals who enjoys solving operational problems, automating repetitive work and progressively taking ownership of complex systems at scale., * Solid hands-on Kubernetes knowledge, including workloads, resource management, and production operations * Strong experience with Infrastructure as Code, and the lifecycle of cloud infrastructure * Solid experience building and operating CI/CD pipelines and automated deployment workflows * Hands-on experience with at least one major cloud provider: GCP, AWS or Azure * Good understanding of networking, distributed systems and reliability engineering * Experience with monitoring, observability and troubleshooting production systems * Strong automation mindset and the ability to take ownership of well-defined production systems and progressively tackle more complex problems * Excellent written and spoken English NICE TO HAVE: * Go and/or Python engineering experience * Exposure to AI/ML infrastructure and inferences * Comfortable working AI-first, using coding agents, agentic workflows and AI-assisted debugging to accelerate engineering and operations, * GRIT - Problem-solving and perseverance capability in an ever-changing and growing environment. * TRUST - Willingness to trust our co-workers and to take ownership. * CANDOR - Ability to receive and give constructive feedback. ## Description * Build and operate production infrastructure supporting AI-related workloads and services * Operate and improve highly available Kubernetes-based platforms * Improve reliability through SLOs, observability, alerting and capacity management * Investigate production issues and turn findings into durable fixes and improvements * Work across networking, databases, compute and service infrastructure * Improve CI/CD pipelines, deployment automation and developer experience * Build and maintain infrastructure using Infrastructure as Code * Participate in on-call, incident response and operational improvements * Collaborate with experienced engineers across AI Platform and progressively take ownership of broader production areas ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [This App Reached 10,000 Users in One Week. Here's How.](https://www.wearedevelopers.com/videos/100329-this-app-reached-10-000-users-in-one-week-here-s-how) - [Developer Tools for Microsoft Azure](https://www.wearedevelopers.com/videos/450-developer-tools-for-microsoft-azure) - [Azure AI Foundry for Developers: Open Tools, Scalable Agents, Real Impact](https://www.wearedevelopers.com/videos/1541-azure-ai-foundry-for-developers-open-tools-scalable-agents-real-impact) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)