> Markdown version of [/jobs/ext/48235-senior-ai-ml-engineer](https://www.wearedevelopers.com/jobs/ext/48235-senior-ai-ml-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI/ML Engineer - **Company:** CLERA, LLC - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $180,000.0 - $280,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Software Quality, Distributed Computing Environment, Python (Programming Language), Performance Tuning, Recommender Systems, Pytorch, Large Language Models, Data Pipelines - **Published:** May 22, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=c65ea5f4d78d3f08 ## About the Role Do you have experience in Production systems?, * 4+ years of applied ML engineering in production environments * Hands-on experience with LLMs, fine-tuning, RAG, or large-scale recommender systems * Strong Python and PyTorch (or JAX) fundamentals * Experience with distributed training, GPU optimization, or inference serving * Pragmatic about trade-offs between research-grade and ship-grade work ## Description * Design and ship end-to-end ML systems: data pipelines, training, evaluation, deployment * Own model performance, latency, and cost trade-offs in production * Build evaluation harnesses and offline benchmarks for fast iteration * Work directly with product to translate ambiguous goals into measurable model improvements * Mentor other engineers on ML best practices and code quality ## Related Videos - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Why and when should we consider Stream Processing frameworks in our solutions](https://www.wearedevelopers.com/videos/1085-why-and-when-should-we-consider-stream-processing-frameworks-in-our-solutions) - [Let's Talk Quality! - Lilia Gargouri](https://www.wearedevelopers.com/videos/1815-let-s-talk-quality-lilia-gargouri) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [How AI Models Get Smarter](https://www.wearedevelopers.com/videos/1374-how-ai-models-get-smarter) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix)