> Markdown version of [/jobs/ext/2529341-remote](https://www.wearedevelopers.com/jobs/ext/2529341-remote). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Remote - **Company:** Lyra Health - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $161,000.0 - $222,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Web Services, Data Analysis, Cloud Computing, Cloud Engineering, Relational Databases, Protocol Buffers, Python (Programming Language), Machine Learning, Performance Tuning, Systems Development Life Cycle, Software Engineering, Systems Architecture, AI Infrastructure, Large Language Models, Model Validation, Generative AI, Data Layers, Kotlin, AI Platforms, Kubernetes, Low Latency, Production Code, Apache Kafka, Machine Learning Operations, Celery, Restful APIs, Data Pipelines, Docker, Microservices - **Published:** August 8, 2026 - **Apply:** https://jobs.lever.co/lyrahealth/d4806b14-3818-4aca-af97-96bb88dd0b4c/apply?lever-origin=applied&lever-source%5B%5D=BuiltInNationwide ## About the Role * 8+ years of experience deploying complex ML/AI solutions into mission-critical production environments, with a proven track record of technical leadership. * Deep System & Software Engineering Expertise: Mastery of Python, RESTful API design, Protobuf, and microservices architecture. * Production AI/ML Infrastructure Mastery: Deep expertise with Docker, Kubernetes, container orchestration, and real-time inference services. * Modern Generative AI Architecture: Hands-on experience designing and deploying RAG pipelines, fine-tuning LLMs, managing vector databases, and building robust LLM evaluation/guardrail frameworks. * Data Layer Systems: Expertise with relational databases, low-latency key-value stores, distributed queueing system architectures (e.g., Celery, Kafka), and data pipeline orchestration. * Cloud Architecture: Strong experience architecting cloud-native solutions on AWS (or equivalent cloud providers). * Strategic Communication: Exceptional ability to distill highly ambiguous technical problems into clear strategic priorities and influence leadership across engineering, product, and business domain disciplines., * Polyglot Engineering Background: Experience writing high-performance production code in Java or Kotlin. * Healthcare & Sensitive Data Expertise: Experience architecting AI/ML systems within highly regulated environments (HIPAA compliance, SOC2, handling PHI/PII). * MLOps / Platform Productization: Experience building internal developer platforms or ML tooling used by dozens of data scientists and engineers. ## Description We are seeking a Staff ML/AI Engineer to define and drive the architectural vision for Lyra's machine learning and generative AI technology landscape. In this role, you will serve as a technical anchor across the engineering and data organizations-architecting enterprise-scale AI platforms, setting technical strategy for high-impact AI/ML initiatives, and ensuring our AI products operate with top-tier reliability, security, and medical precision. The ideal candidate is a seasoned technical leader who excels at translating complex healthcare challenges into scalable platform solutions, building consensus across cross-functional leadership, and elevating the technical bar for the entire engineering organization. Lyra is for you if you * Thrive on working with brilliant teammates to solve complex, meaningful problems * Are passionate about making a social impact and supporting people at their most challenging moments * Enjoy cross-functional collaboration with physicians, therapists, data scientists, data analysts and product managers Responsibilities * Drive AI Platform Architecture: Design and execute the long-term roadmap for Lyra's machine learning and generative AI platform, enabling fast, safe, and reliable deployment of frontier models across the company. * Lead AI Infrastructure Vision: Architect end-to-end training, fine-tuning, and low-latency inference platforms, including centralized RAG architecture, vector databases, and enterprise evaluation/guardrail frameworks. * Set Engineering Excellence Standards: Establish organizational standards for the full AI/ML SDLC-from dataset lineage and CI/CD pipelines to automated model evaluation, red-teaming, and production monitoring. * Cross-Functional Technical Leadership: Partner closely with Product Management, Data Science, Security, and Clinical leaders to translate strategic clinical goals into foundational AI capability roadmaps. * Mentor and Multiply Impact: Elevate the engineering culture by mentoring Senior ML Engineers, conducting high-leverage architectural reviews, and establishing engineering best practices across teams. * Hands-On Leadership: Lead by example through technical prototypes, critical-path architecture, and strategic coding contributions. * And of course, you will be coding!, By applying for this position, you acknowledge that your personal information will be processed as per the Lyra Health Workforce Privacy Notice. Through this application, to the extent permitted by law, we will collect personal information from you including, but not limited to, your name, email address, gender identity, employment information, and phone number for the purposes of recruiting and assessing suitability, aptitude, skills, qualifications, and interests for employment with Lyra. We may also collect information about your race, ethnicity, and sexual orientation, which is considered sensitive personal information under the California Privacy Rights Act (CPRA) and special category data under the UK and EU GDPR. Providing this information is optional and completely voluntary, and if you provide it you consent to Lyra processing it for the purposes as described at the point of collection, for example for diversity and inclusion initiatives. If you are a California resident and would like to limit how we use this information, please use the Limit the Use of My Sensitive Personal Information form. This information will only be retained for as long as needed to fulfill the purposes for which it was collected, as described above. Please note that Lyra does not "sell" or "share" personal information as defined by the CPRA. Outside of the United States, for example in the EU, Switzerland and the UK, you may have the right to request access to, or a copy of, your personal information, including in a portable format; request that we delete your information from our systems; object to or restrict processing of your information; or correct inaccurate or outdated personal information in our systems. These rights may be subject to legal limitations. To exercise your data privacy rights outside of the United States, please contact [email protected]. For more information about how we use and retain your information, please see our Workforce Privacy Notice." ## Related Videos - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Celery on AWS ECS - the art of background tasks & continuous deployment](https://www.wearedevelopers.com/videos/561-celery-on-aws-ecs-the-art-of-background-tasks-continuous-deployment) - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Walking into the era of Supply Chain Risks](https://www.wearedevelopers.com/videos/376-walking-into-the-era-of-supply-chain-risks) - [Why Kotlin is the better Java and how you can start using it](https://www.wearedevelopers.com/videos/661-why-kotlin-is-the-better-java-and-how-you-can-start-using-it) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)