> Markdown version of [/jobs/ext/673445-ai-engineer](https://www.wearedevelopers.com/jobs/ext/673445-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Engineer - **Company:** Lever, Inc. - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $182,300.0 - $220,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Automated Storage and Retrieval Systems, Code Review, Python (Programming Language), Open Source Technology, Software Engineering, Systems Integration, Large Language Models, Kubernetes, Low Latency, Software Coding, Software Version Control - **Published:** June 27, 2026 - **Apply:** https://jobs.lever.co/ro/81dd41da-ba24-435b-83c4-bd19a312744b ## About the Role We're hiring a Senior AI Engineer to join a new team we're building from the ground up alongside a Senior AI Lead and an Engineering Manager. You'll own AI features from prototype through production and spend as much time improving workflows with clinical operators as writing code. Most of your work will involve building orchestration layers, prompts, evaluations, retrieval systems, and tooling that make LLMs reliable enough for real-world healthcare operations., * 5+ years building production software, with at least 1-2 years building and deploying LLM-powered applications in production (or equivalent depth through substantial side projects or open-source work). * Strong Python engineer with experience building backend services and integrating modern AI APIs and frameworks. * Comfortable using AI-assisted development tools throughout the software development lifecycle. * You've shipped AI features that served real users, and you can clearly articulate how you measured quality, reliability, latency, cost, and business impact. * You have strong product instincts and thrive in ambiguous environments, translating operational problems into technical solutions without relying on detailed specifications. * Bonus: Experience with orchestration frameworks (LangChain, LangGraph, AgentCore), RAG architectures, vector databases, prompt/version management, and evaluation or observability platforms (e.g., LangSmith, Braintrust, Arize Phoenix). ## Description Ro is building a team focused on shipping LLM-powered products across the patient experience, clinical operations, and internal tooling., * Build the AI application layer, including prompt orchestration, tool calling, retrieval (RAG), embeddings, agent workflows, and structured outputs. * Develop evaluation infrastructure, including regression suites, LLM-as-a-judge evaluations, synthetic datasets, human review workflows, and quality metrics that catch regressions before users do. * Design observability systems that measure accuracy, latency, cost, hallucination rates, and model behavior in production. * Build safety and reliability guardrails that ensure AI systems meet quality, compliance, and operational requirements. * Evaluate and integrate third-party tools, including orchestration frameworks, evaluation platforms, vector databases, and model providers; and make pragmatic build-versus-buy decisions. * Raise engineering standards across the team through technical leadership, code reviews, mentorship, and thoughtful system design. * Sit with EHR operators, Patient Advocates, and clinical teams to understand the workflows you're automating. * Prototype quickly, ship early, measure outcomes, and iterate based on real-world usage. ## Related Videos - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Are Code Reviews Worth It? Insights from 16 Years of Review Data](https://www.wearedevelopers.com/videos/1135-are-code-reviews-worth-it-insights-from-16-years-of-review-data) - [Swapping Low Latency Data Storage Under High Load](https://www.wearedevelopers.com/videos/746-swapping-low-latency-data-storage-under-high-load) - [Building AI Applications with LangChain and Node.js](https://www.wearedevelopers.com/videos/1512-building-ai-applications-with-langchain-and-node-js) - [AI Killed DevOps... What Now? - Lee Faus](https://www.wearedevelopers.com/videos/1759-ai-killed-devops-what-now-lee-faus) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)