> Markdown version of [/jobs/ext/1189253-sr-ml-engineer-remote](https://www.wearedevelopers.com/jobs/ext/1189253-sr-ml-engineer-remote). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr ML Engineer - REMOTE - **Company:** Insight Global - **Location:** Raleigh, NC, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Microsoft Azure, Decision-Making Software, Distributed Systems, Python (Programming Language), Lexis, Cloud Platform System, Large Language Models, Containerization, AI Platforms, Kubernetes, Machine Learning Operations - **Published:** July 5, 2026 - **Apply:** https://www.juju.com/job/00000000gdz5vj ## About the Role 10+ years building production ML systems (not research-only) Hands-on LLM experience in production, including: RAG architectures Inference performance, reliability, and monitoring Experience designing agentic AI systems (models calling tools/APIs, multi-step workflows) Strong distributed systems architecture experience in cloud environments (AWS, Azure, or GCP) Kubernetes + containerization experience in production environments Strong Python engineering background (platform-level code, not just notebooks) Experience building or contributing to enterprise AI platforms used by multiple teams Proven ability to lead technically (set standards, mentor engineers, influence architecture) Comfortable working in regulated or high-reliability environments Direct experience with Model Context Protocol (MCP) servers or structured tool-calling frameworks Deep experience with vector databases and large-scale search systems Experience designing LLMOps / MLOps standards at the platform level Prior work in legal, financial, healthcare, or other regulated industries Exposure to Responsible AI governance, auditing, or compliance frameworks Experience building internal AI platforms rather than just end-user applications Background mentoring senior engineers or leading cross-team technical initiatives ## Description Day to Day: Designing and owning the architecture for enterprise AI platforms used across multiple LexisNexis products Building and scaling LLM-powered systems (including RAG pipelines) that support legal research and decision-making tools Designing agentic AI workflows where models reason, call tools/APIs, and execute multi-step tasks Creating high-availability, low-latency inference systems for global, enterprise users Establishing platform standards for model deployment, monitoring, evaluation, and reliability Defining guardrails, permissions, and auditability for AI systems in a regulated legal environment Working closely with product, platform, and engineering teams to ensure AI systems are reusable and scalable Mentoring senior engineers and influencing technical direction across teams Ensuring Responsible AI principles are embedded into system design (safety, reliability, governance) ## Related Videos - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [This App Reached 10,000 Users in One Week. Here's How.](https://www.wearedevelopers.com/videos/100329-this-app-reached-10-000-users-in-one-week-here-s-how) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [How to Avoid LLM Pitfalls - Mete Atamel and Guillaume Laforge](https://www.wearedevelopers.com/videos/1328-how-to-avoid-llm-pitfalls-mete-atamel-and-guillaume-laforge) - [From AI Assistance to Agentic Systems: Scaling Sovereign AI in Banking](https://www.wearedevelopers.com/videos/100070-from-ai-assistance-to-agentic-systems-scaling-sovereign-ai-in-banking) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)