Machine Learning Engineer Lead

Compunnel Inc.
Raleigh, NC, United States
16 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Architectural Patterns Microsoft Azure Continuous Integration Python (Programming Language) Google Cloud Cloud Platform System Large Language Models Reliability of Systems Generative AI
+5 more
Containerization AI Platforms Kubernetes Low Latency Machine Learning Operations

Requirements

reasoning-driven agentic AI systems. Design orchestration patterns for tool use, API invocation, and structured function calling. Lead the implementation and governance of Model Context Protocol (MCP) servers to standardize tool integration and context management. Define guardrails, permissions, security controls, and audit mechanisms for enterprise-safe AI systems. Establish and maintain best practices for MLOps, CI/CD, observability, scalability, and system reliability. Design and implement scalable inference systems using containerization and Kubernetes. Drive the deployment and optimization of LLM, Generative AI, and RAG solutions in production environments. Design cloud-based AI/ML architectures across AWS, Azure, or Google Cloud Platform. Establish technical standards and architectural patterns for AI/ML and agentic systems across engineering teams. Embed Responsible AI principles into platform architecture and engineering practices. Provide technical leadership, mentorship, and guidance to senior engineers and engineering teams. Collaborate with cross-functional teams to influence technical direction and ensure alignment with enterprise AI platform strategy. Support people management, leadership, and team development activities as required. Required Qualifications 10+ years of experience building and deploying production-grade machine learning systems at scale. Strong experience with LLMs, Generative AI, and RAG deployments in production environments. Strong Python development background. Expertise designing and implementing AI/ML systems in cloud environments such as AWS, Azure, or Google Cloud Platform. Hands-on experience with Kubernetes, containerization, and scalable inference systems. Experience designing agentic AI systems and tool orchestration frameworks. Experience implementing and governing MCP servers or structured architectures for tool integration and context management. Experience with large-scale distributed ML systems and enterprise platform engineering. Experience establishing MLOps, CI/CD, observability, and system reliability practices. Demonstrated people management, technical leadership, or mentorship experience. Strong understanding of high-availability and low-latency AI/ML architectures. Ability to define technical standards and influence architecture and engineering decisions across teams. Education: Bachelors Degree

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

6:10 min

Unlocking free learning credits via Google Cloud Innovators

Asrar Asrar · World Congress 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

Videos

See all

Related articles

See all