AI Model Deployment Administrator
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+7 more
Job description
Design and deploy agentic AI workflows that connect Large Language Models (LLMs) and custom machine learning models into backend architectures, using frameworks like LangChain, LangGraph, PydanticAI or SemanticKernel. Design systems for multi-agent coordination and multi-step reasoning loops/memory management. Build and maintain data ingestion pipelines and vector databases (e.g., Pinecone, Weaviate) to support Retrieval-Augmented Generation for AI agents. Implement tracing (e.g., LangSmith, Langfuse) and build evaluation pipelines to measure accuracy, latency, and cost. Design testing strategies for non-deterministic AI outputs, including the implementation of guardrails and safety constraints. Deploy and orchestrate containerized services using Docker and Kubernetes on cloud platforms like AWS, Google Cloud Platform, or Azure (Azure preferred)
Preferred: Experience with Azure Foundry/ Azure OpenAI
Requirements
- University degree in Computer Science or a related discipline.
- 12+ years of professional experience in software engineering or backend systems.
- 2 3 years relevant experience in AI technologies building APIs, integrating LLMs, designing RAG pipelines, and deploying scalable microservices.
- Strong applied expertise with frameworks like LangChain, FastAPI, CrewAI, and cloud platforms.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How to Become an AI Engineer
What Are Large Language Models?
From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path
Navigating the AI Shift