AI Model Deployment Administrator

OpenKyber LLC
Piscataway, United States
3 months ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Memory Management Machine Learning Software Engineering Systems Integration Google Cloud Data Ingestion Retrieval-Augmented Generation Large Language Models Multi-Agent Systems
+7 more
Backend Fastapi Information Technology Machine Learning Operations Api Design Docker Microservices

Job description

Design and deploy agentic AI workflows that connect Large Language Models (LLMs) and custom machine learning models into backend architectures, using frameworks like LangChain, LangGraph, PydanticAI or SemanticKernel. Design systems for multi-agent coordination and multi-step reasoning loops/memory management. Build and maintain data ingestion pipelines and vector databases (e.g., Pinecone, Weaviate) to support Retrieval-Augmented Generation for AI agents. Implement tracing (e.g., LangSmith, Langfuse) and build evaluation pipelines to measure accuracy, latency, and cost. Design testing strategies for non-deterministic AI outputs, including the implementation of guardrails and safety constraints. Deploy and orchestrate containerized services using Docker and Kubernetes on cloud platforms like AWS, Google Cloud Platform, or Azure (Azure preferred)

Preferred: Experience with Azure Foundry/ Azure OpenAI

Requirements

  • University degree in Computer Science or a related discipline.
  • 12+ years of professional experience in software engineering or backend systems.
  • 2 3 years relevant experience in AI technologies building APIs, integrating LLMs, designing RAG pipelines, and deploying scalable microservices.
  • Strong applied expertise with frameworks like LangChain, FastAPI, CrewAI, and cloud platforms.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

2:03 min

Solving complex engineering challenges in artificial intelligence deployment

Nico Axtmann · WWC 2022

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

3:14 min

Building a community-governed LAMP stack for open AI

Raffi Krikorian Raffi Krikorian · WWC Europe 2026

Videos

See all

Related articles

See all