FS Ai Engineer/ AI/ML Engineer

Connvertex Technologies Inc.
Atlanta, GA, United States
11 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Continuous Integration Software Debugging Programming Tools Distributed Systems Python (Programming Language) Node.Js Open Source Technology Software Engineering Systems Integration Data Logging
+11 more
Google Cloud Retrieval-Augmented Generation Large Language Models Multi-Agent Systems Prompt Engineering Mttr Containerization Kubernetes Infrastructure Automation Frameworks Machine Learning Operations Docker

Job description

We are seeking an experienced AI/ML Engineer to design, build, and operate AI/ML infrastructure and agentic systems. This role involves developing MCP servers and agents, integrating LLMs, and implementing RAG pipelines for production environments. Key Responsibilities

  • Design, build and operate MCP servers and MCP agents that host, orchestrate and monitor AI/agent workloads
  • Develop agentic AI, prompt engineering patterns, LLM integrations and developer tooling for production use.
  • Own deployment, scaling, reliability and cost-efficiency on Kubernetes/Docker and Google Cloud with automated CI/CD
  • Design and implement RAG (Retrieval Augmented Generation) pipelines and integrations with vector stores and retrieval tooling; use LangChain and Langfuse for orchestration, chaining, and observability.

Core Responsibilities:

  • Implement and maintain MCP server and agent code, APIs, and SDKs for model access and agent orchestration.
  • Design agent behavior, workflows and safety guards for agentic AI systems.
  • Create, test and iterate prompt templates, evaluation harnesses and grounding/chain of thought strategies.
  • Integrate LLMs and model providers (self hosted and cloud APIs) with unified adapters and telemetry.
  • Build developer tooling: CLI, local runner, simulators, and debugging tools for agents and prompts.
  • Containerize services (Docker), manage orchestration (Kubernetes/GKE), and optimize nodes, autoscaling and resource requests.
  • Ensure observability: logging, metrics, traces, dashboards, alerting and SLOs for model infra and agents.
  • Create runbooks, playbooks and incident response procedures; reduce MTTR and perform postmortems.
  • Design and maintain RAG workflows: document chunking, embeddings, vector indexing, retrieval strategies, re ranking and context injection.
  • Integrate and instrument LangChain for composable chains, agents and tooling; use Langfuse (or equivalent tracing) to capture prompts, model calls, RAG traces and evaluation telemetry., Job Title: Production AI Platform Engineer (GCP) Location : Dallas, TX/ Basking Ridge, NJ/Atlanta, GA (Hybrid position, 2-3 days onsite) Long term position We are looking for c…
  • 15 hours ago +

Requirements

  • 5+ years of Strong Software Engineering (Python/NodeJS), system design and production service experience.
  • 2+ years of Experience with LLMs, prompt engineering, and agent frameworks.
  • 2+ years of Experience Practical experience implementing RAG: embeddings, vector DBs and retrieval tuning.
  • 2+ years of Experience with LangChain patterns and with toolchain telemetry (Langfuse or similar) for prompt/model traceability.
  • 5+ years of Experience with Kubernetes, Docker, CI/CD and infrastructure as code experience.
  • 2+ years of Experience with Practical experience with Google Cloud Platform services
  • 2+ years of Experience with Observability, testing, and security best practices for distributed systems.
  • 2+ years of Experience with evaluating and mitigating retrieval/augmentation failures, hallucinations, and leakage risks in RAG systems.
  • Familiarity with vendor and open source vector stores and embedding providers

About the company

Founded in 2012, H2O.ai is on a mission to democratize AI. As the world’s leading agentic AI company, H2O.ai converges Generative and Predictive AI to help enterprises and public s…

  • 9 days ago, Founded in 2012, H2O.ai is on a mission to democratize AI. As the world’s leading agentic AI company, H2O.ai converges Generative and Predictive AI to help enterprises and public s…
  • 1 month ago

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:45 min

Fusing developer experience and platform engineering for agentic SDLC

Julia Kordick Julia Kordick · WWC Europe 2026

45 sec

Working securely with Node.js path application programming interfaces

Sonya Moisset · WWC 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · WWC 2021

3:55 min

Identifying underlying Node.js runtime vulnerabilities using fuzzing tools

Sonya Moisset · WWC 2023

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all