> Markdown version of [/jobs/ext/3012283-enterprise-agent-development-platform-project](https://www.wearedevelopers.com/jobs/ext/3012283-enterprise-agent-development-platform-project). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Enterprise Agent Development Platform project - **Company:** EPAM Systems, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Continuous Integration, Python (Programming Language), Machine Learning, Large Language Models, Multi-Agent Systems, AI Platforms, Deployment Automation, Virtual Agents, Cloudwatch - **Published:** September 20, 2026 - **Apply:** https://www.dice.com/job-detail/d38f64d8-3ba2-4c0f-9e7b-276cd83aa5c3 ## About the Role behavioral checks Define evaluation criteria, metrics, thresholds, and acceptance rules Evaluate agent behavior across individual responses, tool calls, and complete workflows Work with OpenTelemetry traces and spans as evaluation data Integrate evaluations into CI/CD pipelines and automated deployment gates Enable continuous quality monitoring of solutions in production Establish reusable evaluation patterns and engineering standards Work closely with AI Engineers, Architects, and Platform Engineers to embed quality into the development process Requirements 5+ years of experience in ML Engineering, AI Engineering, or AI Platform Engineering Strong Python development experience Hands-on experience with LLM/GenAI evaluation Experience designing and implementing evaluation frameworks Experience developing custom or deterministic evaluators Experience integrating AI/ML quality checks into CI/CD Good understanding of LLM and AI agent architectures Nice to have Hands-on experience with AWS AgentCore Evaluation Experience with AWS Bedrock Guardrails, including PII detection Knowledge of CloudWatch metrics and production monitoring Experience with OpenTelemetry Familiarity with LangGraph, Strands Agents, or similar agent frameworks ## Related Videos - [Beyond the Benchmark: How to Evaluate AI Agents in the Real World](https://www.wearedevelopers.com/videos/100269-beyond-the-benchmark-how-to-evaluate-ai-agents-in-the-real-world) - [From Black Box to Glass Box : Bedrock AgentCore Observability](https://www.wearedevelopers.com/videos/2126-from-black-box-to-glass-box-bedrock-agentcore-observability) - [Guiding Agentic AI with Vue](https://www.wearedevelopers.com/videos/2033-guiding-agentic-ai-with-vue) - [This App Reached 10,000 Users in One Week. Here's How.](https://www.wearedevelopers.com/videos/100329-this-app-reached-10-000-users-in-one-week-here-s-how) - [On a Secret Mission: Developing AI Agents](https://www.wearedevelopers.com/videos/1510-on-a-secret-mission-developing-ai-agents) - [How to Stop Your Agents From Going Rogue - Arnav Gupta](https://www.wearedevelopers.com/videos/2152-how-to-stop-your-agents-from-going-rogue-arnav-gupta) ## Related Articles - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [What is Agentic Programming and Why Should Developers Care?](https://www.wearedevelopers.com/magazine/625-what-is-agentic-programming-and-why-should-developers-care) - [A 5-Step Open-Source Setup for Agentic Engineering](https://www.wearedevelopers.com/magazine/738-a-5-step-open-source-setup-for-agentic-engineering) - [Introducing Redis Agent Memory Server](https://www.wearedevelopers.com/magazine/699-introducing-redis-agent-memory-server) - [Never delegate the understanding](https://www.wearedevelopers.com/magazine/749-never-delegate-the-understanding) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j)