> Markdown version of [/jobs/ext/277243-ml-platform-engineer](https://www.wearedevelopers.com/jobs/ext/277243-ml-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # ML Platform Engineer - **Company:** OpenKyber LLC - **Location:** United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Databases, Information Engineering, Extract Transform Load (ETL), Distributed Systems, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Machine Learning, Performance Tuning, Azure Machine Learning, Search Technologies, Systems Integration, Workflow Management Systems, AWS Cdk, Cloud Platform System, Large Language Models, Informatica Cloud, Generative AI, Backend, Cloudformation, Information Technology, Low Latency, AWS Glue, AWS Data Analytics, Machine Learning Operations, Cloudwatch, Terraform, GPT, Data Pipelines - **Published:** May 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=838425ec33e5fff1 ## About the Role 12+ years of overall IT development experience, with a strong background in backend and distributed systems. 7+ years of experience in Machine Learning, Data Engineering, or Applied AI engineering. Strong proficiency in Python, with experience building modular, production-grade services. Proven experience implementing Retrieval-Augmented Generation (RAG) and semantic search architectures. Hands-on experience integrating and operationalizing OpenAI LLM APIs in production environments. Solid experience deploying and managing systems within AWS environments, including services such as S3, Lambda, ECS/EKS, and IAM. Experience building scalable, secure, and observable AI/ML systems in production., * Experience working with Amazon SageMaker and/or Amazon Bedrock for model development, deployment, or managed LLM services. * Strong familiarity with AWS data services, including AWS Glue, Amazon Athena, Amazon OpenSearch Service, and Amazon Aurora. * Hands-on experience designing and implementing ETL/ELT data pipelines in cloud environments. * Experience building LLM orchestration pipelines, including reasoning workflows, tool usage, and multi-step agent architectures. * Knowledge of LLM benchmarking, evaluation frameworks, and performance optimization (latency, cost, quality metrics). * Experience integrating enterprise systems using SnapLogic. * Exposure to Craxel Black Forest Time-Series Database (or similar time-series platforms); willingness to learn/train if not previously experienced. * Experience implementing Infrastructure as Code (IaC) using AWS CDK, CloudFormation, or Terraform. Key Skills: Machine Learning, LLM, AWS, Amazon SageMaker / Amazon Bedrock, RAG, Python ## Description We are seeking a ML Engineer LLM Platforms & Assistants who will design, build, and operate production-grade large language model (LLM) pipelines primarily within AWS-based environments. This role focuses on integrating OpenAI models into modular Python services, implementing Retrieval-Augmented Generation (RAG) and semantic search, and deploying scalable, secure, and observable AI assistants., * Design and maintain LLM integrations using OpenAI APIs within AWS environments. * Build Python-based LLM services deployed on AWS compute platforms (ECS, EKS, Lambda, or EC2). * Implement RAG workflows and semantic search using AWS data and storage services. * Develop LangChain or agentic workflows supporting reasoning and tool use. * Integrate LLM pipelines with ETL/ELT workflows and enterprise data systems. * Deploy and integrate MCP servers and emerging orchestration tools. * Apply AWS security best practices using IAM, KMS, and Secrets Manager. * Implement monitoring and observability using CloudWatch and related tools. * Migrate custom GPT solutions into production-grade AWS-hosted assistants. ## Related Videos - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Building Reliable Serverless Applications with AWS CDK and Testing](https://www.wearedevelopers.com/videos/812-building-reliable-serverless-applications-with-aws-cdk-and-testing) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [The power of Cloud Development Kit (CDK): How to get the most out of it](https://www.wearedevelopers.com/videos/740-the-power-of-cloud-development-kit-cdk-how-to-get-the-most-out-of-it) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)