AI Foundation Model Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+15 more
Job description
We are seeking a Senior AI Foundation Model Engineer to build and deploy secure, scalable, enterprise-grade AI solutions using LLMs, RAG and agentic workflows. The role involves developing production AI applications and reusable services for an AWS-hosted, cloud-agnostic AI platform., · Build LLM applications, RAG pipelines, knowledge assistants, document intelligence solutions and workflow agents.
· Develop embeddings, semantic search, reranking, grounding and citation capabilities.
· Deploy and manage AI services using APIs, Docker, Kubernetes, CI/CD and cloud-native infrastructure.
· Collaborate on Terraform/IaC, environment promotion, release controls and rollback procedures.
· Optimize models and inference for accuracy, latency, throughput, token usage, reliability and cost.
· Implement LLMOps/MLOps covering evaluation, monitoring, observability, feedback loops and continuous improvement.
· Ensure security, privacy, Responsible AI, governance and audit readiness.
· Maintain production documentation, runbooks and release records.
Requirements
· Strong hands-on experience with LLMs, transformers, GenAI, RAG, embeddings and vector databases.
· Advanced Python skills and experience with PyTorch, TensorFlow, Hugging Face, LangChain, LlamaIndex, Semantic Kernel or similar frameworks.
· Production deployment experience using APIs, containers, Kubernetes, CI/CD and monitoring tools.
· Practical AWS AI/cloud experience, preferably with Bedrock, SageMaker, OpenSearch, Lambda and EKS/ECS.
· Working knowledge of Terraform/IaC, MLOps/LLMOps, model evaluation, inference optimization and secure data handling.
Preferred Experience
· Banking, risk, compliance, financial crime or enterprise technology experience.
· Experience with Kendra, Azure OpenAI, Vertex AI, Databricks, vLLM, Triton, MLflow, Kubeflow or model gateways.
· Knowledge of LoRA, PEFT, instruction tuning, quantization, model governance and private/open-source LLM deployments.
Alternate Titles: LLM Engineer, GenAI Engineer, AI Platform Engineer, RAG Engineer, Applied ML Engineer or NLP Engineer.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
What Are Large Language Models?
MLOps And AI Driven Development
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
MLOps – What’s the deal behind it?