> Markdown version of [/jobs/ext/1486487-foundation-model-engineer](https://www.wearedevelopers.com/jobs/ext/1486487-foundation-model-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Foundation Model Engineer - **Company:** Bright Vision Technologies - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $200,000.0 - $230,000.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Amazon Web Services, Automation of Tests, Microsoft Azure, Code Review, Computer Programming, Continuous Integration, Data Cleansing, Software Debugging, Python (Programming Language), Machine Learning, Open Source Technology, Performance Tuning, Reinforcement Learning, Google Cloud, Feature Engineering, Chatbots, Pytorch, Large Language Models, Prompt Engineering, Generative AI, Git, Kubernetes, Information Technology, ONNX (Open Neural Network Exchange) Format, HuggingFace, Machine Learning Operations, TensorRT, Docker - **Published:** July 29, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=b2b7bed1ab7fd4d2 ## About the Role Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position., engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production., * Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, or a related field. Equivalent industry experience will also be considered. * 6+ years of experience in Machine Learning or AI engineering, including at least 2 years working with Large Language Models (LLMs) or Generative AI. * Strong programming skills in Python. * Hands-on experience with PyTorch and Hugging Face Transformers. * Experience fine-tuning open-source LLMs such as Llama, Mistral, Falcon, Gemma, or similar transformer-based models. * Knowledge of parameter-efficient fine-tuning techniques including LoRA, QLoRA, or PEFT. * Experience building data preprocessing and model training pipelines. * Familiarity with vector databases, embeddings, and Retrieval-Augmented Generation (RAG) concepts. * Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform. * Working knowledge of Docker, Kubernetes, Git, and CI/CD practices. * Strong analytical, problem-solving, and debugging skills. * Excellent communication and collaboration skills., * Experience with LangChain, LlamaIndex, DSPy, or similar LLM orchestration frameworks. * Familiarity with distributed model training technologies such as DeepSpeed or FSDP. * Experience deploying LLMs using vLLM, TensorRT-LLM, or ONNX Runtime. * Knowledge of ML lifecycle and MLOps tools such as MLflow, Kubeflow, or Weights & Biases. * Experience building enterprise AI chatbots, copilots, document intelligence, or RAG-based applications. * Exposure to prompt engineering, AI evaluation frameworks, and LLM safety best practices. * Contributions to open-source AI or machine learning projects are a plus. * Experience working in Agile software development environments. ## Description We are looking for an Foundation Model Engineer to design, execute, and operationalize fine-tuning workflows for large language models across supervised, preference-based, and reinforcement learning approaches. The role requires deep practical experience with modern training stacks, careful dataset construction, rigorous evaluation methodology, and the engineering discipline to operate complex training pipelines reliably. The ideal candidate combines strong ML intuition with production-grade engineering practices, and is comfortable navigating the trade-offs between data quality, compute budget, evaluation rigor, and shipping velocity. In this role you will work closely with cross-functional partners - product, design, engineering, operations, and business stakeholders - to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong, * Develop, fine-tune, and optimize Large Language Models (LLMs) for enterprise AI applications. * Design and implement end-to-end model training and fine-tuning pipelines using PyTorch, Hugging Face Transformers, and related frameworks. * Prepare, clean, and curate training datasets for supervised fine-tuning and instruction tuning. * Implement parameter-efficient fine-tuning techniques such as LoRA, QLoRA, and PEFT to improve training efficiency. * Evaluate model performance using standard NLP benchmarks, automated metrics, and task-specific evaluations. * Collaborate with data scientists, software engineers, and product teams to integrate LLM solutions into production applications. * Optimize model inference performance, latency, and resource utilization for scalable deployment. * Develop data preprocessing, feature engineering, and automation scripts using Python. * Deploy and monitor machine learning models on cloud platforms using MLOps best practices. * Troubleshoot model training issues, optimize hyperparameters, and improve overall model accuracy. * Document technical designs, experiments, and implementation details for knowledge sharing. * Stay current with advancements in Generative AI, transformer architectures, and open-source LLM technologies. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [The Best Large Language Models on The Market](https://www.wearedevelopers.com/magazine/319-the-best-large-language-models-on-the-market) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)