Foundation Model Engineer

Bright Vision Technologies
United States
13 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$200,000.0 - $230,000.0
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Automation of Tests Microsoft Azure Code Review Computer Programming Continuous Integration Data Cleansing Software Debugging Python (Programming Language) Machine Learning
+18 more
Open Source Technology Performance Tuning Reinforcement Learning Google Cloud Feature Engineering Chatbots Pytorch Large Language Models Prompt Engineering Generative AI Git Kubernetes Information Technology ONNX (Open Neural Network Exchange) Format HuggingFace Machine Learning Operations TensorRT Docker

Job description

We are looking for an Foundation Model Engineer to design, execute, and operationalize fine-tuning workflows for large language models across supervised, preference-based, and reinforcement learning approaches. The role requires deep practical experience with modern training stacks, careful dataset construction, rigorous evaluation methodology, and the engineering discipline to operate complex training pipelines reliably. The ideal candidate combines strong ML intuition with production-grade engineering practices, and is comfortable navigating the trade-offs between data quality, compute budget, evaluation rigor, and shipping velocity. In this role you will work closely with cross-functional partners - product, design, engineering, operations, and business stakeholders - to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong, * Develop, fine-tune, and optimize Large Language Models (LLMs) for enterprise AI applications.

  • Design and implement end-to-end model training and fine-tuning pipelines using PyTorch, Hugging Face Transformers, and related frameworks.
  • Prepare, clean, and curate training datasets for supervised fine-tuning and instruction tuning.
  • Implement parameter-efficient fine-tuning techniques such as LoRA, QLoRA, and PEFT to improve training efficiency.
  • Evaluate model performance using standard NLP benchmarks, automated metrics, and task-specific evaluations.
  • Collaborate with data scientists, software engineers, and product teams to integrate LLM solutions into production applications.
  • Optimize model inference performance, latency, and resource utilization for scalable deployment.
  • Develop data preprocessing, feature engineering, and automation scripts using Python.
  • Deploy and monitor machine learning models on cloud platforms using MLOps best practices.
  • Troubleshoot model training issues, optimize hyperparameters, and improve overall model accuracy.
  • Document technical designs, experiments, and implementation details for knowledge sharing.
  • Stay current with advancements in Generative AI, transformer architectures, and open-source LLM technologies.

Requirements

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position., engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production., * Master’s degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, or a related field. Equivalent industry experience will also be considered.

  • 6+ years of experience in Machine Learning or AI engineering, including at least 2 years working with Large Language Models (LLMs) or Generative AI.
  • Strong programming skills in Python.
  • Hands-on experience with PyTorch and Hugging Face Transformers.
  • Experience fine-tuning open-source LLMs such as Llama, Mistral, Falcon, Gemma, or similar transformer-based models.
  • Knowledge of parameter-efficient fine-tuning techniques including LoRA, QLoRA, or PEFT.
  • Experience building data preprocessing and model training pipelines.
  • Familiarity with vector databases, embeddings, and Retrieval-Augmented Generation (RAG) concepts.
  • Experience with cloud platforms such as AWS, Azure, or Google Cloud Platform.
  • Working knowledge of Docker, Kubernetes, Git, and CI/CD practices.
  • Strong analytical, problem-solving, and debugging skills.
  • Excellent communication and collaboration skills., * Experience with LangChain, LlamaIndex, DSPy, or similar LLM orchestration frameworks.
  • Familiarity with distributed model training technologies such as DeepSpeed or FSDP.
  • Experience deploying LLMs using vLLM, TensorRT-LLM, or ONNX Runtime.
  • Knowledge of ML lifecycle and MLOps tools such as MLflow, Kubeflow, or Weights & Biases.
  • Experience building enterprise AI chatbots, copilots, document intelligence, or RAG-based applications.
  • Exposure to prompt engineering, AI evaluation frameworks, and LLM safety best practices.
  • Contributions to open-source AI or machine learning projects are a plus.
  • Experience working in Agile software development environments.

Benefits & conditions

4.24.2 out of 5 stars Remote $200,000 - $230,000 a year - Full-time

About the company

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

2:08 min

Applying large language models to infrastructure tasks

Alfonso Sandoval Rosas Alfonso Sandoval Rosas · Europe 2026 Virtual

Videos

See all

Related articles

See all