Foundation Model Engineer

Bright Vision Technologies
United States
24 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$100,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

Computer Clusters Code Review Python (Programming Language) Machine Learning Language Modeling Performance Tuning Reinforcement Learning Pytorch Large Language Models Deep Learning Information Technology Data Lineage
+4 more
Optimization Algorithms Free and Open-Source Software GPT Data Generation

Job description

We are looking for an Foundation Model Engineer to design, execute, and operationalize fine-tuning workflows for large language models across supervised, preference-based, and reinforcement learning approaches. The role requires deep practical experience with modern training stacks, careful dataset construction, rigorous evaluation methodology, and the engineering discipline to operate complex training pipelines reliably. The ideal candidate combines strong ML intuition with production-grade engineering practices, and is comfortable navigating the trade-offs between data quality, compute budget, evaluation rigor, and shipping velocity. In this role you will work closely with cross-functional partners - product, design, engineering, operations, and business stakeholders - to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong, * Design and execute fine-tuning experiments for large language models using supervised, DPO, RLHF, and related techniques.

  • Lead dataset construction, curation, and quality assurance processes for instruction tuning and preference data.
  • Build scalable training pipelines on top of modern distributed training frameworks.
  • Tune hyperparameters, optimizer configurations, and training stability strategies for large-model fine-tuning.
  • Implement parameter-efficient fine-tuning techniques such as LoRA, QLoRA, and adapter-based methods.
  • Design rigorous evaluation suites including automated benchmarks, human evaluation, and capability-specific probes.
  • Implement safety, refusal, and policy evaluations to track model behavior across releases.
  • Operate large-scale training jobs on GPU clusters, diagnosing failures and recovering training state reliably.
  • Optimize training throughput using mixed precision, sequence packing, and efficient attention implementations.
  • Manage model artifacts, lineage tracking, and reproducibility across many concurrent experiments.
  • Collaborate with product, research, and platform teams to align fine-tuning roadmaps with business needs.
  • Document training methodology, results, and decisions clearly for technical and non-technical audiences.
  • Mentor engineers on fine-tuning best practices, evaluation rigor, and responsible deployment.
  • Stay current with LLM research and translate advances into production-ready fine-tuning recipes.

Requirements

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position., engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production., * Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent experience.

  • Six or more years of combined ML research and engineering experience, with significant LLM exposure.
  • Strong proficiency in Python and modern deep learning frameworks, especially PyTorch.
  • Hands-on experience fine-tuning transformer-based language models at non-trivial scale.
  • Familiarity with distributed training strategies including FSDP, ZeRO, and pipeline parallelism.
  • Experience with RLHF, DPO, or other preference optimization techniques.
  • Strong understanding of evaluation methodology, benchmarks, and human evaluation design.
  • Experience operating training jobs on GPU clusters and recovering from failures.
  • Strong written and verbal communication skills.
  • Track record of shipping or publishing impactful LLM work., * Publications at top-tier ML venues.
  • Experience with multimodal model fine-tuning.
  • Familiarity with synthetic data generation and dataset distillation.
  • Open-source contributions to LLM training libraries.
  • Exposure to responsible AI evaluation and red-teaming practices.

Benefits & conditions

4.24.2 out of 5 stars Remote $100,000 - $150,000 a year - Full-time

About the company

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

3:39 min

Addressing code review surrender and process exploitation

Laura Tacho Laura Tacho · WWC Europe 2026

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · WWC Europe 2026

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · WWC 2025

2:15 min

Open-source community and machine learning frameworks

Gian Marco Iodice Gian Marco Iodice · WWC 2025

Videos

See all

Related articles

See all