Staff AI/Machine Learning Engineer

Tonic AI, Inc.
United States
8 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Customer Data Management Distributed Computing Environment Information Extraction Machine Learning Operational Databases Pytorch Large Language Models Model Validation Build Management Machine Learning Operations Data Generation

Job description

The models you build here are load-bearing. The environments you generate decide whether an agent is ready to ship or only looked good in a demo. The synthesis and de-identification models you train decide whether a bank can safely put its data near a model at all. And the work spans real range: in one week you might build evaluation that separates the best models from the rest on real tasks, train a synthesis model where both fidelity and downstream utility have to hold, and improve entity detection on messy production data. Real enterprise data, real stakes, and problems that don’t have textbook answers yet.

What You’ll Do

  • Design and build the systems that generate longitudinally coherent synthetic environments for agent training and evaluation, including persona modeling, task generators, and verifiable ground truth.
  • Build and maintain synthesis models that generate realistic replacement values at very large scale, preserving format, statistical distribution, and semantic consistency so de-identified data stays useful downstream.
  • Train and improve the NER models behind our entity detection, driving accuracy and recall across free text, structured fields, and mixed enterprise data at scale.
  • Build evaluation infrastructure that grades agent outcomes, not just traces, and produces real discrimination between frontier models on real tasks.
  • Fine-tune and evaluate open-weight models on Tonic-generated data, and turn benchmark results into product and research direction.
  • Expand coverage into new domains, languages, and entity types, and handle the long tail of formats and edge cases that real customer data throws off.
  • Own model evaluation across the board: precision and recall on detection, utility preservation on synthesis, and outcome-level grading for agents.
  • Optimize inference so models run efficiently on large volumes of sensitive data inside customer environments.
  • Partner directly with frontier labs and enterprise ML team to turn hard data problems into shipped model improvements.
  • Set technical direction for a small, senior team and raise the bar on rigor, reproducibility, and shipping.

Requirements

  • 8+ years (or PhD with 3+ years) building production ML systems, with real depth in some combination of LLMs, agents, RL, NER, or information extraction.
  • Hands-on experience training and shipping models to production, and a pragmatic bar for quality: you know how to measure it, where it breaks, and when it’s good enough to ship.
  • Experience with generative or synthesis models where output fidelity and downstream utility both matter, not just plausibility.
  • Strong software engineering fundamentals. You write code others build on.
  • Fluency with modern training and eval stacks (PyTorch, distributed training, standard agent and benchmark frameworks).
  • Comfort working with messy, sensitive, real-world data and the privacy constraints that come with it.
  • A track record of framing ambiguous problems and driving them to measurable, shipped results.
  • Bonus: synthetic data generation, data privacy or de-identification, or benchmark construction.

Benefits & conditions

Pulled from the full job description

  • Parental leave
  • 401(k)
  • Health insurance
  • Vision insurance
  • Dental insurance
  • Unlimited paid time off, * Competitive salary and equity
  • Unlimited paid time off
  • 401k plan with employer contribution
  • Medical, dental, and vision insurance
  • Generous parental leave policy
  • Remote-friendly work environment

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

1:46 min

Overcoming data scarcity with synthetic data generation

Anshul Jindal Anshul Jindal +1 · WWC Europe 2026

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · WWC 2025

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

Videos

See all

Related articles

See all