Senior Machine Learning Engineer

Retell AI, Inc
Redwood City, CA, United States
6 days ago
Apply on www.careerboard.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$200,000.0 - $280,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Continuous Delivery Python (Programming Language) Machine Learning Pytorch Large Language Models Machine Learning Operations

Job description

This is a hands-on, high-ownership role for ML engineers who want to build production models that actually ship and perform under real-world constraints. As a Founding Senior Machine Learning Engineer at Retell, you’ll work across the ML stack to power human-like voice agents that handle millions of Real Time phone conversations. You’ll fine-tune large language models and audio models, evaluate them with rigorous benchmarks (and human feedback), and deploy them into latency-sensitive, high-traffic systems. You’ll own model performance end-to-end-from training pipelines to post-deployment monitoring-and shape our ML strategy alongside the founding team. If you’re excited by hard technical challenges, fast iteration, and the opportunity to define how voice AI works at scale, this role is a rare chance to do it from the ground up., * Train & Tune Models - Fine-tune LLMs and audio models to maximize speed, accuracy, and production-readiness-pushing the frontier of Real Time AI voice experiences.

  • Benchmark & Evaluate - Build datasets, define rigorous metrics, and measure model performance across high-impact voice AI tasks to guide development.
  • Deploy to Production - Work closely with engineering to ship models, monitor them in the wild, and ensure they stay fast, reliable, and accurate at scale.
  • Run Human Evaluations - Build scalable pipelines to collect structured human feedback, benchmark subjective quality, and inform model iterations.
  • Level Up Infrastructure - Design and maintain the ML infrastructure needed for fast experimentation, robust training, and continuous deployment.

Requirements

  • ML Engineer with Real-World Experience - You’ve trained and shipped models in production. Bonus if you’ve worked with LLMs or audio models.
  • Fluent in Modern ML Stack - You know your way around Python, PyTorch, and today’s ML tools-from training pipelines to evaluation benchmarks.
  • Execution-Oriented - You move fast, take ownership, and focus on solving real problems over perfect ones.
  • Startup-Ready - You’re adaptable, resilient, and energized by ambiguity and fast-changing priorities.
  • Clear Communicator & Team Player - You collaborate well across functions and push decisions forward.

Benefits & conditions

Other Benefits

  • 100% coverage for medical, dental, and vision insurance
  • $70/day DoorDash credit for unlimited breakfast, lunch, dinner, and snacks
  • $200/month wellness reimbursement (gym, fitness classes, etc.)
  • $300/month commuter reimbursement (gas, Caltrain, etc.)
  • $75/month phone bill reimbursement
  • $50/month Internet reimbursement

Compensation Philosophy

  • Best Offer Upfront: Choose from three cash-equity balance options, no negotiation needed.
  • Top 1% Talent: Above-market pay (top 5 percentile) to attract high performers.
  • High Ownership: Small teams, >$1M revenue/employee, and significant equity.
  • Performance-Based: Offers tied to interview performance, not experience or past salaries.

INTERVIEW PROCESS

  • Online Assessment (30 min): Coding questions on practical problem-solving (7 days to complete).
  • Talent Screen (15 min): Chat with our recruiter to get a better sense of the role, the team, and what it’s like to work here.
  • Technical Interview (45 min): Machine Learning specific coding interview.
  • Technical Interview (45 min): Live Practical Systems Design and Coding interview.
  • Onsite/Virtual Interviews (3 hrs): Hosted in our office if located in the Bay Area or virtual, with three rounds:
  • ML System Design: A non-coding interview focused on white-boarding and high-level system architecture.
  • ML Question Deep Dive: In-depth discussion exploring your approach to a machine learning problem.
  • Backend + AI Practical: A hands-on coding interview combining Back End development with AI integration.
  • Offer: Final stage, pending decision and offer discussion.

About the company

Retell AI is using first principles to reimagine the call center with cutting-edge voice AI. We believe voice is still the most natural way humans communicate, yet it has been trapped in outdated call centers for decades. Our mission is to bring intelligence, empathy, and speed to every phone conversation between businesses and their customers.

Since launching 18 months ago, thousands of companies now use Retell’s AI voice agents to handle sales, support, and logistics calls that once required large teams of human agents. Backed by Y Combinator, Altman Capital, and other leading investors, we have scaled to $30M ARR with a team of 20 people, up from $5M at the start of 2025.

Now, we’re scaling fast, and we’re looking for bold, ambitious people to help us build the gold standard for voice automation. If you want to work on deeply technical challenges, move fast, and make an outsized impact at one of the fastest-growing Voice AI startups in the world, you’ll love it here.

Let’s build the future of voice automation together.

  • We’re a top 50 AI app in a16z list: https://tinyurl.com/5853dt2x
  • We’re also one of the top ranking startups on: https://leanaileaderboard.com

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerboard.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Implementing continuous delivery architecture for machine learning

Dubravko Dolic +1 · World Congress 2021

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

3:31 min

Differentiating continuous delivery against continuous deployment models

Pawel Piwosz · LIVE

2:15 min

Open-source community and machine learning frameworks

Gian Marco Iodice Gian Marco Iodice · World Congress 2025

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all