Senior AI/ML Engineer

CLERA, LLC
San Francisco, CA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$180,000.0 - $280,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Software Quality Distributed Computing Environment Python (Programming Language) Performance Tuning Recommender Systems Pytorch Large Language Models Data Pipelines

Job description

  • Design and ship end-to-end ML systems: data pipelines, training, evaluation, deployment
  • Own model performance, latency, and cost trade-offs in production
  • Build evaluation harnesses and offline benchmarks for fast iteration
  • Work directly with product to translate ambiguous goals into measurable model improvements
  • Mentor other engineers on ML best practices and code quality

Requirements

Do you have experience in Production systems?, * 4+ years of applied ML engineering in production environments

  • Hands-on experience with LLMs, fine-tuning, RAG, or large-scale recommender systems
  • Strong Python and PyTorch (or JAX) fundamentals
  • Experience with distributed training, GPU optimization, or inference serving
  • Pragmatic about trade-offs between research-grade and ship-grade work

About the company

Our client is a well-funded AI startup building production-grade ML infrastructure used by enterprise customers. They are looking for a Senior AI/ML Engineer to own model training pipelines, evaluation systems, and inference serving at scale. Full-time, on-site in San Francisco.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

6:08 min

Applying software engineering environments and testing to data pipelines

Matthias Niehoff Matthias Niehoff · WWC 2024

1:48 min

Balancing code generation velocity with software quality standards

Lilia Gargouri Lilia Gargouri · Coffee With Developers

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

2:36 min

Exploring high-level Python frameworks for accelerated enterprise artificial intelligence

Paul Graham Paul Graham · LIVE

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

Videos

See all

Related articles

See all