Senior AI/ML Engineer in San Francisco

Energy Jobline
San Francisco, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Distributed Computing Environment Python (Programming Language) Performance Tuning Recommender Systems Pytorch Large Language Models

Requirements

  • 4+ years of applied ML engineering in production environments
  • Hands-on experience with LLMs, fine-tuning, RAG, or large-scale recommender systems
  • Strong Python and PyTorch (or JAX) fundamentals
  • Experience with distributed training, GPU optimization, or inference serving
  • Pragmatic about trade-offs between research-grade and ship-grade work

About the company

Our client is a well-funded AI startup building production-grade ML infrastructure used by enterprise customers. They are looking for a Senior AI/ML Engineer to own model training pipelines, evaluation systems, and inference serving at scale. Full-time, on-site in San Francisco.

What you will do

  • Design and ship end-to-end ML systems: data pipelines, training, evaluation, deployment
  • Own model performance, latency, and cost trade-offs in production
  • Build evaluation harnesses and offline benchmarks for fast iteration
  • Work directly with product to translate ambiguous goals into measurable model improvements
  • Mentor other engineers on ML best practices and code quality

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.energyjobline.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:36 min

Exploring high-level Python frameworks for accelerated enterprise artificial intelligence

Paul Graham Paul Graham · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

8:32 min

Benchmarking GitOps engine constraints for extensive multi-cluster environments

Artem Lajko · Europe 2026 Virtual

1:50 min

Real life recommendation systems and final project conclusions

Lutske van der Meer Lutske van der Meer · World Congress 2024

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

5:01 min

Leveraging large language models for code optimization and development

Stephan Gillich Stephan Gillich +3 · World Congress 2024

Videos

See all

Related articles

See all