AI Engineer job in San Francisco

Little Maintenance Co Inc
San Francisco, CA, United States
11 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Python (Programming Language) PostgreSQL Open Source Technology Redis Session Management Large Language Models Backend Fastapi Api Design

Job description

As the Backend LLM Engineer, you’ll be responsible for ensuring LiteLLM unifies the format for calling LLM APIs in the broader OpenAI + Anthropic spec. This involves writing transformations to convert API requests from OpenAI/Anthropic spec to various LLM provider formats, building provider-agnostic unification functionality (e.g. session management across non-openai models for /v1/responses API, etc.). You’ll work directly with the CEO and CTO on critical projects including:

  • System design for supporting provider level features at the gateway level (skills, compaction, batches, etc.)
  • Thinking about the developer experience - to support millions of users who leverage your work via LiteLLM’s Python SDK
  • Building across LLM’s, MCP’s, and Agents by maintaining an excellent LLM interoperability layer, supporting a wide range of MCP Auth flows and enabling various Agent use-cases

What is our tech stack

The tech stack includes Python, FastAPI, Redis, Postgres.

Requirements

  • 1-2 years of backend/full-stack experience with production systems
  • Passion for open source and user engagement
  • Experience working with the OpenAI api (understand the difference between /chat/completions and /responses, and can speak to API-specific nuances)
  • Strong work ethic and ability to thrive in small teams
  • Eagerness to talk to users and help solve real problems

Benefits & conditions

Why Join LiteLLM?

  • Work directly with the founders every day.
  • Build products used by thousands of engineering teams around the world.
  • Own meaningful product decisions-not just implementation.
  • Move incredibly fast and see your work in production immediately.
  • Help define the future of AI infrastructure and developer tooling.
  • Competitive salary, equity, health, dental, and vision benefits.

About LiteLLM

LiteLLM is the world’s most widely adopted AI Gateway. We provide a unified API for 100+ LLM providers along with routing, authentication, budgets, observability, and governance for production AI systems. Thousands of companies rely on LiteLLM to power AI in production.

If you love talking to users, building products, and shipping software that people actually use, we’d love to meet you.

About the company

LiteLLM is the world’s most popular AI Gateway, trusted by companies like Adobe, Netflix, and NASA. We’re building the next generation of infrastructure for AI applications-and now we’re building our second product.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.diversity.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

1:24 min

Building client-facing AI agents for engineering teams

Alfonso Graziano Alfonso Graziano · Coffee With Developers

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · World Congress 2022

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all