Senior Engineering Manager, Model Infrastructure

Harvey, Inc.
San Francisco, CA, United States
19 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$272,000.0 - $355,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Airflow Microsoft Azure Cloud Computing Distributed Systems Failover Machine Learning Open Source Technology Software Engineering AI Infrastructure Large Language Models
+10 more
Apache Spark Model Validation AI Platforms Kubernetes Apache Flink Apache Kafka Data Management Machine Learning Operations Hardware Infrastructure Data Pipelines

Job description

  • Lead and grow a high-performing team of software engineers responsible for Harvey’s Model Infrastructure platform.
  • Define the technical roadmap for model reliability, scalability, and operational excellence.
  • Build highly reliable systems for model provisioning, capacity management, failover, and incident response across multiple AI providers.
  • Own Harvey’s multi-provider model platform, including provider integrations, SDK upgrades, API migrations, and onboarding new model providers.
  • Drive the evolution of our Unified Model Controller (UMC) and Model Selector platform to automatically detect degraded models and intelligently route traffic based on health, latency, quality, compliance, and cost.
  • Improve observability through health dashboards, alerting, token usage analytics, cost reporting, and end-to-end model telemetry.
  • Partner with Product Engineering to support new model launches, capacity planning, experimentation, and proactive production monitoring.
  • Lead initiatives to improve inference efficiency, reduce infrastructure costs, and increase model utilization across providers.
  • Build the infrastructure foundation for Harvey’s future model training efforts, including data pipelines, model operations, training environments, and AI platform capabilities.
  • Partner with executive leadership on long-term AI infrastructure strategy and vendor relationships.
  • Recruit, mentor, and develop exceptional engineering talent while fostering a culture of technical excellence and operational ownership.

Requirements

  • 8+ years of software engineering experience, including multiple years managing high-performing engineering teams.
  • Experience leading teams responsible for large-scale distributed systems or cloud infrastructure.
  • Strong technical background that enables you to guide architectural decisions and mentor senior engineers.
  • Experience operating highly available production services with strong reliability and operational excellence.
  • Experience building platforms that require scalability, observability, automation, and cost optimization.
  • Strong cross-functional leadership skills with the ability to partner effectively across Engineering, Research, Product, and external vendors.
  • Excellent communication skills and the ability to influence technical strategy across organizations.
  • A passion for building teams and developing engineering talent.

Nice to Have

  • Experience with AI infrastructure, LLM serving, or machine learning platforms.
  • Experience working with multiple model providers such as OpenAI, Anthropic, Azure OpenAI, Fireworks, Baseten, or open-source model ecosystems.
  • Experience building inference platforms, model gateways, traffic routing systems, or policy-based serving infrastructure.
  • Experience with Kubernetes, cloud infrastructure, distributed systems, and large-scale observability platforms.
  • Experience supporting GPU infrastructure, model training platforms, or ML infrastructure.
  • Familiarity with data platforms and technologies such as Spark, Kafka, Flink, Airflow, or Iceberg.
  • Experience leading organizations through periods of rapid growth and technical transformation.

Benefits & conditions

Pulled from the full job description

  • 401(k) 4% Match
  • Health insurance
  • 401(k) matching
  • Paid time off
  • Vision insurance
  • Dental insurance, * $272K - $355K * Offers Equity * Offers Bonus

Additionally, this role is eligible to participate in our equity plan and benefits program. Benefits include, but not limited to: Comprehensive health, dental and vision coverage, retirement benefits (401k match up to 4%), and flexible PTO., $272,000 - $355,000 USD

About the company

At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come.

This is a rare chance to help build a generational company at a true inflection point. With 1500+ customers in 60+ countries, strong product-market fit, and world-class investor support, we’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth - personal, professional, and financial - is unmatched.

Our team moves fast, takes ownership, and is deeply committed to the mission - operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job’s Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we’re never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we’d love to build with you.

At Harvey, the future of professional services is being written today - and we’re just getting started., As the Engineering Manager for Model Infrastructure, you’ll lead the team responsible for the platform powering every model request across Harvey. You’ll partner closely with AI Research, Product Engineering, Infrastructure, and external AI providers to ensure our platform remains reliable, scalable, and cost-efficient as our business grows.

Model Infrastructure is one of Harvey’s most strategic engineering organizations. Every product capability-from chat experiences and agents to document workflows and future reasoning systems-depends on this platform.

Over the next several years, the team will evolve beyond operating third-party models to building the infrastructure that enables Harvey to train, evaluate, deploy, and operate our own frontier AI models. This role offers the opportunity to shape the technical foundation of Harvey’s AI platform and build an organization that will power the company’s next phase of growth.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

2:59 min

Scaling clusters and handling automated replica failover

Jürgen Pilz · WWC 2023

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · WWC 2022

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

2:03 min

Solving complex engineering challenges in artificial intelligence deployment

Nico Axtmann · WWC 2022

Videos

See all

Related articles

See all