Infrastructure Engineer

David Joseph & Company
San Francisco, CA, United States
24 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$150,000.0 - $250,000.0
Working hours
Shift work
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Continuous Integration Graphics Processing Unit (GPU) Google Cloud Cloud Platform System Multi-Cloud Hardware Infrastructure Terraform Docker

Job description

Own the full multi-region GPU infrastructure stack end to end as the sole infra hire - global low-latency serving, multi-cloud and on-premise deployments, reliability, and cost efficiency. Your work sits directly in the critical path of live customer traffic. This is a high-ownership, in-person role at a fast-moving early-stage team, working a 996 pace (9am-9pm, six days a week) in San Francisco.

What you’ll be doing

  • Own the full infrastructure stack end to end across multi-region GPU deployments, major public clouds, and on-premise enterprise environments.
  • Build and maintain super low-latency GPU serving infrastructure that sits in the critical path of live customer traffic.
  • Manage multi-cloud deployments, including cloud marketplace integrations and provider relationships.
  • Design and iterate on deployment, scaling, reliability, and cost-efficiency systems as the sole infra owner.
  • Support on-premise deployments for enterprise clients and ensure performance and reliability at each site.
  • Research and adopt new infrastructure solutions continuously as the stack and customer base grow.

Tech stack: AWS, GCP, Terraform, Docker, CI/CD, GPU/ML inference infrastructure

Requirements

  • Own cloud systems serving compression API end-to-end
  • Build and operate global low-latency high-throughput GPU ML inference infrastructure
  • Work with AWS, Terraform, Docker and CI/CD
  • Have built and operated production infrastructure at a startup or larger company
  • Learn new solutions and technologies quickly
  • Improve and research infrastructure solutions continuously
  • Based in or willing to relocate to San Francisco to work in person at the hacker house
  • Willingness to work startup hours in a 996-style environment (9am-9pm, six days a week)

Green Flags

  • Quick learner who grasps products and systems fast
  • Experience building for performance and reliability at scale
  • Research and product focus mindset
  • High ownership mentality
  • Startup-minded operator who prioritizes learning and growth over work-life balance
  • GPU infrastructure experience in production
  • First infra hire at a startup
  • Background at an infrastructure company, * Infra scope limited to model training pipelines only
  • 20+ years of experience with a slow-moving, process-heavy background
  • Prioritizes work-life balance as a primary requirement
  • No production infra ownership

Benefits & conditions

Pulled from the full job description

  • Paid housing
  • Food provided
  • Dental insurance
  • Visa sponsorship, * Sole infra owner with full-stack ownership from day one, directly in the critical path of live customer traffic.
  • Well-funded seed-stage company with strong early traction and an experienced backer base.
  • Significant equity, housing and food provided at the SF hacker house, visa sponsorship, laundry and cleaning, company off-sites, infinite DoorDash, and health & dental.

About the company

A seed-stage LLM interpretability and context-optimization company building custom machine learning models that analyze and compress token contexts before they reach the underlying model - delivering roughly 50% inference cost reduction, lower latency, and higher accuracy for the enterprises and scale-ups integrating LLMs into their products. Venture-backed, with strong early traction (~1,000 customers within its first seven months).

Founded 2025 · 1-10 people · Industry: AI Tools / LLM infrastructure

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:22 min

Overcoming developer challenges in multi-cloud environments

Sandeep Pal Sandeep Pal · Coffee With Developers

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

Videos

See all

Related articles

See all