Senior Infrastructure Engineer

VAST, INC.
San Francisco, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$180,000.0 - $300,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Big Data C++ (Programming Language) Computer Programming Linux Distributed Systems Fault Tolerance Python (Programming Language) Kernel-Based Virtual Machine PostgreSQL Redis
+11 more
Zero Trust Network Access Virtualization Technology Cloud Platform System Backend Usage Tracking Kubernetes Hardware Infrastructure Restful APIs Terraform Grpc Docker

Job description

As a Senior Infrastructure Engineer, you will help design and scale the core systems that power Vast.ai’s global GPU marketplace. You’ll work closely with our founders and core engineering team to extend the underlying compute infrastructure - from GPU provisioning and scheduling to billing, orchestration, and marketplace dynamics. We’re looking for someone who has previously built large-scale infrastructure platforms - systems with similarities to Vast.ai, or distributed compute orchestration frameworks. Tech stack: Python, C++, PostgreSQL, Linux, Docker, KVM, Redis, Terraform, AWS, REST/gRPC APIs., * Improve the backend systems that power Vast.ai’s compute marketplace

  • Integrate GPU provider onboarding, usage tracking, billing, and orchestration APIs
  • Develop scalable infrastructure for workload scheduling and resource management
  • Optimize pricing and marketplace logic for efficiency and transparency
  • Benchmark, profile, and harden systems for performance, reliability, and fault tolerance
  • Collaborate with product and infrastructure teams to shape the future of decentralized compute

Requirements

  • Distributed Systems: Experience building high-throughput backend systems or compute clouds
  • Compute Orchestration: Familiarity with Docker, or custom scheduling frameworks
  • GPU Infrastructure: Understanding of GPU provisioning, driver management, and workload scheduling
  • Billing & Metering: Implemented or integrated usage-based billing and account credit systems
  • Marketplace Dynamics: Knowledge of dynamic pricing, spot instances, or supply-demand balancing mechanisms
  • Security & Multi-Tenancy: Experience designing secure, multi-tenant systems in cloud environments
  • Programming: Strong programming skills in Python and C++; ability to write performant, maintainable, well-architected code
  • Database Expertise: Comfortable designing schemas and queries for large-scale data systems (PostgreSQL preferred)

Nice to Have

  • Experience with GPU security, virtualization, or zero-trust compute isolation
  • Prior startup experience or end-to-end product ownership

Benefits & conditions

  • Comprehensive health, dental, vision, and life insurance
  • 401(k) with company match
  • Meaningful early-stage equity
  • Onsite meals, snacks, and close collaboration with founders and tech leads
  • Ambitious, fast-paced startup culture where initiative is rewarded

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:00 min

Overview of gRPC and language-agnostic environments

Max Hausner Max Hausner +1 · WWC 2025

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

1:11 min

Evaluating architectural trade-offs between REST and gRPC

Sakshi Nasha Sakshi Nasha · Europe 2026 Virtual

Videos

See all

Related articles

See all