Software Engineer - AI Infrastructure

Andromeda Cluster, Inc.
San Francisco, CA, United States
7 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Linux Distributed Systems Python (Programming Language) Ansible Software Engineering Web Services AI Infrastructure Scripting Graphics Processing Unit (GPU) Computer Network Operations
+5 more
Kubernetes Bare Metal Restful APIs Terraform Golang

Job description

As an Infrastructure Product Engineer, you will play a pivotal role in building the backbone of Andromeda’s platform. You’ll transform complex, real-world infrastructure challenges into scalable product capabilities that benefit our customers.

Positioned at the intersection of infrastructure and product engineering, this role is deeply technical and systems-oriented, yet laser-focused on building solutions with broad leverage.

What You’ll Do

  • Design and develop core platform components, including infrastructure orchestration, provisioning, and lifecycle management solutions.
  • Build robust APIs, services, and control planes that abstract over diverse infrastructure types (VMs, Kubernetes, bare metal, schedulers).
  • Translate customer usage patterns into product requirements, delivering impactful features and improvements.
  • Create automation and internal tooling to eliminate manual or ad-hoc operational work.
  • Enhance reliability, performance, and observability at the platform level, emphasizing durable improvements over quick fixes.
  • Collaborate with peer teams to define clear ownership boundaries between platform capabilities and customer-specific solutions.
  • Write clean, maintainable, and well-documented code with a focus on long-term sustainability.
  • Participate in technical design discussions and contribute to the architectural evolution of our platform.

Requirements

  • 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.
  • Strong systems fundamentals: deep understanding of Linux, networking, storage, and distributed systems.
  • Proven expertise with Kubernetes, VMs, or bare-metal environments.
  • Advanced software engineering skills; capable of building production-grade APIs and services (Python, Go, or similar).
  • Extensive experience with infrastructure as code and automation tools (Terraform, Ansible, Helm, etc.).
  • Demonstrated ability to navigate ambiguity and distill complex problems into clear, maintainable abstractions.
  • Product-focused mindset: care about interfaces, defaults, reliability, and sustainable operations.
  • Excellent written and verbal communication skills; effective collaborator across engineering and product functions.

Nice to Have:

  • Hands-on experience with GPU or AI infrastructure.
  • Experience with control-plane or orchestration systems.
  • Background spanning both infrastructure and application/backend engineering.
  • Experience architecting multi-tenant systems.
  • Strong skills in technical writing and design documentation.
  • Early-stage startup experience., Advertising Operations, Ansible, Application Programming Interface (API), Architectural Services, Artificial Intelligence (AI), Automation, Communication Skills, Distributed Computing, Documentation, Finance, GPU (Graphics Processing Unit), Go Programming Language (Golang), Linux Operating System, Machine Tool, Network Operations Center, Politics, Presentation/Verbal Skills, Product Engineering, Productivity Model, Python Programming/Scripting Language, Research Laboratory, Risk, Sales, Software Engineering, Technical Writing, Technical/Engineering Design, Underwriting, Writing Skills

About the company

Andromeda is a market and infrastructure platform to buy, sell, and operate compute.

We believe demand for compute will grow exponentially. So fast that a handful of vertically integrated providers won’t be able to scale across operations, capital, supply chains, and politics to serve it. The result is a massive wave of fragmentation, with AI factories of every shape and size coming to market to fill this demand. Our job is to enable all of that fragmented compute to flow through one platform, delivering reliable capacity to model builders, research labs, and inference providers when they need it. We believe every spare electron should be made productive for AI and we’re building the platform that makes that possible.

We sit at the center of three forces:

  • Companies that need reliable, high-performance compute fast
  • A fragmented global supply of GPUs across hyperscalers, neoclouds, and independent data centers
  • Capital, risk, and operational complexity that most teams are not equipped to manage

When we succeed, trillions of dollars of compute will flow through Andromeda. Builders get capacity when they need it. Providers get a reliable way to monetize, operate, and finance infrastructure at scale. Capital gets an easy way to deploy, hedge, and underwrite.

In five years, Andromeda won’t just participate in the AI infrastructure market. We will shape it.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all