Software Engineer - AI Infrastructure
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+5 more
Job description
As an Infrastructure Product Engineer, you will play a pivotal role in building the backbone of Andromeda’s platform. You’ll transform complex, real-world infrastructure challenges into scalable product capabilities that benefit our customers.
Positioned at the intersection of infrastructure and product engineering, this role is deeply technical and systems-oriented, yet laser-focused on building solutions with broad leverage.
What You’ll Do
- Design and develop core platform components, including infrastructure orchestration, provisioning, and lifecycle management solutions.
- Build robust APIs, services, and control planes that abstract over diverse infrastructure types (VMs, Kubernetes, bare metal, schedulers).
- Translate customer usage patterns into product requirements, delivering impactful features and improvements.
- Create automation and internal tooling to eliminate manual or ad-hoc operational work.
- Enhance reliability, performance, and observability at the platform level, emphasizing durable improvements over quick fixes.
- Collaborate with peer teams to define clear ownership boundaries between platform capabilities and customer-specific solutions.
- Write clean, maintainable, and well-documented code with a focus on long-term sustainability.
- Participate in technical design discussions and contribute to the architectural evolution of our platform.
Requirements
- 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.
- Strong systems fundamentals: deep understanding of Linux, networking, storage, and distributed systems.
- Proven expertise with Kubernetes, VMs, or bare-metal environments.
- Advanced software engineering skills; capable of building production-grade APIs and services (Python, Go, or similar).
- Extensive experience with infrastructure as code and automation tools (Terraform, Ansible, Helm, etc.).
- Demonstrated ability to navigate ambiguity and distill complex problems into clear, maintainable abstractions.
- Product-focused mindset: care about interfaces, defaults, reliability, and sustainable operations.
- Excellent written and verbal communication skills; effective collaborator across engineering and product functions.
Nice to Have:
- Hands-on experience with GPU or AI infrastructure.
- Experience with control-plane or orchestration systems.
- Background spanning both infrastructure and application/backend engineering.
- Experience architecting multi-tenant systems.
- Strong skills in technical writing and design documentation.
- Early-stage startup experience., Advertising Operations, Ansible, Application Programming Interface (API), Architectural Services, Artificial Intelligence (AI), Automation, Communication Skills, Distributed Computing, Documentation, Finance, GPU (Graphics Processing Unit), Go Programming Language (Golang), Linux Operating System, Machine Tool, Network Operations Center, Politics, Presentation/Verbal Skills, Product Engineering, Productivity Model, Python Programming/Scripting Language, Research Laboratory, Risk, Sales, Software Engineering, Technical Writing, Technical/Engineering Design, Underwriting, Writing Skills
About the company
Andromeda is a market and infrastructure platform to buy, sell, and operate compute.
We believe demand for compute will grow exponentially. So fast that a handful of vertically integrated providers won’t be able to scale across operations, capital, supply chains, and politics to serve it. The result is a massive wave of fragmentation, with AI factories of every shape and size coming to market to fill this demand. Our job is to enable all of that fragmented compute to flow through one platform, delivering reliable capacity to model builders, research labs, and inference providers when they need it. We believe every spare electron should be made productive for AI and we’re building the platform that makes that possible.
We sit at the center of three forces:
- Companies that need reliable, high-performance compute fast
- A fragmented global supply of GPUs across hyperscalers, neoclouds, and independent data centers
- Capital, risk, and operational complexity that most teams are not equipped to manage
When we succeed, trillions of dollars of compute will flow through Andromeda. Builders get capacity when they need it. Providers get a reliable way to monetize, operate, and finance infrastructure at scale. Capital gets an easy way to deploy, hedge, and underwrite.
In five years, Andromeda won’t just participate in the AI infrastructure market. We will shape it.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Navigating the AI Shift
Highest Paying Tech Companies for Developers
MLOps And AI Driven Development
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence