Platform Engineer (10x Openings)

Volta
Greater London, UK
7 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Clean Code Principles Java (Programming Language) Application Programming Interfaces (APIs) Agile Methodology Border Gateway Protocol C Sharp (Programming Language) C++ (Programming Language) Software as a Service Code Review Continuous Integration Software Debugging Linux
+27 more
Distributed Data Store Distributed Systems Infrastructure as a Service (IaaS) Python (Programming Language) Network Control Networking Basics Object-Oriented Software Development Platform as a Service (PAAS) Performance Tuning Scrum Methodology Remote Direct Memory Access Prometheus Software Engineering System Programming Systems Integration Management of Software Versions Ceph (Software) Datadog Data Logging High Performance Computing Grafana Backend Kubernetes Bare Metal Hardware Infrastructure Software Version Control Serverless Computing

Job description

  • Design and implement Kubernetes operators and controllers that manage the lifecycle of platform resources.
  • Work closely with the product team to turn roadmap requirements into the platform capabilities that support them.
  • Collaborate with the bring-up teams to identify operational pain points and turn them into scalable platform features.
  • Integrate security guidance from the security engineering team into platform-level controls, and remediate findings at the platform layer.
  • Treat observability as a platform concern: instrument services, define meaningful metrics, and build tooling that gives the team visibility into platform health.
  • Own the services you build in production, including participation in an on-call rotation, incident response, and the follow-up work that closes structural gaps rather than only the immediate issue.
  • Hold to clean interface and versioning practice on anything other teams or customers depend on, including disciplined handling of breaking changes.
  • Participate in code review, technical design discussions, and cross-team collaboration in an Agile (Kanban or Scrum) environment.

Depending on your team and background, you will go deep in some of the following:

  • Control plane and APIs: improve and extend the API layer between user-facing services and the underlying platform, with disciplined versioning and backward compatibility.
  • Compute: build and operate the lifecycle of virtualized and bare metal compute resources.
  • Storage: provisioning workflows, attachment reliability, performance tuning, and failure handling.
  • Confidential computing: build and extend confidential computing capabilities across the stack, from secure bare metal and confidential VMs to Confidential Containers (CoCo).
  • Customer-facing services: the platform surfaces customers interact with directly, including the APIs and interfaces through which they consume capacity, working alongside product and UX.

Requirements

  • 3 to 5 years of software engineering experience, with a meaningful portion spent on infrastructure or platform systems.
  • Strong backend or systems programming experience in a production environment. Our working languages are Python, Go, and Rust; we welcome strong engineers from other compiled or object-oriented languages (for example C++, C#, or Java) who are ready to work across our stack as it evolves.
  • Solid understanding of Kubernetes internals: the control loop model, CRDs, controllers and operators, and reliable reconciliation logic.
  • Comfortable working close to the infrastructure layer: Linux, networking fundamentals, and distributed systems behavior.
  • Experience designing, building, and versioning production-grade APIs or service interfaces that other teams depend on, including disciplined handling of breaking changes and backward compatibility.
  • Experience operating what you build: debugging production systems, and taking part in on-call or incident response.
  • Strong engineering fundamentals: clean code, testing, version control, code review, and CI/CD practices., * Fluency with AI-assisted development: agentic CLI tools, IDE assistants, and orchestrating multiple coding agents through MCP, skills, or APIs to amplify delivery.
  • Depth in Go or Rust beyond working proficiency.
  • Familiarity with confidential computing technologies: TEEs, AMD SEV, Intel TDX, or Confidential Containers (CoCo).
  • Experience integrating security requirements into platform or infrastructure systems.
  • Familiarity with high-performance networking: overlay protocols, BGP, RDMA, or packet-processing frameworks.
  • Hands-on experience with distributed storage systems (Ceph or similar) at an engineering level.
  • Background building Kubernetes operators using frameworks such as Kopf, controller-runtime, or similar.
  • Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems.
  • Experience building SaaS or PaaS layers on top of an IaaS platform.
  • Exposure to serverless or inference serving infrastructure.
  • Exposure to GPU infrastructure or HPC environments.
  • Experience working distributed across time zones with counterparts in other regions.
  • REQ-79

Benefits & conditions

At Volta, we believe people do their best work when they feel supported, trusted and able to grow. We’re building a company where you can make an impact, keep a healthy balance between work and life, and build a career you’re proud of. As a global team, we do our best to provide great benefits wherever you’re based. While some benefits vary by country due to local regulations, we believe looking after our people is simply the right thing to do.

  • Competitive salary based on the work you do here, not your previous salary
  • Equity in Volta, giving you the opportunity to share in the company’s long-term success
  • Retirement/pension contributions
  • Comprehensive health, wellbeing and insurance benefits
  • Generous number of vacation days each year

About the company

Volta is the category-defining, fully vertically integrated AI infrastructure platform - from capital to clusters to software, under a founder-led enterprise. Our mission is The Utility of Compute : AI infrastructure as dependable and available as electricity, for every organization that needs it. Launched with a $10B strategic partnership with one of the leading frontier AI labs, a Series A led by Andreessen Horowitz, and a $5B AI Infrastructure Fund, Volta is building the infrastructure layer of the AI era from the ground up. We are 100+ people across London, Palo Alto, and New York, with rapid growth expectations to hundreds., Volta builds and operates large scale GPU compute infrastructure for AI workloads. Our platform is Kubernetes-native, spans multiple regions, and delivers virtual machines, storage, and networking through a fully automated infrastructure stack built on custom Kubernetes operators.

We are building out several platform engineering teams that together own the full stack, from managed bare metal and IaaS through to higher-order platform services. Each team owns a different part of that stack: compute, networking, storage, the control plane and API layer, confidential computing, and the customer-facing surface. The area you work in depends on the team you join and your prior expertise, so no single engineer is expected to cover all of it.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

4:18 min

Prioritizing communication and structural awareness over strict tool mastery

Liam Hurrel +1 · World Congress 2021

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all