Cloud Infrastructure Engineer

Ultralytics
London, UK
12 days ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Computer Vision Microsoft Azure Cloud Computing Cloud Computing Security Cloud Storage DevOps Github Monitoring of Systems Identity and Access Management Python (Programming Language)
+18 more
Machine Learning Networking Basics Reliability Engineering Prometheus Data Processing Google Cloud System Availability Grafana Infrastructure as Code (IaC) Backend Containerization AI Platforms Kubernetes Machine Learning Operations Terraform Serverless Computing Docker Microservices

Job description

Overview

At Ultralytics, we relentlessly drive innovation in AI, building the world’s leading Ultralytics YOLO models. We’re looking for passionate individuals obsessed with AI, eager to make a global impact, and ready to excel in a dynamic, high-energy environment. Join our team and help shape the future of AI.

Location and Legalities

This full-time Cloud Infrastructure Engineer position is based onsite in our brand-new Ultralytics office in London, UK. Applicants must have legal authorization to work in the UK, as Ultralytics does not provide visa sponsorship.

What You’ll Do

As a Cloud Infrastructure Engineer at Ultralytics, you will design, build, and maintain the scalable and robust cloud infrastructure that powers our entire suite of AI services, including the Ultralytics HUB. You’ll be at the core of our operations, ensuring our systems can handle the demands of training cutting-edge AI models and serving millions of inference requests globally. Key responsibilities include:

  • Architecting and managing our cloud environment on Google Cloud Platform (GCP) using Infrastructure as Code (IaC) with Terraform.
  • Designing, building, and maintaining our CI/CD pipelines using GitHub Actions to ensure smooth and reliable software delivery.
  • Containerizing applications with Docker and deploying them as microservices on serverless platforms like Google Cloud Run and orchestrators like GKE.
  • Writing automation scripts and infrastructure-related code, primarily in Python.
  • Ensuring the high availability, performance, and security of our production systems that support everything from data processing to model deployment.
  • Collaborating closely with AI and backend engineers to streamline the MLOps lifecycle and build a world-class platform for computer vision.
  • Monitoring system health and costs, proactively identifying bottlenecks and implementing optimizations.

Your expertise will be critical to the stability and scalability of Ultralytics’ solutions, enabling our team to innovate at an unprecedented pace.

Skills and Experience

  • 5+ years of experience in a Cloud Infrastructure, DevOps, or Site Reliability Engineering (SRE) role.
  • Strong proficiency with Python for scripting and automation.
  • Extensive hands-on experience with Google Cloud Platform (GCP) and its core services (Cloud Run, GKE, IAM, Cloud Storage).
  • Expertise in writing production-grade Infrastructure as Code using Terraform.
  • Deep understanding of containerization with Docker and container orchestration with Kubernetes.
  • Proven experience designing and operating microservices architectures, as described by thought leaders like Martin Fowler.
  • Hands-on experience building and maintaining complex CI/CD pipelines, preferably with GitHub Actions.
  • Familiarity with monitoring and observability tools (e.g., Prometheus, Grafana, Google Cloud’s operations suite).
  • A solid understanding of networking principles and cloud security best practices.
  • Experience with other cloud platforms like Amazon Web Services (AWS) or Microsoft Azure is a plus.

Cultural Fit

Intensity Required

Ultralytics is a high-performance environment for world-class talent obsessed with achieving extraordinary results. We operate at a relentless pace, demanding exceptional dedication and an unwavering commitment to excellence, guided by our mission, vision, and values. Our team thrives on audacious goals and absolute ownership. This is not a conventional workplace. If your priority is predictable comfort or a standard work-life balance over the relentless pursuit of progress, Ultralytics is not for you. We seek driven individuals prepared for the profound personal investment required to make a defining contribution to the future of AI.

Compensation and Benefits

  • Competitive Salary: Highly competitive based on experience.
  • Startup Equity: Participate directly in our company’s growth and success.
  • Hybrid Flexibility: 3 days per week in our brand-new office - 2 days remote
  • Generous Time Off: 24 days vacation, your birthday off, plus local holidays.
  • Flexible Hours: Tailor your working hours to suit your productivity.
  • Tech: Engage with cutting-edge AI projects.
  • Gear: Brand-new Apple MacBook and Apple Display provided.
  • Team: Become part of a supportive and passionate team environment.

If you are driven to redefine the capabilities of machine learning and eager to make a significant impact, Ultralytics offers an exceptional career opportunity. Check out our careers page to apply.

Requirements

  • 5+ years of experience in a Cloud Infrastructure, DevOps, or Site Reliability Engineering (SRE) role.
  • Strong proficiency with Python for scripting and automation.
  • Extensive hands-on experience with Google Cloud Platform (GCP) and its core services (Cloud Run, GKE, IAM, Cloud Storage).
  • Expertise in writing production-grade Infrastructure as Code using Terraform.
  • Deep understanding of containerization with Docker and container orchestration with Kubernetes.
  • Proven experience designing and operating microservices architectures, as described by thought leaders like Martin Fowler.
  • Hands-on experience building and maintaining complex CI/CD pipelines, preferably with GitHub Actions.
  • Familiarity with monitoring and observability tools (e.g., Prometheus, Grafana, Google Cloud’s operations suite).
  • A solid understanding of networking principles and cloud security best practices.
  • Experience with other cloud platforms like Amazon Web Services (AWS) or Microsoft Azure is a plus.

Benefits & conditions

  • Competitive Salary: Highly competitive based on experience.
  • Startup Equity: Participate directly in our company’s growth and success.
  • Hybrid Flexibility: 3 days per week in our brand-new office - 2 days remote
  • Generous Time Off: 24 days vacation, your birthday off, plus local holidays.
  • Flexible Hours: Tailor your working hours to suit your productivity.
  • Tech: Engage with cutting-edge AI projects.
  • Gear: Brand-new Apple MacBook and Apple Display provided.
  • Team: Become part of a supportive and passionate team environment.

About the company

At Ultralytics, we relentlessly drive innovation in AI, building the world’s leading Ultralytics YOLO models. We’re looking for passionate individuals obsessed with AI, eager to make a global impact, and ready to excel in a dynamic, high-energy environment. Join our team and help shape the future of AI., Ultralytics is a high-performance environment for world-class talent obsessed with achieving extraordinary results. We operate at a relentless pace, demanding exceptional dedication and an unwavering commitment to excellence, guided by our mission, vision, and values. Our team thrives on audacious goals and absolute ownership. This is not a conventional workplace. If your priority is predictable comfort or a standard work-life balance over the relentless pursuit of progress, Ultralytics is not for you. We seek driven individuals prepared for the profound personal investment required to make a defining contribution to the future of AI.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all