IT Platform Engineer - HPC & Linux

Red Bull
Milton Keynes, UK
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Computing Platforms Application Services Microsoft Azure Bash Shell Configuration Management Computer Programming Computer Networks Continuous Delivery Linux DevOps Ethernet
+25 more
Github Identity and Access Management InfiniBand Python (Programming Language) Key Management Octopus Deploy OpenStack Ansible Software Engineering Vault (Revision Control System) Virtualization Technology Software Vulnerability Management Data Logging Cloud Platform System High Performance Computing Grafana HybridCloud Gitlab Git Kubernetes Infrastructure Automation Frameworks Apache Kafka Devsecops Docker Golang

Job description

Design, build, operate and continuously improve secure, scalable and reliable platform services supporting engineering, simulation and HPC workloads across hybrid cloud and on-premises environments. The role combines platform engineering, automation, infrastructure operations and developer enablement, delivering self-service capabilities and platform products that improve the productivity, reliability and efficiency of engineering teams.

Accountabilities:

  • Design, implement and maintain innovative platforms end-to-end for the Technology Campus
  • Design and develop self-service platform capabilities that enable engineering teams to provision and consume infrastructure, compute, storage and application services consistently and securely.
  • Define and maintain service reliability objectives, capacity plans and operational metrics to ensure platform availability, performance and scalability.
  • Provide advanced technical support and manage problem escalations from users communicating well and recording in ticketing management system.
  • Partner with software engineering teams to understand application requirements, improve developer experience, support cloud-native delivery practices and ensure platform services effectively meet the needs of engineering workloads.
  • Own on-premises and cloud-hosted platform services, proactively identifying reliability, performance and scalability improvements and leading the design and implementation of appropriate solutions.
  • Drive automation and Infrastructure as Code practices to improve reliability, consistency, and operational efficiency of platform services
  • Contribute to platform governance forums, providing technical input, sharing knowledge, and supporting alignment across infrastructure and engineering teamsEnable efficient and reliable consumption of platform services by internal teams, with a focus on usability, repeatability, and reduced operational friction

Additional Accountabilities:

  • Participate in wider team projects, change management and take an active role in reviewing architecture of new solutions across Platforms infrastructure and HPC
  • Contribute to platform architecture, technology roadmaps and lifecycle management decisions to ensure services remain fit for purpose and aligned with future business requirements.
  • Maintain accurate platform documentation, asset records, and operational runbooks to support effective operation and knowledge sharing
  • Ensure platform solutions comply with organisational security standards, policies, and regulatory requirementsWork closely with vendors to leverage their expertise and solutions, making use of technical partnerships to improve performance, inform future technical direction, execute proof of concepts and to implement new technology

Requirements

  • Expert knowledge administering Linux/Unix systems
  • Experience in scripting and programming e.g. Python, Bash, Go
  • Experience implementing CI/CD and GitOps workflows using tools such as GitHub Actions, GitLab, ArgoCD or Flux.
  • Working knowledge of platform security principles including identity management, secrets management, vulnerability remediation, hardening and least-privilege access controls.
  • Network knowledge of InfiniBand, MPI and Ethernet concepts
  • Working knowledge of platform security principles.
  • Experience with configuration management and infrastructure as code tools (Ansible, Git, Vault)
  • Knowledge of container orchestration tools, virtualisation and observability stacks e.g. Kubernetes, Grafana, Kafka, Docker, OpenStack, OLVM.
  • Experience implementing platform and application observability solutions including monitoring, logging, tracing, alerting and telemetry.
  • Awareness of secure software development and DevSecOps principles.
  • Experience working within Agile and DevOps delivery models.
  • Knowledge of software delivery platforms such as GitHub, GitLab, Azure DevOps or equivalent.
  • Experience designing, deploying and operating infrastructure services within public cloud platforms.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on uk.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all