HPC Network Engineer

Mirantis
Berlin, Germany
4 days ago
Apply on jobs.smartrecruiters.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Bash Shell Command-Line Interface Computer Networks Linux Distributed Systems Ethernet Monitoring of Systems InfiniBand Subnetting Virtual Private Networks (VPN) Python (Programming Language)
+17 more
Network Security Network Configuration and Change Management Networking Basics Network Monitoring Routing Remote Direct Memory Access Ansible TCP/IP Virtual Local Area Networks AI Infrastructure Scripting Software Troubleshooting Data Center Networking Kubernetes Low Latency Fortinet Firewall Services Module

Job description

We are looking for a motivated HPC (High Performance Computing) Network Engineer to support and grow within our infrastructure team. This role is ideal for engineers with a strong foundation in networking who are interested in developing deep expertise in InfiniBand, HPC/AI infrastructure, and network security using Fortinet technologies., You will work alongside senior engineers to operate, troubleshoot, and improve high-performance network environments, with a clear growth path toward becoming a Senior HPC Network Engineer., * Support the deployment, configuration, and maintenance of InfiniBand and Ethernet network infrastructure.

  • Assist in troubleshooting network issues, including connectivity, latency, and performance degradation.
  • Monitor network health and performance using standard tools and assist in identifying bottlenecks.
  • Work with InfiniBand technologies (switches, HCAs, subnet managers) under guidance from senior engineers.
  • Support management and troubleshooting of Fortinet solutions (FortiGate, VPNs, firewall rules).
  • Assist in implementing network configurations, routing, and segmentation.
  • Collaborate with compute and storage teams to support HPC and AI workloads.
  • Participate in incident response and operational support activities.
  • Contribute to documentation of network configurations and procedures.
  • Learn and adopt automation tools and practices (e.g., Ansible, scripting).

Requirements

  • 2-4 years of experience in networking, system engineering, or infrastructure roles.
  • Basic understanding of networking fundamentals (TCP/IP, routing, switching, VLANs).
  • Familiarity with Linux systems and command-line tools.
  • Exposure to data center networking or distributed systems environments.
  • Strong willingness to learn and develop expertise in InfiniBand and HPC networking., * Exposure to InfiniBand or high-performance networking concepts.
  • Familiarity with Fortinet or other firewall technologies.
  • Basic scripting skills (e.g., Bash or Python).
  • Understanding of monitoring tools and troubleshooting methodologies.

Growth Opportunities:

  • Hands-on training with InfiniBand fabrics and HPC networking tools.
  • Mentorship from senior engineers and involvement in complex troubleshooting scenarios.
  • Opportunity to gain experience with RDMA, MPI workloads, and large-scale cluster environments.
  • Path toward a Senior HPC Network Engineer role with increasing ownership and responsibility., * A strong problem-solver with curiosity about how systems work at a deeper level.
  • Someone eager to learn and take on increasingly complex networking challenges.
  • A team player who can collaborate effectively and grow within a high-performance engineering environment.

About the company

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment-on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen. Learn more at www.mirantis.com., * Work with some of the most advanced AI infrastructure environments in production today.

  • Gain exposure to NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments.
  • Help define how next-generation AI infrastructure is operated and supported.
  • Be part of a team shaping the future of AI-powered operations through k0rdent AI.
  • Join a growing organisation investing heavily in AI infrastructure and platform services.

It is understood that Mirantis, Inc. may use automated decision-making technology (ADMT) for specific employment-related decisions. Opting out of ADMT use is requested for decisions about evaluation and review connected with the specific employment decision for the position applied for. You also have the right to appeal any decisions made by ADMT by sending your request to [email protected]

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.smartrecruiters.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:19 min

Executing complex workflows using Ansible Automation Platform

Goetz Rieger Goetz Rieger · World Congress 2025

3:50 min

Queues in TCP stacks and continuous network connections

Clemens Vasters Clemens Vasters · World Congress 2022

Videos

See all

Related articles

See all