HPC Administrator
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+11 more
Job description
We are looking for an experienced HPC Administrator to support the design, build, and operation of Linux-based high-performance computing clusters (CPU and GPU). You will manage system administration tasks, ensure cluster stability, maintain software stacks, and provide support for HPC users and scientific computing environments., * Design and build Linux-based HPC CPU/GPU clusters.
- Perform system administration, software maintenance, system monitoring, and troubleshooting for HPC/GPU clusters.
- Manage and operate HPC facilities and infrastructure.
- Configure and support parallel computing environments:CUDA,OpenMP,MPI
- Deploy and manage Infiniband (ultra low-latency networks) and Ethernet high-performance networks.
- Administer batch scheduling systems such as Slurm
- Manage compute resources (CPU, GPU, RAM) for serial and parallel workloads.
- Compile, install, update, and tune scientific software and libraries (commercial and open-source) on compute nodes and HPC file systems.
- Maintain and update firmware and drivers related to HPC systems.
- Provide user support for HPC environments, including troubleshooting, software assistance, guidance on cluster usage, and performance optimization., Work at the heart of cutting-edge supercomputing projects in Europe, alongside top engineers, in an environment that supports learning, ownership, and professional growth. You will collaborate with highly skilled HPC and AI specialists across Europe and contribute to infrastructures that support major scientific and industrial innovation. #Bull
Here, your ideas, your curiosity and your technical excellence directly shape the next era of advanced computing - unlocking enterprise value, accelerating scientific progress and driving positive impact for society.
Requirements
- Experience designing and managing HPC Linux clusters (CPU/GPU).
- Strong background in Linux system administration (preferably RHEL-based).
- Knowledge of parallel programming environments (CUDA, OpenMP, MPI).
- Experience administering Slurm and managing HPC compute resources.
- Hands-on experience with Infiniband and high-performance networking.
- Proficiency with compiling and maintaining scientific software stacks.
- Strong troubleshooting skills in complex HPC environments.
Nice to Have
- Familiarity with HPC storage systems.
- Scripting knowledge (Bash, Python).
- Experience in performance tuning of HPC environments.
About the company
Bull is a story. One with a century of European innovation and a working environment where experts design powerful, sustainable, and sovereign digital solutions, enabling states and industries to retain full control over their data and their AI.
Bull is also thousands of engineers, researchers and passionate tech people shaping the future of high-performance computing, AI, and quantum technologies.
Every day, our teams push the boundaries of what is technologically possible - from next-generation HPC architectures to exascale supercomputers - supported by world-class R&D, more than 1,600 patents, and unique end-to-end capabilities spanning hardware design, software engineering, data science and quantum research.
We are a people-centric, innovation-driven company, where collaboration spans Europe, the Americas and India. We share a common vision of a responsible and sustainable innovation that delivers concrete impact for our customers.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Top 6 Hackathons for Developers in 2023
7 Cloud Computing Trends Coming in 2025 for Developers
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production
Making Data Warehouses Fast: A Developer’s Story