Senior HPC Engineer - Hybrid - Inside IR35
Hamilton Barnes
Stevenage, UK
1 day ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.careerboard.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source
Tech stack
Compilers
Nvidia CUDA
Linux
DevOps
General-Purpose Computing on Graphics Processing Units
Web Servers
InfiniBand
IT Management
Red Hat Enterprise Linux
Ansible
Scientific Computating
SSL Certificate Management
+6 more
Graphics Processing Unit (GPU)
High Performance Computing
Containerization
Slurm
Docker
Servicenow
Job description
We are looking for an experienced Senior HPC Engineer to support and maintain a scientific computing environment, with a strong focus on RHEL, Slurm and HPC infrastructure.
You will work closely with research scientists and technical teams to ensure HPC services are secure, reliable and high performing., * Administer, patch and maintain RHEL 7, 8 and 9 across HPC clusters and workstations.
- Deploy, configure and manage Slurm, including queues, partitions and scheduling.
- Monitor cluster health, performance, storage, networking and resource utilisation.
- Install and support scientific applications, compilers, libraries and MPI environments.
- Work with scientists to optimise workloads and resolve application issues.
- Troubleshoot hardware, operating system, scheduler and application problems.
- Manage incidents and service requests through ServiceNow.
- Collaborate with storage, networking, security and DevOps teams.
Requirements
- 10+ years’ enterprise IT experience, including 3-5+ years in HPC or research computing.
- Strong hands-on experience with RHEL 7, 8 and 9.
- Proven experience managing HPC clusters and Slurm.
- Experience supporting scientific or research applications on Linux.
- Strong troubleshooting and root-cause analysis skills.
- Experience with ServiceNow or a similar ITSM platform.
- Strong communication and stakeholder-management skills.
- Able to work onsite in Stevenage 3 days per week.
Desirable Skills
- Docker or container technologies.
- Ansible or similar automation tools.
- GPU computing, CUDA and GPU-accelerated workloads.
- OpenMPI, MPICH or other MPI libraries.
- InfiniBand and high-speed networking.
- Web server and SSL certificate management.
- RHCSA, RHCE or equivalent Red Hat certification.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.careerboard.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
IK
Igor Khokhriakov
28 days ago
LM
Luis Minvielle
Top 6 Hackathons for Developers in 2023
about 3 years ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago
EM
Eli McGarvie
Fullstack Developer Salary UK
about 3 years ago
EM
Eli McGarvie
Best Companies to work for in London: Top 25 Companies in 2023
over 3 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago