TS/SCI HPC Systems Engineer

Insight Global
Ravanna, MO, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Systems Engineering Bash Shell Command-Line Interface Linux Python (Programming Language) Linux System Administration Scripting High Performance Computing Git Information Technology Slurm Software Version Control
+1 more
Server Operating Systems & Platforms

Job description

A federal IT services client of Insight Global is hiring for a highly skilled HPC Systems Engineer to join their team full time in Charlottesville, VA. This role requires an active TS/SCI clearance and is 5 days/week on site. Relocation packages are available!

The HPC (High Performance Computing) Systems Engineer will work directly with engineers, analysts, and researchers to support job execution, troubleshoot workload failures, and improve the performance and efficiency of compute workloads running on HPC clusters. The Engineer will assist users with scheduler job scripts, application execution, and workload performance troubleshooting while promoting HPC best practices for efficient cluster utilization. This role serves as the primary interface between mission users and HPC platform infrastructure teams.

Key Responsibilities:

  • Provide direct support to users running computational workloads on HPC clusters (classified & unclassified)

  • Assist with creating, submitting, and troubleshooting job scripts (Slurm, PBS), including CPU/GPU resource allocation

  • Diagnose and resolve failing, slow, or hanging jobs (including MPI, parallel, and GPU workloads)

  • Support application setup, compilation, and execution in Linux-based HPC environments

  • Advise users on best practices to improve job performance, efficiency, and resource utilization

  • Monitor workload usage and recommend optimizations to maximize cluster throughput

  • Develop and maintain automation scripts/tools (Bash/Python) and manage them in version control (Git)

  • Collaborate with infrastructure teams and maintain documentation to resolve system issues and support users

Requirements

Active TS/SCI clearance

  • ONE of the following certifications: Security+, CCNA Security, CySA+, GICSP, GSEC, CND, SSCP, CAP, CASP+, CISM, CISSP, GSLC, CCISO, HCISPP

  • 5+ years of experience working in Linux environments supporting distributed compute workloads or HPC cluster platforms

  • Experience executing or troubleshooting workloads using HPC workload schedulers such as Slurm, PBS, Torque, or similar systems

  • Experience administering command-line Linux systems including scripting (Bash, python, etc.) and troubleshooting applications in multi-user server environments.

  • Experience supporting systems within DoD/DoW or IC environments

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · WWC Europe 2026

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · WWC Europe 2026

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all