nHPC Systems Engineer
SBS SYSTEMS LLC
Chantilly, VA, United States
23 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on jobs.military.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source
Tech stack
Amazon Web Services
Systems Engineering
Bash Shell
Configuration Management
Distributed Systems
Job Scheduling
Python (Programming Language)
Performance Tuning
Scripting
High Performance Computing
Containerization
Slurm
+1 more
Docker
Requirements
The ideal candidate will have strong experience with large-scale, on-premises or hybrid HPC infrastructures and a passion for optimizing system performance. While the primary focus is on traditional HPC environments, familiarity with \nAWS-based technologies is beneficial for supporting hybrid or cloud-extended HPC use cases.\n, * Demonstrated experience supporting High Performance Computing (HPC) or large-scale distributed computing environments.\n
- Strong proficiency in Linux/Unix system administration.\n
- Experience with job schedulers such as Slurm, PBS, or LSF.\n
- Understanding of parallel and distributed computing concepts.\n
- Familiarity with high-performance networking and parallel storage systems.\n
- Scripting experience using Bash, Python, or similar languages.\n
- Strong analytical and problem-solving skills.\n
- Excellent communication and collaboration abilities.\n, * Experience supporting federal or defense customers.\n
- Familiarity with GPU-accelerated computing and performance optimization.\n
- Experience with container technologies such as Singularity/Apptainer or Docker.\n
- Exposure to configuration management or automation tools.\n
- Experience with AWS-based technologies (e.g., AWS services supporting HPC or hybrid environments).\n
- Relevant industry or technical certifications.\n
Benefits & conditions
- Manage and support job scheduling and workload management platforms (e.g., Slurm, PBS, or similar).\n
- Monitor and tune system performance to ensure efficient utilization of compute, storage, and network resources.\n
- Support high-speed interconnects (e.g., InfiniBand) and parallel file systems such as Lustre, GPFS, or BeeGFS.\n
- Collaborate with engineers, scientists, and mission stakeholders to translate computational requirements into effective technical solutions.\n
- Implement automation using scripting languages such as Bash or Python.\n
- Ensure system reliability, security, and compliance with organizational and government standards.\n
- Provide technical documentation and user support.\n
- Contribute to capacity planning and infrastructure enhancements.\n
- Nice to Have: Support the integration or extension of HPC workloads using AWS-based technologies for hybrid computing scenarios.\n
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on jobs.military.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
IK
Igor Khokhriakov
25 days ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago
LM
Luis Minvielle
Top 6 Hackathons for Developers in 2023
about 3 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
KD
Krissy Davis
Best Paying Jobs in Technology
about 3 years ago
LM
Luis Minvielle
Is Software Engineering Over-Saturated?
over 2 years ago