IT045 High Performance Computing (HPC) and Storage...

Adnet Systems, Inc.
Greenbelt, MD, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Access Control List Active Directory Agile Methodology Computing Platforms JIRA Bash Shell Ubuntu (Operating System) Command-Line Interface Cyber Security Nvidia CUDA Information Systems Computer Networks
+20 more
System Configuration Linux File Systems Identity and Access Management Job Scheduling Python (Programming Language) Lightweight Directory Access Protocols (LDAP) Linux System Administration Performance Tuning Scrum Methodology Software Maintenance Red Hat Enterprise Linux Shell Script Software Engineering High Performance Computing Computerised Systems Firewalls (Computer Science) Gitlab Patch Management Slurm

Job description

  • Full Operational Management: Perform day-to-day operations and management of large-scale, supercomputing clusters to meet the required availability, and performance, including, but not limited to, integration, provisioning, software stack deployment, updates, hardware and software maintenance, and decommissioning.

  • High-Performance Storage Administration: Deploy, tune, configure, maintain, and operate massive parallel file systems.

  • Workload and Schedule Management: Manage, configure, optimize, and troubleshoot cluster management and job scheduling software.

  • Security, Patches, and Compliance: Proactively implement security updates, coordinate systematic Operating System kernel patches, and mitigate vulnerabilities across computing and storage environments without compromising system stability.

  • Preventative and Corrective Maintenance: Coordinate vendor-supported maintenance schedules, conduct hardware and software diagnostics, and participate in rapid-response resolution during service degradations or system blackouts.

  • User Support: Provide specialized, tiered technical assistance ranging from software provisioning and workflow optimization to advanced, expert-level troubleshooting for complex research challenges.

  • GPU System Administration: Provision, configure, and maintain GPU-accelerated computing systems, including driver management, library configuration, and performance optimization for workload acceleration.

Requirements

This job description is for a High Performance Computing and Storage System Administrator to support the operations of the Integrated Modeling Computing Center (IMCC), formerly known as the NASA Center for Climate Simulation (NCCS). The IMCC will directly support the Integrated Modeling Virtual Institute (IMVI) to meet the Earth science modeling needs for NASA. The following describes the core duties and responsibilities and technical skills. Ideal candidates should have excellent communication skills, problem solving, and the ability to work efficiently within a highly performing team environment., + Expert Linux System Administration: Advanced, production-level expertise in enterprise Linux distributions (RHEL, Rocky Linux, AlmaLinux, or Ubuntu Server), incorporating expert-level command-line proficiency, kernel tuning, and automated shell scripting (Bash, Python).

  • Parallel File Systems Architecture: Hands-on experience in the design, deployment, scaling, and/or optimization of high-performance file systems. Experience in deploying, configuring, and operating IBM Spectrum Scale and/or Lustre.

  • Scheduling Proficiency: Working familiarity with HPC resource management, including experience with Slurm.

  • Systems Security Alignment: Robust foundation in core security frameworks, containing firewalls, identity management (LDAP/Active Directory), access control lists (ACLs), SSH hardening, and continuous patch management cycles.

  • Agile Methodologies: Experience operating within modern Agile frameworks (Scrum, Kanban), leveraging iterative workflows, participating in sprint reviews, and utilizing collaborative project boards (Jira, Gitlab) to track milestones.

  • GPU Accelerator Management: Proficiency in configuring and maintaining GPU-accelerated computing environments, including driver installation/management, CUDA or similar library configuration, and performance tuning for accelerated workloads.

  • A MS degree and 5+ years’ experience in relevant work areas.

  • US Citizenship required.

  • Ability to obtain and maintain a Tier 1 or Tier 2 Investigation through NASA.

Team ADNET brings over 30+ years of experience to information systems and professional services for the federal government. With a history of expertise in software development, computer network design, IT security, mission operations support, and educational outreach, Team ADNET is deeply embedded in the Space and Earth Science at NASA’s Goddard Space Flight Center (GSFC) in Greenbelt, MD.

About the company

ADNET Systems, Inc. is working with Goddard Space Flight Center to fulfill NASA’s vision for space exploration, and working with the Science and Exploration Directorate to fulfill its many missions.

ADNET Systems, Inc. is an employee-centric company, committed to providing premier benefits that support our employees and their families. With affordable medical and dental plans coupled with leading disability and life insurance options, ADNET offers our employees the benefits most sought after by today’s professional candidate. Furthermore, our benefits package features the extras that distinguish us from other small businesses, ensuring our high employee retention that our customers appreciate.

Some features of our compensation plans and environment perks include

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · WWC Europe 2026

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · WWC Europe 2026

Videos

See all

Related articles

See all