Systems Administrator V

General Atomics
San Diego, CA, United States
24 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

Proxmox Private Networks Artificial Intelligence Bash Shell Configuration Management System Configuration Data Centers Data Loss Data Structures Software Debugging Linux RAID
+32 more
File Systems Ethernet General-Purpose Computing on Graphics Processing Units Networking Hardware Python (Programming Language) Logical Volume Manager Machine Learning Message Broker Networking Basics Network Protocols Open Source Technology Performance Tuning Systems Development Life Cycle Remote Direct Memory Access Red Hat Enterprise Linux Redis Ansible Scientific Computating Transmission Control Protocol (TCP) Virtualization Technology Scripting Caching Storage Technologies Information Technology Low Latency Bare Metal Apache Kafka Operational Systems Slurm ZFS File System Data Pipelines Nvme

Job description

General Atomics (GA), and its affiliated companies, is one of the world’s leading resources for high-technology systems development ranging from the nuclear fuel cycle to remotely piloted aircraft, airborne sensors, and advanced electric, electronic, wireless and laser technologies.

With consultative direction within the Core Services group, this position is responsible for architecting, optimizing, and maintaining a high-velocity, on-premises Linux data pipeline supporting real-time, AI-driven machine learning models for the DIII-D national fusion reactor program.

As a hands-on technical practitioner, you will manage bare-metal deployments, orchestrate high-availability parallel storage arrays, and troubleshoot ultra-low-latency network protocols across a heavily segmented, secure infrastructure. This role focuses on deep, OS-level infrastructure and physical data center operations, bridging the gap between raw hardware capabilities and mission-critical scientific computing., + Bare-Metal & OS Administration: Plan, manage, and optimize day-to-day operations of on-premises, physical x86 and GPU server infrastructure running Red Hat Enterprise Linux (RHEL).

  • High-Velocity Networking: Provision and troubleshoot core network sharing protocols (NFS) and ultra-low-latency networking hardware supporting 100GbE backbones and RDMA over Converged Ethernet (RoCEv2).

  • Operational Technology (OT) Security: Maintain and fortify a heavily firewalled, push-only segmented internal network architecture, ensuring strict “Deny All Inbound” rules protect sensitive reactor control systems.

  • Storage Architecture Lifecycle: Build, configure, and maintain high-density NVMe storage tiers and software-defined, parallel filesystems (BeeGFS/ZFS) to handle massive multi-gigabyte payload dumps with zero ingestion bottlenecks.

  • Performance & I/O Benchmarking: Conduct granular system-level I/O benchmarking, diagnose deep kernel-level platform anomalies, and implement kernel parameter tuning to maximize data velocity and prevent data loss.

  • Automation & Provisioning: Develop and maintain automated infrastructure workflows and configuration management frameworks using Bash shell scripting and Ansible to streamline bare-metal node deployment and environment validation.

  • Data Center Operations: Manage physical data center footprints, including high-density server rack layouts, equipment delivery, hardware diagnostics, power infrastructure/UPS, and asset lifecycle tracking.

  • Collaborative Consultation: Act as a technical infrastructure expert, guiding the development of innovative solutions to unique computing challenges for data acquisition systems and scientific teams.

  • Vendor & Planning Coordination: Analyze new hardware architectures, engineer custom hardware Bill of Materials (BOM) alongside OEMs/ODMs, and represent the organization as a primary technical contact with suppliers.

  • Compliance & Safety: Observe all laws, regulations, and facility safety obligations, ensuring that all system upgrades and physical data center modifications are designed with established personnel operating procedures properly considered.

We recognize and appreciate the value and contributions of individuals with diverse backgrounds and experiences and welcome all qualified individuals to apply.

Requirements

  • Typically requires a bachelor’s degree in information technology or a related discipline and fifteen or more years of progressive professional experience in an information technology department primarily in systems administration. May substitute equivalent working experience in the field in lieu

  • Expert Linux Skills: Detailed and extensive technical expertise in Red Hat Enterprise Linux (RHEL) system design, installation, configuration, and low-level kernel troubleshooting.

  • Storage Mastery: Proven hands-on experience implementing, configuring, and performance-tuning ZFS, RAID, and LVM storage structures.

  • Networking Fundamentals: Comprehensive understanding of TCP/IP networking, host-based firewalls, network bonding, and debugging network sharing protocols across segmented topologies.

  • Automation Background: Demonstrated proficiency writing advanced automation scripts in Bash or Python, coupled with configuration management toolsets (Ansible).

  • Physical Infrastructure Competency: Direct experience handling physical server deployment, rack configuration, hardware component replacement, and diagnostic testing within an enterprise data center environment.

Preferred (Nice-to-Have) Pipeline Skills:

  • Familiarity with High-Performance Computing (HPC) cluster environments and workload scheduling managers (e.g., SLURM, PBS Pro).

  • Experience configuring or administering parallel file systems (e.g., BeeGFS).

  • Familiarity with distributed, event-driven streaming architectures or message brokers (e.g., Apache Kafka).

  • Familiarity with in-memory data structures or caching layers (e.g., Redis).

  • Experience with open-source hypervisors and virtualization platforms (e.g., Proxmox).

  • Exposure to GPU computing clusters and high-performance environments (e.g., NVIDIA/Mellanox).

About the company

General Atomics

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

7:31 min

Essential foundational skills and concepts for infrastructure roles

Megha Kadur · LIVE

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · WWC Europe 2026

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · WWC Europe 2026

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all