> Markdown version of [/jobs/ext/2866286-hpc-linux-systems-engineer](https://www.wearedevelopers.com/jobs/ext/2866286-hpc-linux-systems-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # HPC Linux Systems Engineer - **Company:** Cadre5, LLC. - **Location:** Knoxville, TN, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Computer Clusters, Configuration Management, Linux, File Systems, General Parallel File Systems, Python (Programming Language), Scripting, Slurm, Puppet - **Published:** September 12, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=f5ca72362683e27c ## About the Role * Bachelor's Degree in a scientific or technical field with a combination of 5-8 years of Linux systems experience is required. An equivalent combination of education and experience will be considered. * 8+ years of Linux systems experience is required * Demonstrated experience working in classified environments, including a thorough understanding of security policies, compliance frameworks, and associated standard processes (e.g., NIST, DISA STIGs). * Demonstrated experience designing and deploying of HPC systems, ensuring they meet the computational needs and security requirements of a classified environment. * The ability to obtain and maintain a Department of Energy "Q" clearance with SCI is required. This requires US Citizenship., * Experience managing Linux operating systems in a large-scale system environment * Solid understanding of networked computing environment concepts * 5 + years of experience with Linux Cluster Administration * Ability to develop and maintain programs and scripts that aid in the operation and automation of administrative tasks * using various shell and scripting languages (bash, Python, Go) * Experience with Lustre and GPFS file systems * Experience with batch schedulers (particularly SLURM) * Experience deploying and maintaining automated configuration management software such as Puppet ## Description * Install, integrate, and administer HPC Linux clusters and high-speed networks * Diagnosing system operational problems quickly and effectively * Coordinating with vendors to resolve hardware and software problems * Recommending, planning, and coordinating hardware and software changes with customer participation using * change management processes * Porting and writing system management tools * Documenting system administration procedures for routine and complex tasks * Participating in a 24-hour, 7-day on-call support rotation and off-hours maintenance windows * System implementation/integration into the NCCS environment and systems performance analysis. * Lead system deployment, integration and troubleshooting of a large-scale computer system. * Participate in relevant systems topics with the internal and external community of peers contributing experiences * and solutions. * Mentor junior-level staff as they join the group. ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Automate everything via NodeJS and Puppeteer](https://www.wearedevelopers.com/videos/322-automate-everything-via-nodejs-and-puppeteer) - [From Cloud Racks to Control Cabinets: Operating Kubernetes on Edge Devices](https://www.wearedevelopers.com/videos/100160-from-cloud-racks-to-control-cabinets-operating-kubernetes-on-edge-devices) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [The Memory Leak That Ate Our Cluster: A Postmortem](https://www.wearedevelopers.com/videos/2057-the-memory-leak-that-ate-our-cluster-a-postmortem) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Top 6 Hackathons for Developers in 2023](https://www.wearedevelopers.com/magazine/263-top-6-hackathons-for-developers-in-2023) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany)