> Markdown version of [/jobs/ext/2804891-data-center-lab-operations-engineer](https://www.wearedevelopers.com/jobs/ext/2804891-data-center-lab-operations-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Center Lab Operations Engineer - **Company:** CareerCircle - **Location:** Reno, NV, United States - **Experience:** Expert - **Salary:** $72,800.0 - **Contract:** Temporary to permanent - **Skills:** Artificial Intelligence, Systems Engineering, Software Quality, Data Centers, Data Center Infrastructure Management (CIM), Software Debugging, Linux, Domain Name System (DNS), Firmware, Networking Hardware, IP Addressing, Python (Programming Language), Linux System Administration, Machine Learning, Networking Basics, Network File Systems, Network Administration, Ansible, Software Engineering, System Testing, TCP/IP, Network Switches, Network Routing, Graphics Processing Unit (GPU), Transport Layer Security, High Performance Computing, Infrastructure Automation Frameworks, Operational Systems, Slurm, Hardware Infrastructure, Network Server - **Published:** September 9, 2026 - **Apply:** https://www.careercircle.com/jobs/all/all/usa/nv/reno/ba27320c-30a1-4894-b725-b75610d32417 ## About the Role Operations Leadership Management Automation Data Centers Communication Problem Solving Asset Management Safety Assurance Machine Learning Influencing Skills Business Valuation Process Improvement Systems Engineering Mechanical Aptitude Network File Systems Full Stack Development Hardware Installations Artificial Intelligence Business Transformation Performance Engineering Hardware Troubleshooting Infrastructure Management Software Quality (SQA/SQC) High Performance Computing Python (Programming Language) Influencing Without Authority Troubleshooting (Problem Solving) Slurm (Batch Scheduling Software), * 5+ years of experience supporting data centers, engineering labs, compute environments, or similar infrastructure operations. * Strong understanding of Linux and Windows administration. * Experience troubleshooting hardware, servers, networking, and operating systems. * Knowledge of networking fundamentals including TCP/IP, DNS, NFS, SSL, and related protocols. * Experience with scripting and automation tools such as Python, Shell, or Ansible. * Familiarity with DCIM tools and asset management platforms. * Strong documentation, communication, and problem-solving skills. * Ability to work cross-functionally with both technical and non-technical stakeholders., * Experience supporting HPC or AI/ML compute clusters. * Familiarity with Slurm, BCM, or other cluster management platforms. * Experience working with GPUs, server hardware, PCBs, and system validation environments. * CCNA or similar infrastructure certification. * Knowledge of storage, networking, and dense data center architectures. * Experience with liquid cooling technologies. * Strong mechanical aptitude and comfort working with tools and hardware installations., Individual compensation offered for this position within this range will depend on many factors, including qualifications, skills, relevant experience, job knowledge, geographic location, internal equity, and other pertinent job-related factors. ## Description We are seeking a Data Center Lab Operations Engineer to support a high-performance compute environment used for developing and validating next-generation technologies. This role will partner closely with engineering, QA, and operations teams to maintain critical infrastructure, troubleshoot complex hardware and software issues, and drive continuous improvements across compute and testing operations., * Support and maintain a large-scale compute farm consisting of builders, testers, packagers, and supporting infrastructure. * Collaborate with hardware, software, QA, and systems engineering teams to deploy, troubleshoot, and optimize engineering environments. * Monitor system health, availability, and performance while ensuring SLA targets are achieved. * Lead troubleshooting and recovery efforts for infrastructure, hardware, and software incidents. * Assist with server deployments, hardware installations, upgrades, and maintenance activities. * Develop and maintain SOPs, documentation, and operational processes. * Collect and analyze infrastructure metrics to drive improvements in reliability and efficiency. * Implement automation and process improvements using scripting and infrastructure management tools., Use of Artificial Intelligence (AI): We may use Artificial Intelligence (AI) to support parts of our hiring process, including sourcing, screening, and evaluating candidates. AI helps assess applications and qualifications, but final decisions are made by our hiring team. By applying, you acknowledge and agree that your application may be reviewed using AI tools. Related Jobs Data Center Technician TEKsystems Sparks, NV*On-Site Operations IP Addressing Network Routing Firmware Updates Business Valuation Category 6 Cabling Category 5 Cabling Full Stack Development Network Administration Artificial Intelligence Business Transformation Submittals (Construction) Troubleshooting (Problem Solving) +0 Data Center Engineer TEKsystems Sparks, NV*On-Site Linux Debugging Operations System Recovery Safety Assurance Hardware Support Quality Assurance Business Valuation Mechanical Aptitude Software Engineering Linux Administration Full Stack Development Artificial Intelligence Business Transformation Hardware Troubleshooting High Performance Computing Standard Operating Procedure Slurm (Batch Scheduling Software) +0 Google IT Automation with Python Data Center Technician TEKsystems Sparks, NV*On-Site Firmware Operations Leadership IP Addressing Problem Solving Desktop Support Network Switches Help Desk Support Business Valuation Category 6 Cabling Category 5 Cabling Networking Hardware Full Stack Development Data Center Operations Artificial Intelligence Business Transformation Hardware Troubleshooting Troubleshooting (Problem Solving) +0 ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Best Paying Jobs in Technology](https://www.wearedevelopers.com/magazine/256-best-paying-jobs-in-technology) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)