Data Center Technician / Engineer

TRUSTIT LLC
Reno, NV, United States
3 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Microsoft Windows Apple Mac Systems Data Centers Data Center Infrastructure Management (CIM) Software Debugging Linux Domain Name System (DNS) Python (Programming Language) Network File Systems Network Protocols Ansible TCP/IP
+6 more
Scripting Graphics Processing Unit (GPU) Transport Layer Security Break Fix Slurm Hardware Infrastructure

Job description

  • Core Specialty: GPU & Server Hardware Break/Fix, Linux/Windows OS Administration, DCIM Tooling, Scripting (Python/Shell/Ansible)
  • High-Value Skills: High-Performance Computing (HPC) clusters (Slurm, Bright Cluster Manager), liquid cooling, dense rack layout, networking protocols (TCP/IP, DNS, NFS, SSL), * Hardware & Compute Farm Management: Maintain a high-performing compute farm of builders, packagers, testers, and core server infrastructure.
  • Server & GPU Break/Fix: Perform hands-on troubleshooting and replacement for PCBs, GPUs, power supplies, memory, and high-density compute nodes.
  • Automation & Scripting: Use Shell, Python, or Ansible to automate recurring tasks, run operational scripts, and manage DCIM tooling (e.g., Nautobot).
  • Cross-Functional Operations: Collaborate with system architects, software developers, and QA engineers to debug hardware/software edge cases and meet availability SLAs.
  • Process Documentation: Author Standard Operating Procedures (SOPs), collect key operational metrics, and manage system recovery efforts.

Requirements

  • Associate s or Bachelor s degree in a technical major (or equivalent hands-on experience).
  • 5 to 8 years of direct experience in data center environments or large engineering labs.
  • Strong operating system administration across Linux, Windows, and macOS.
  • Hands-on scripting proficiency with Python, Shell, or Ansible.
  • Working knowledge of network protocols: TCP/IP, DNS, NFS, SSL.
  • Direct experience with DCIM tools (Nautobot or similar inventory/rack management systems).

Preferred / Standout Skills:

  • Experience managing HPC clusters using Slurm or Bright Cluster Manager (BCM).
  • Knowledge of dense server infrastructure, including liquid cooling systems.
  • Network certifications such as CCNA or equivalent.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:29 min

Recommended community resources for cloud engineers

Piet Van Dongen · LIVE

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

3:50 min

Queues in TCP stacks and continuous network connections

Clemens Vasters Clemens Vasters · World Congress 2022

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all