Data Center Technician / Engineer
TRUSTIT LLC
Reno, NV, United States
3 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source
Tech stack
Microsoft Windows
Apple Mac Systems
Data Centers
Data Center Infrastructure Management (CIM)
Software Debugging
Linux
Domain Name System (DNS)
Python (Programming Language)
Network File Systems
Network Protocols
Ansible
TCP/IP
+6 more
Scripting
Graphics Processing Unit (GPU)
Transport Layer Security
Break Fix
Slurm
Hardware Infrastructure
Job description
- Core Specialty: GPU & Server Hardware Break/Fix, Linux/Windows OS Administration, DCIM Tooling, Scripting (Python/Shell/Ansible)
- High-Value Skills: High-Performance Computing (HPC) clusters (Slurm, Bright Cluster Manager), liquid cooling, dense rack layout, networking protocols (TCP/IP, DNS, NFS, SSL), * Hardware & Compute Farm Management: Maintain a high-performing compute farm of builders, packagers, testers, and core server infrastructure.
- Server & GPU Break/Fix: Perform hands-on troubleshooting and replacement for PCBs, GPUs, power supplies, memory, and high-density compute nodes.
- Automation & Scripting: Use Shell, Python, or Ansible to automate recurring tasks, run operational scripts, and manage DCIM tooling (e.g., Nautobot).
- Cross-Functional Operations: Collaborate with system architects, software developers, and QA engineers to debug hardware/software edge cases and meet availability SLAs.
- Process Documentation: Author Standard Operating Procedures (SOPs), collect key operational metrics, and manage system recovery efforts.
Requirements
- Associate s or Bachelor s degree in a technical major (or equivalent hands-on experience).
- 5 to 8 years of direct experience in data center environments or large engineering labs.
- Strong operating system administration across Linux, Windows, and macOS.
- Hands-on scripting proficiency with Python, Shell, or Ansible.
- Working knowledge of network protocols: TCP/IP, DNS, NFS, SSL.
- Direct experience with DCIM tools (Nautobot or similar inventory/rack management systems).
Preferred / Standout Skills:
- Experience managing HPC clusters using Slurm or Bright Cluster Manager (BCM).
- Knowledge of dense server infrastructure, including liquid cooling systems.
- Network certifications such as CCNA or equivalent.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
CS
Christina Schaireiter
4 months ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
about 1 month ago
LM
Luis Minvielle
A Guide to Green Tech and Green IT Careers
over 2 years ago
LM
Luis Minvielle
The Most Popular IT Jobs on the Market
over 2 years ago
KD
Krissy Davis
Best Paying Jobs in Technology
about 3 years ago