Hardware Engineer
STN, inc.
United States
about 2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source
Tech stack
Computing Platforms
Systems Engineering
Intelligent Platform Management Interface
Bash Shell
BIOS
Configuration Management Databases
Computer Engineering
Data Centers
Linux
Firmware
Python (Programming Language)
Systems Architecture
+3 more
Scripting
Graphics Processing Unit (GPU)
Information Technology
Job description
The Hardware Engineer owns hardware lifecycle for GPU and supporting infrastructure assets, including fleet health monitoring, RMA workflows, firmware management, and long-range capacity planning. The role is the technical owner of the physical compute platform., * Monitor GPU and server health including thermal, error rates, and component failures
- Drive the RMA process with vendors (NVIDIA, Supermicro, HPE, and others) end-to-end
- Manage firmware, BIOS, and BMC upgrade campaigns across the fleet
- Develop hardware burn-in and acceptance test procedures, including NCCL and stress tests
- Investigate hardware failures and produce vendor-grade root cause analyses
- Maintain hardware inventory, asset records, and CMDB accuracy
- Drive capacity planning across compute, storage, and networking
- Coordinate with Procurement on spare parts strategy and stocking levels
- Author hardware engineering runbooks and operational procedures
- Support new platform bring-up, qualification, and reference architecture validation
Requirements
Do you have experience in System architecture?, * 5+ years in hardware engineering, systems engineering, or data center engineering
- Deep knowledge of x86 server architecture, GPU systems, and modern storage
- Hands-on experience with NVIDIA HGX, DGX, or hyperscale-class systems
- Strong Linux fundamentals and scripting skills (Python, Bash)
- Bachelor’s degree in computer science, electrical engineering, or related field, * Experience with NVIDIA Mission Control, Base Command Manager, or Bright Cluster Manager
- Familiarity with IPMI, Redfish, and vendor management interfaces
- Knowledge of liquid cooling and high-density power architectures
- Experience operating fleets of 1,000+ GPUs
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
CS
Christina Schaireiter
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
2 months ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago
LM
Luis Minvielle
Top 6 Hackathons for Developers in 2023
about 3 years ago
EM
Eli McGarvie
The Best X (Twitter) Accounts for Developers
almost 3 years ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago