HPC Deployment and Cable Validation
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
The Candidate will provide deployment services for HPC and AI cluster infrastructure with an emphasis on
physical build quality, cable validation, rack integration, and cluster readiness. This role is for a senior hands-on
operator who can turn physical deployment work into repeatable, auditable infrastructure quality.
WHAT THIS CANDIDATE WILL BE DOING
Execute and coordinate rack-level deployment of compute, network, and storage infrastructure for AI and
HPC environments.
Validate rack elevations, port maps, cable maps, power distribution, labeling, and physical connectivity
before cluster handoff.
Perform detailed cable validation across Ethernet, InfiniBand, management, and storage interconnects.
Detect and resolve cabling defects including polarity issues, incorrect port destinations, unsupported optics
pairings, breakout errors, damaged media, and inconsistent labeling.
Support hardware bring-up by validating BIOS baselines, BMC reachability, inventory accuracy, and initial
connectivity tests.
Partner with network and Linux teams during cluster turn-up to quickly isolate physical-layer versus logical-
layer failures.
Create deployment checklists, as-built documentation, and signoff criteria for production acceptance.
Drive structured remediation during expansion, re-cabling, or failed build events.
Requirements
7+ years in data center deployment, HPC infrastructure installation, or large-scale hardware integration.
Strong practical knowledge of structured cabling, optics, transceivers, breakout schemes, and rack-level
physical validation.
Experience reading and validating rack diagrams, patch plans, cable matrices, and topology documentation.
Strong troubleshooting skill for L1 issues that manifest as network instability, missing hosts, degraded
performance, or failed cluster readiness checks.
Experience with asset tracking, labeling discipline, and deployment quality control.
Ability to work across physical infrastructure, server hardware, and network operations teams without
losing detail.
PREFERRED EXPERIENCE
Experience in GPU cluster deployment or AI factory buildouts.
Familiarity with automated cable validation or DCIM-integrated deployment workflows.
Working knowledge of host provisioning and network validation enough to accelerate multi-team turn-up.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
Top 6 Hackathons for Developers in 2023
Why Attend a Developer Event in 2026?
Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters