> Markdown version of [/jobs/ext/1403042-senior-ai-gpu-deployment-engineer](https://www.wearedevelopers.com/jobs/ext/1403042-senior-ai-gpu-deployment-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI GPU Deployment Engineer - **Company:** 5C DATA CENTERS USA INC. - **Location:** Springfield, OH, United States (Remote available) - **Experience:** Expert - **Salary:** $120,000.0 - $150,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Computing Platforms, Bash Shell, BIOS, Computer Clusters, Configuration Management, Data Centers, Ethernet, Network Interface Controllers, Firmware, InfiniBand, Subnetting, Python (Programming Language), Linux System Administration, Remote Direct Memory Access, Ansible, SQL Databases, AI Infrastructure, Graphics Processing Unit (GPU), High Performance Computing, Infrastructure Automation Frameworks, Information Technology, Hardware Infrastructure - **Published:** July 23, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=f876384524162e95 ## About the Role The ideal candidate brings deep technical expertise in GPU infrastructure, network fabrics, storage, automation, and Linux systems administration, combined with strong execution and troubleshooting skills., * Bachelor's degree in Computer Science, Engineering, IT, or related field (or equivalent experience) * 5+ years of infrastructure engineering or datacenter deployment experience * 3+ years deploying large-scale AI, HPC, or GPU infrastructure * Hands-on experience deploying and operating large GPU clusters in enterprise or hyperscale environments * Strong expertise with: + GPU architectures + InfiniBand (NDR/XDR) and Ethernet GPU fabrics (Spectrum-X) + NVLink, NVSwitch, and GPU-direct technologies + Canonical MaaS and automated provisioning systems + VAST Data or similar high-performance storage platforms + Linux systems administration for HPC/AI workloads + Infrastructure-as-Code and configuration management (Ansible) + Python, Shell, and SQL for infrastructure automation and diagnostics * Strong understanding of: + RDMA, RoCE, and lossless Ethernet fabrics + Cluster automation, observability, and lifecycle management ## Description We are seeking an experienced Senior AI GPU Deployment Engineer to plan, deploy, and operationalize large-scale GPU AI infrastructure environments. This role delivers production-grade GPU clusters supporting AI training, inference, and high-performance computing workloads in our hyperscale data centers., We're looking for a Senior AI GPU Deployment Engineer to join our Cloud Ops team in the United States. This person will share our company values and play an important role in supporting our continued growth. What You Will Do * Deploy integrate and validate multi-rack GPU-based compute platform deployments * Deploy fabric configuration engines (Subnet Manager), observability platforms (UFM) and validate interconnect and fabric performance (nccl) * Collaborate with network engineering team on topology implementation and optimization and storage engineering team on deployment and integration of high-performance storage environments supporting AI workloads (e.g. VAST Data) * Configure settings and manage firmware updates for GPUs, NICs, BMC, BIOS and other components across large-scale clusters * Contribute to infrastructure-as-code automation development for cluster provisioning and lifecycle management * Contribute to improving and documenting repeatable deployment methodologies and scalable operational standards * Query and analyze deployment outcomes using SQL for diagnostics and operational reporting ## Related Videos - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [10M Data Records Lost, Underwater Computing, and Psychedelic Fish - Matthias Geniar](https://www.wearedevelopers.com/videos/1908-10m-data-records-lost-underwater-computing-and-psychedelic-fish-matthias-geniar) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud)