> Markdown version of [/jobs/ext/2710222-staff-ai-ml-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2710222-staff-ai-ml-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff AI/ML Infrastructure Engineer - **Company:** The Constant Company, LLC - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $145,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Automation of Tests, Intelligent Platform Management Interface, Bash Shell, BIOS, Cloud Computing, Computer Clusters, Linux, Device Drivers, Network Interface Controllers, Firmware, InfiniBand, Python (Programming Language), Machine Learning, Package Management Systems, PCI Express, Software Systems, AI Infrastructure, Infrastructure Automation Frameworks, Bare Metal, Hardware Infrastructure - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/staff-ai-ml-infrastructure-engineer-vultr-com-8094908 ## About the Role * 5+ years experience working with bare metal infrastructure and hardware automation * Hands-on experience with modern NVIDIA/AMD GPU platforms and high-performance networking (RoCE, InfiniBand) * Deep knowledge of BIOS, BMC, firmware, NICs, Redfish/IPMI, and PCIe systems * Strong Linux systems experience including device drivers and package management * Experience building infrastructure automation using Python and Bash * Familiarity with GPU drivers, firmware ecosystems, and vendor collaboration * Experience designing and delivering complex infrastructure products * Proven ability to lead projects and mentor engineers * Experience optimizing multi-cluster GPU environments * Exposure to Machine Learning software stacks and GPU workloads, This salary can vary based on location, years of experience, background and skill set. ## Description Vultr is seeking a highly skilled and experienced Staff AI/ML Infrastructure Engineer to drive the design, performance, and reliability of our AI infrastructure platform. The ideal candidate is a hands-on infrastructure expert with deep GPU systems knowledge, strong automation experience, and a track record of technical leadership in high-performance environments. This is a highly visible role in a high-growth technology company, requiring ownership of complex hardware and software systems, collaboration across engineering and vendor partners, and a relentless focus on operational excellence. This is your opportunity to build the foundation powering next-generation AI workloads and leave a lasting mark on Vultr and the future of cloud infrastructure., * Design and maintain GPU and bare metal infrastructure in containerized and physical environments * Build scalable GPU clusters in partnership with networking and provisioning teams * Ensure reliable, high-performance provisioning of GPU infrastructure * Develop automated testing systems for GPU-based platforms * Implement infrastructure solutions for diverse AI/ML workloads * Benchmark, test, and troubleshoot GPU performance at scale * Collaborate with hardware vendors on drivers, firmware, and support * Resolve hardware, software, and performance issues across environments * Optimize rail and cluster performance across architectures * Lead technical direction and mentor engineers on infrastructure best practices ## Related Videos - [10M Data Records Lost, Underwater Computing, and Psychedelic Fish - Matthias Geniar](https://www.wearedevelopers.com/videos/1908-10m-data-records-lost-underwater-computing-and-psychedelic-fish-matthias-geniar) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Playing Pong on a shoulder press machine](https://www.wearedevelopers.com/videos/100140-playing-pong-on-a-shoulder-press-machine) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [How to Submit CFPs and Get into Public Speaking - Moran Weber](https://www.wearedevelopers.com/videos/2112-how-to-submit-cfps-and-get-into-public-speaking-moran-weber) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)