> Markdown version of [/jobs/ext/2805137-infrastructure-ops-specialist](https://www.wearedevelopers.com/jobs/ext/2805137-infrastructure-ops-specialist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Ops Specialist - **Company:** Massed Compute, Inc. - **Location:** United States - **Experience:** Experienced - **Contract:** Temporary contract - **Skills:** Artificial Intelligence, Spreadsheets, Data Centers, Data Center Infrastructure Management (CIM), Inventory Management Software, Hardware Infrastructure - **Published:** September 9, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=07bfea08818c6857 ## About the Role * Experience in supply chain management, analytics, or logistics, and a track record of tracking physical assets accurately at scale, and specific opinions about where tracking breaks down * Working fluency in data center fundamentals - racks, power, cooling, cabling, density - enough to make placement decisions and be credible with site teams * Experience coordinating work across organizations that do not report to you, including vendors who have their own priorities * The instinct to build a repeatable process the second time you do something, rather than the tenth * Strong personal organization and comfort holding many parallel threads without dropping any, and a system for it that is not memory * Strong written communication, especially in making status legible to people who do not want a status meeting * High ownership and follow-through; you stay with a deployment through acceptance rather than handing it off at delivery * Fluency with AI tools and excitement about using them to automate the repetitive parts of operations work Desired background (we encourage you to apply even if you don't check every box) * Typically 3-8+ years of professional experience in data center operations, infrastructure deployment, hardware or IT asset management, supply chain and logistics operations, technical program management, or a similarly high-responsibility role * Experience at a high-growth technology company, rapidly scaling startup, infrastructure operator, or similarly demanding environment * Hands-on experience with hardware deployment lifecycles - receiving, installation, commissioning, failure, RMA, redeployment, decommissioning * Experience implementing or materially rebuilding an asset management, inventory, or DCIM system that other teams then relied on * A bachelor's degree or equivalent combination of education, training, and experience Bonus points if you… * Have worked with GPU infrastructure specifically, and understand what is different about it - density, cooling, failure modes, RMA timelines, and how long a dead node actually costs you * Have run deployments across multiple sites or colocation providers simultaneously * Have managed warranty and RMA processes at volume with OEMs, integrators, or resellers * Have built lightweight internal tooling or automation yourself to make an operations process work without more headcount * Are fluent with AI tools such as Claude, Claude Code, and Codex, and related workflow automation tools ## Description Massed Compute is looking for an Infrastructure Operations Specialist to make sure every GPU we buy ends up in the right place, in service, and earning - and that we always know where all of it is You'll be the connective tissue between procurement, data center leasing, site operations, and solutions engineering. Those four functions each hold part of the picture: what we've bought, where we have space and power, what is actually racked and running, and what we've promised customers. That means deciding where new equipment goes across our site portfolio, tracking every shipment from purchase order to production, running deployments and redeployments, driving warranty and RMA work to resolution, and building the process that tells us what state our infrastructure is in now and what it will be in ninety days. This role involves extensive logistics and coordination that will directly drive utilization and operational efficiency, and transparency that informs decisions and prioritization for multiple teams across the company. Additionally, you'll define what has to be tracked and how, and partner with engineering and use coding agents to build out the tooling behind the monitoring and analytics. This is a builder role. We're looking for someone who can both run today's deployments personally and build the tracking, process, and vendor discipline that let us run ten times as many without ten times the people., In your first 8-12 months, you will have built a single, trusted view of what hardware we own and what state it is in, cut the time between equipment arriving and equipment earning, established a warranty and RMA process that resolves failures without anyone chasing them, and given procurement, capacity, and the commercial team a forward view of the fleet they can plan against. Critically, you will have replaced a picture that lives in a handful of people's heads and a dozen spreadsheets with a system that stays true on its own - so that "where is it, what state is it in, and when is it available" stops being a question anyone has to ask around to answer. Responsibilities * Decide where equipment goes. Own placement across our site portfolio, balancing power and cooling headroom, rack density, customer adjacency, deployment speed, and what each site should be used for over the next year. * Track everything from PO to production. Maintain the chain of custody for every shipment - ordered, in transit, received, staged, racked, cabled, burned in, accepted - and know at any moment what is where and what is late. * Run deployments and redeployments. Coordinate the sequence across facilities, networking, storage, and platform teams, and move capacity between sites when demand or economics change. * Own warranty, RMA, and servicing. Drive hardware failures through vendor and depot processes to resolution, track what we are owed, and hold vendors to their commitments. * Own the fleet system of record. Select, implement, and maintain the asset and inventory system that tells us what we own, where it is, and what state it is in - including the reconciliation discipline that keeps it accurate * Build the forward view. Design the reporting that shows what is landing when, what capacity frees up when, and where we are at risk so leadership and the commercial team can plan against reality rather than intent * Support solutions engineering. Be the source of truth on what we can credibly commit to a customer and by when, and flag early when a promise is at risk * Manage vendors and remote hands. Own the relationships, SLAs, and escalation paths with the integrators, remote-hands providers, and logistics partners we depend on * Close the loop on failures. When a deployment slips or hardware goes missing, find the actual cause and change the process, rather than absorbing it and moving on * Bring leverage through AI. Use AI tools aggressively to automate tracking, reporting, vendor follow-up, and documentation - this role has more repeatable, automatable surface area than almost any other on the team ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Launching a marketplace on-time: A lesson in taking shortcuts using spreadsheets!](https://www.wearedevelopers.com/videos/477-launching-a-marketplace-on-time-a-lesson-in-taking-shortcuts-using-spreadsheets) - [The Sustainability Race: AI's Promises, Pitfalls and Potential](https://www.wearedevelopers.com/videos/100155-the-sustainability-race-ai-s-promises-pitfalls-and-potential) - [Edit Your Future: Queerverse Radical AI](https://www.wearedevelopers.com/videos/909-edit-your-future-queerverse-radical-ai) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [Building the Nervous System of AI - Michael Kagan (NVIDIA)](https://www.wearedevelopers.com/videos/2133-building-the-nervous-system-of-ai-michael-kagan-nvidia) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [6 Emerging Technologies We’ll Learn About in 2025](https://www.wearedevelopers.com/magazine/381-6-emerging-technologies-we-ll-learn-about-in-2025) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025)