> Markdown version of [/jobs/ext/2717596-systems-administrator-gpu-ai-infrastructure](https://www.wearedevelopers.com/jobs/ext/2717596-systems-administrator-gpu-ai-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Systems Administrator, GPU/AI Infrastructure - **Company:** Lenovo - **Location:** Morrisville, NC, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Nvidia CUDA, Data Center Infrastructure Management (CIM), InfiniBand, Role-Based Access Control, Management of Software Versions, AI Infrastructure, Multi-Agent Systems, HybridCloud, Kubernetes, Machine Learning Operations, Hardware Infrastructure, Docker, Service Stack - **Published:** September 4, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/88231234/1 ## About the Role * 2+ years of systems administration experience with hands-on breadth across network, compute, storage, and operating system layers, plus a GPU/AI infrastructure specialization. * Production Kubernetes Administration: demonstrated experience with cluster operations, RBAC, cluster networking, and version upgrades in a production environment. * Docker: hands-on experience administering Docker container runtime environments. * Hands-on NVIDIA administration: demonstrated experience administering NVIDIA drivers and the CUDA toolkit. * InfiniBand Fabric: working familiarity with InfiniBand fabric in a multi-node GPU environment. * GPU Orchestration: experience with Run:AI, ClearML, or a comparable GPU orchestration and MLOps toolset. * Physical Infrastructure: comfortable with hands-on hardware work, including rack-and-stack, structured cabling, and smart hands, in addition to higher-level administration. Core Professional Skills * Incident Response: able to own after-hours incident response for a shared, multi-team infrastructure environment. * Cross-Team Coordination: comfortable supporting multiple HC/AI teams across a shared lab environment with competing priorities, and coordinating day to day with the Network Engineer and Junior Systems Administrator roles on adjacent, overlapping infrastructure., * Kubernetes Certification: Certified Kubernetes Administrator (CKA) or equivalent. * Multi-Tenant Lab Experience: experience supporting a multi-tenant shared lab environment serving distributed teams. * Familiarity with NVIDIA certification and reference architecture (Lenovo Validated Design) processes. * Experience with DCIM, IPAM, or monitoring tooling comparable to the lab's stack (Hyperview-class DCIM, BlueCat / Infoblox-class IPAM, NVIDIA DCGM monitoring). ## Description Lenovo seeks a Systems Administrator to own end-to-end deployment, support, and administration of the consolidated Morrisville AI lab, across the network, compute, storage, operating system, and container orchestration (Docker and Kubernetes) layers, plus the NVIDIA technology stack across 50-plus servers. This is a hands-on infrastructure role: the same person who designs and administers these layers also does the day-to-day admin, support, and engineering work in the lab, including rack-and-stack and smart hands. You are the only dedicated support resource for the Lab Operations Manager, and you keep the lab's FY26/27 Net CapEx investment operational at the scale the AI Lab Consolidation requires. You will work alongside the lab's Network Engineer and Junior Systems Administrator roles, who own adjacent, more specialized slices of network engineering and routine rack-and-stack work respectively; you provide the full-stack administration that connects their work end to end. You will also provide secondary technical support for capital equipment hosted in Bangalore and validate lab configurations against NVIDIA reference architecture certification standards. The role sits within Lenovo's Hybrid Cloud and AI Infrastructure Services (HCAIS) Lab Operations organization. Key Responsibilities Infrastructure Deployment and Administration Network: deploy, support, and administer lab networking, in coordination with the dedicated Network Engineer role for switch and fabric configuration. Compute: deploy, support, and administer physical and virtual compute infrastructure across the lab. Storage: deploy, support, and administer lab storage systems across block and file tiers. Operating Systems: install, patch, and administer operating systems across lab infrastructure. Container Orchestration Docker: deploy and administer Docker container runtime environments. Kubernetes: own full production Kubernetes administration, including cluster operations, role-based access control (RBAC), cluster networking, and version upgrades. NVIDIA Platform Administration NVIDIA Technology Stack: administer NVIDIA drivers, the CUDA toolkit, InfiniBand fabric, Run:AI orchestration, and ClearML across the lab's GPU infrastructure. Configuration Validation: validate lab configurations against NVIDIA reference architecture (Lenovo Validated Design) certification requirements. Lab Operations and Smart Hands Day-to-Day Support: provide day-to-day administration, support, and engineering functions for the network, compute, and storage layers in the physical lab. Rack and Stack: perform rack-and-stack, structured cabling, and smart-hands support for lab hardware alongside higher-level administration duties. Cross-Site Support and Coordination Bangalore Secondary Support: provide secondary technical support for NVIDIA technology stack capital equipment hosted in Bangalore. Lab Coordination: coordinate with the adjacent ISG AI COE lab (Tech Marketing, customer proof-of-concepts) and the Morrisville Executive Briefing Center given shared physical proximity. NVIDIA Relationship: maintain a direct working relationship with NVIDIA field contacts based in Morrisville. ## Related Videos - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [A Deep Dive on How To Leverage the NVIDIA GB200 for Ultra-Fast Training and Inference on Kubernetes](https://www.wearedevelopers.com/videos/1625-a-deep-dive-on-how-to-leverage-the-nvidia-gb200-for-ultra-fast-training-and-inference-on-kubernetes) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [Microservices: how to get started with Spring Boot and Kubernetes](https://www.wearedevelopers.com/videos/242-microservices-how-to-get-started-with-spring-boot-and-kubernetes) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)