> Markdown version of [/jobs/ext/2035530-machine-learning-ai-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2035530-machine-learning-ai-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Machine Learning & AI Infrastructure Engineer - **Company:** Kforce Inc. - **Location:** Austin, TX, United States (Remote available) - **Salary:** $175,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Bash Shell, Big Data, Ubuntu (Operating System), File Systems, InfiniBand, Python (Programming Language), Linux System Administration, Machine Learning, NetApp Applications, Data ONTAP (Server Appliance), Remote Direct Memory Access, AI Infrastructure, Scripting, High Performance Computing, Containerization, AI Platforms, Kubernetes, Storage Technologies, Machine Learning Operations, Isilon - **Published:** August 12, 2026 - **Apply:** https://www.dice.com/job-detail/3e11b289-347c-4098-910f-eca5137e0dbe ## About the Role * Strong experience administering High Performance Computing (HPC) environments * Experience with AI cluster administration and infrastructure operations * Hands-on Kubernetes administration experience in on-premises environments * Experience provisioning and managing PV/PVC storage through Kubernetes CSI drivers * Strong Linux administration skills, specifically Ubuntu * Scripting experience with Bash and/or Python * Proven troubleshooting and operational support experience * Ability to manage and maintain production infrastructure environments Experience with one or more of the following: * Dell PowerScale/Isilon * VAST Storage * NetApp ONTAP * DDN IntelliFlash * DDN Exascaler * Lustre Parallel File Systems, * NVIDIA ecosystem experience * NVIDIA Base Command Manager (BCM) * Bright Cluster Manager * MLOps platform exposure * Containerization technologies and orchestration platforms * High-performance networking experience * RDMA technologies * InfiniBand networking * NVIDIA UFM * Parallel file system administration * Storage Technologies (highly desired) ## Description Kforce has a client in Austin, TX that is seeking a Machine Learning & AI Infrastructure Engineer. This is not a traditional AI Engineer or Data Scientist role. The hiring team is specifically seeking a unique blend of: HPC Administrator + Kubernetes Administrator + AI Infrastructure Operations Engineer. Candidates who have owned, operated, supported, and troubleshot production AI or HPC environments will be the strongest fit. Experience administering and maintaining systems is significantly more important than architecture-only experience., * Administer and support AI and HPC cluster environments * Manage day-to-day operations of large-scale compute infrastructure * Deploy, maintain, and troubleshoot Kubernetes-based platforms * Ensure reliability, performance, and scalability across compute, storage, and networking environments * Support AI model training and inference infrastructure * Automate operational processes through scripting and tooling * Partner with engineering teams and customers to optimize platform performance * Troubleshoot complex infrastructure, networking, storage, and containerization issues * Support both internal platforms and customer-facing environments, * AI Infrastructure * Machine Learning Platforms * HPC Operations * Research Computing * Biotechnology * Academic Medical Centers * Digital Biology * Financial Services AI Platforms * Automotive AI Initiatives * Large-Scale Data Science Environments ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Intermediate Bitcoin Script](https://www.wearedevelopers.com/videos/25-intermediate-bitcoin-script) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)