> Markdown version of [/jobs/ext/1917956-ai-and-machine-learning-engineer](https://www.wearedevelopers.com/jobs/ext/1917956-ai-and-machine-learning-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI and Machine Learning Engineer - **Company:** Hewlett-Packard Enterprise - **Location:** Spring, TX, United States - **Experience:** Expert - **Salary:** $120,500.0 - $276,500.0 - **Contract:** Permanent contract - **Skills:** C (Programming Language), Artificial Intelligence, Artificial Neural Networks, C++ (Programming Language), Computer Programming, Continuous Integration, Information Engineering, File Systems, InfiniBand, Python (Programming Language), Machine Learning, Software Engineering, Weka, Scripting, NFT, Large Language Models, Deep Learning, Scalability Testing, Information Technology, Slurm, Machine Learning Operations, Network Server - **Published:** August 4, 2026 - **Apply:** https://hpe.wd5.myworkdayjobs.com/Jobsathpe/job/Spring-Texas-United-States-of-America/AI-and-Machine-Learning-Engineer_1209939-2 ## About the Role * Master's degree or PhD in Computer Science, Engineering, Information Technology or Systems, or relevant field. * 5+ years of experience. Knowledge and Skills: * 5+ years of experience in Machine Learning/Artificial Intelligence and 5+ years of experience in HPC * Experience running NCCL, HPL and AI benchmarks. * Experience working with containers and distributed deep learning and neural networks, to include transformers used in generative AI projects * Experience working with High Performance Computer Servers, High Performance Networking, and associated software, including resource managers like Slurm * Experience working with Weka I/O, NFTS and Lustre File Systems * Programming experience in Python, C, C++ * Strong analytical and critical thinking skills * Scripting, process automation and CI/CD are strongly desired * Must be a self-starter and be able to work with minimum supervision in a semi-remote setting Additional Skills: Artificial Intelligence Technologies, Cross Domain Knowledge, Data Engineering, Data Science, Design Thinking, Development Fundamentals, Full Stack Development, IT Performance, Machine Learning Operations, Scalability Testing, Security-First Mindset. ## Description This role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office., * Installs and configures complex IT infrastructure components (servers, storage, network) * Develop software scripts and configurations for automating deployment. * Study and improve the performance of Large Language Models run on HPE GPU servers * Performs system level analysis of server workloads on various HPE platforms running DL and ML code to include accelerated hardware and high-speed networks like InfiniBand * Writes white papers and other guidance documents for AI workload and model selection * Captures and reviews system performance data, logs, traces to understand workload behaviour * Develops software and scripts that help analyse AI workload performance data * Communicates technical work well and can provide summaries of work to non-technical colleagues * Works with software and hardware partners in optimizing systems and resolving performance issues * Documents and reports issues discovered when testing and evaluating the systems * Communicates project status and concerns to management in a timely manner * Provides guidance to less-experienced staff members. * Runs AI and HPC benchmarks. ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [Transforming Newspaper Readers into NFT holders](https://www.wearedevelopers.com/videos/1162-transforming-newspaper-readers-into-nft-holders) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [Blockchain, NFT and smart contracts for my application](https://www.wearedevelopers.com/videos/877-blockchain-nft-and-smart-contracts-for-my-application) - [Get Started With Blockchain For Your Business](https://www.wearedevelopers.com/videos/483-get-started-with-blockchain-for-your-business) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)