HPC & AI Performance Engineer

Hewlett-Packard Enterprise
Houston, TX, United States
3 days ago
Apply on hpe.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$105,500.0 - $243,000.0
Working hours
Regular working hours

Tech stack

C (Programming Language) Artificial Intelligence C++ (Programming Language) Compilers Profiling Nvidia CUDA Software Debugging Linux Microprocessors Fortran (Programming Language) Python (Programming Language) OpenMP
+10 more
Performance Tuning Shell Script System Software Graphics Processing Unit (GPU) Large Language Models Distributed Programming Parallel Computation Information Technology Deepseek (AI Talent Sourcing Platform) Engineering Base

Job description

This role has been designed as ‘Hybrid’ with a requirement that you will work on average 2 days per week from an HPE office., * Lead HPC and AI benchmarking projects across CPU, GPU, network, memory, and storage platforms.

  • Evaluate HPE and competitive HPC and AI architectures using performance models and benchmark data.
  • Run and analyze scientific, engineering, large language model (LLM), AI training and inference, storage, and I/O workloads.
  • Identify performance bottlenecks and optimize applications, AI frameworks, libraries, compilers, runtimes, and system software.

  • Apply parallel computing, GPU acceleration, profiling, and memory and I/O optimization to deliver credible, repeatable results.

Requirements

  • 6+ years of Experience with HPC and AI workloads, including scientific and engineering applications, MLPerf, large language models such as DeepSeek, Kimi K2.6 or K3, and gpt-oss-120b, AI training and inference, storage benchmarks, and performance-profiling tools. Relevant coursework and internship experience will be considered.
  • Knowledge of HPC and AI system architecture, including CPUs, GPUs and other accelerators, memory, networking, storage, and software stacks, with the ability to explain their impact on application and benchmark performance.
  • Understanding of parallel and distributed programming techniques, including MPI, OpenMP, OpenSHMEM, algorithms, and performance considerations for HPC and AI workloads.
  • Ability to lead complex HPC and AI performance projects, work effectively across technical teams, and translate findings into customer-focused recommendations.
  • Demonstrated ability to analyze and optimize computational applications and benchmarks on Linux-based HPC and AI systems.
  • Experience with HPC and AI software environments, including C, C++, Fortran, Python, Linux scripting, compilers, MPI, MPI-IO, OpenMP, and relevant AI frameworks and libraries.
  • Experience with NVIDIA GPUs and AMD Instinct MI-series accelerators, including offloading computational routines using CUDA, HIP, OpenMP target offload, OpenACC, or comparable programming models for HPC and AI workloads.
  • Experience using performance-profiling, tracing, and debugging tools on Linux-based HPC and AI systems.
  • Ability to interpret benchmark results, identify performance bottlenecks, and recommend improvements for HPC and AI applications and systems.
  • Excellent analytical and problem-solving skills.
  • Excellent written and verbal communication skills, with the ability to present complex technical findings clearly to engineering teams, customers, and business stakeholders; professional proficiency in English required.
  • Master’s degree in computer science, engineering, mathematics, physics, chemistry, environmental science, or a related technical field; PhD preferred.

Benefits & conditions

“The expected salary/wage range for this position is provided below. Actual offer may vary from this range based upon geographic location, work experience, education/training, and/or skill level.

  • United States of America: Annual Salary USD 105,500 - 243,000 in Minnesota & Texas The listed salary range reflects base salary. Variable incentives may also be offered.”

About the company

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today’s complex world. Our culture thrives on finding new and better ways to accelerate what’s next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE., HPE operates at the forefront of a fast-moving HPC and AI industry. We are looking for driven engineers who thrive on complex technical challenges and consistently deliver high-quality results that strengthen the performance, competitiveness, and success of HPE’s HPC & AI solutions for customers worldwide.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on hpe.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:51 min

Evolution of custom compilers and virtual machines

Florian Rappl · LIVE

4:37 min

Simplifying parallel programming with the CUDA ecosystem

Paul Graham Paul Graham · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

3:00 min

Accelerating machine learning research with optimized compilers

Tanmay Bakshi · LIVE

3:37 min

Discovering hardware prize uses for localized artificial intelligence

Chris Heilmann Chris Heilmann +5 · LIVE

Videos

See all

Related articles

See all