Senior GenAI & High Performance Computing (HPC)...

Dell Technologies Inc.
Round Rock, TX, United States
about 2 months ago
Apply on www.juju.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$145,000.0 - $199,100.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Linux InfiniBand Node.Js Performance Tuning Dell PowerEdge Red Hat Enterprise Linux AI Infrastructure High Performance Computing Parallel Computation HybridCloud Kubernetes
+2 more
Slurm Docker

Job description

Dell Technologies has delivered HPC solutions for 25+ years, including support for Bright Cluster Manager (now NVIDIA BCM) since 2011. Today, Dell is NVIDIA’s preferred partner for GenAI Factory systems, using Dell GenAI PowerEdge XE servers and NVIDIA NVAIE to help customers build and scale end to end GenAI and High-Performance Computing environments.

Join us to do the best work of your career and make a profound social impact as a Senior GenAI & High Performance Computing (HPC) Delivery Engineer on our Service Delivery Team in Austin, Texas or Remote United States. 50-70 % National Travel.

What you’ll achieve

We’re seeking a Senior GenAI & HPC Engineer with deep experience in GPU accelerated systems, Linux performance tuning, and benchmarking. This role is highly hands on and customer facing, supporting onsite deployments across the U.S. for advanced HPC and GenAI solutions.

You will work as a part of a team to help build, integrate, and test some of the world’s largest multi GPU systems, benchmark them using industry standard tools, and deliver the next generations of AI and HPC infrastructure.

You will:

  • Deploy, configure, and validate GPU accelerated compute clusters for AI, ML, and HPC with NVIDIA Base Command Manager (Warewulf and OpenHPC knowledge are a plus)

  • Perform benchmarking with HPL GPU, HPL MxP, STREAM, NCCL, RCCL, OSU Microbenchmarks, and related tools

  • Produce as-built documentation, performance reports, and share best practices amongst the team.

Requirements

  • 7+ years with HPC or GenAI clusters, GPU based systems, AI infrastructure, or related fields

  • Deep hands on experience with GPU deployment, configuration, and multi-node testing using NVIDIA Base Command Manager

  • Proficiency with benchmarking tools: HPL, STREAM, NCCL, RCCL, MxP, OSU Microbenchmarks

  • Red Hat certification (RHCSA/RHCE) or 7+ years of relevant RH distros experience

  • Experience with GenAI/HPC networking (InfiniBand and/or RoCE)

  • Experience working in Linux based parallel computing environments at scale

  • Experience with containers/orchestration (Docker, Singularity/Apptainer, Kubernetes, Slurm)

  • Ability to travel up to 70% of the time across the U.S . as needed for projects

  • Strong customer facing and communication skills

Desirable Requirements

  • Bachelor’s degree

  • NVIDIA certifications (NCA, NCE, DGX)

  • Experience with NVIDIA UFM, Infiniband, and SpectrumX fabrics

  • Exposure to hybrid cloud or GPU cloud environments

  • Experience with GPU observability/performance profiling tools

Benefits & conditions

Dell is committed to fair and equitable compensation practices. The salary range for this position is $145,000 to $199,100.

Benefits and Perks of working at Dell Technologies

Your life. Your health. Supported by your benefits. You can explore the overall benefits experience that awaits you as a Dell Technologies team member - right now at MyWellatDell.com

About the company

Who We Are

We believe that each of us has the power to make an impact. That’s why we put our team members at the center of everything we do. If you’re looking for an opportunity to grow your career with some of the best minds and most advanced tech in the industry, we’re looking for you.

Dell Technologies is a unique family of businesses that helps individuals and organizations transform how they work, live and play. Join us to build a future that works for everyone because Progress Takes All of Us.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

Videos

See all

Related articles

See all