Principal Machine Learning Engineer, Accelerated Apache Spark

NVIDIA Ltd.
Santa Clara, CA, United States
1 day ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$272,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Artificial Intelligence Algorithm Design Big Data C++ (Programming Language) Nvidia CUDA Computer Programming Data Centers Extract Transform Load (ETL) Python (Programming Language) Machine Learning NumPy
+17 more
Open Source Technology Tensorflow SciPy SQL Databases Reinforcement Learning Data Processing Graphics Processing Unit (GPU) Feature Engineering Pytorch Delivery Pipeline Large Language Models Apache Spark Pandas Scikit Learn Information Technology Xgboost Machine Learning Operations

Job description

NVIDIA is looking for a Machine Learning (ML) Engineer to join the GPU accelerated Apache Spark team. Apache Spark is the most popular data processing engine in data centers for running large scale workloads for ETL, SQL, and ML/DL model training and inference pipelines, spanning many domains and use cases. NVIDIA GPUs offer a promising avenue for significantly speeding up and/or lowering the cost of running Apache Spark applications at massive scales. You will work with the open source community to accelerate Apache Spark with GPUs. You will apply the latest ML/AI methods to empower enterprises to migrate Spark workloads onto GPUs at scale.

What You’ll Be Doing

  • Design and implement machine learning solutions for performance prediction and optimization of GPU accelerated enterprise Apache Spark workloads.
  • Develop advanced algorithms and adaptive systems to continuously improve the performance of Apache Spark workloads on GPUs.
  • Develop AI-based agents and tools to assist with fixing system issues and application optimization.
  • Collaborate with key partners and customers on the deployment of complex machine learning solutions in various environments.
  • Maintain deep domain expertise by knowing the latest published advances in ML systems and algorithms.
  • Provide technical mentorship and leadership in data science and machine learning to a team of engineers.

Requirements

  • BS, MS, or PhD or equivalent experience in Machine Learning, Data Science, Computer Science or a closely related field.
  • 12+ years of professional experience in designing, implementing, and productionizing high-quality ML/DL solutions.
  • 5+ experience as technical lead in ML model development.
  • Proven hands-on experience (2+ years) with large-scale data processing platforms, such as Apache Spark.
  • Proven ability to employ modern tooling and sound techniques for all aspects of crafting, deploying, and maintaining machine learning models.
  • Excellent programming skills in Python and Python data science related libraries like numpy, pandas, scikit-learn, scipy, pytorch, and tensorflow.
  • Deep experience with sophisticated ML methodologies, including LLM/GenAI, reinforcement learning, and adaptive, on-line ML systems.
  • Strong expertise in feature engineering, feature importance assessment, and developing boosted tree model solutions (e.g., XGBoost).

Ways To Stand Out From The Crowd

  • Understanding of the internal workings and architecture related to Apache Spark.
  • Familiarity with NVIDIA GPUs and CUDA.
  • Experience coding in Scala, Java, and/or C++.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most experienced and dedicated people in the world working for us. If you are passionate about what you do, creative and autonomous, we want to hear from you!

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:54 min

Development history of scientific computation libraries and PyViz tools

Radovan Kavický · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

1:33 min

Summary of machine learning capabilities and engineering opportunities

Jan Zawadzki · LIVE

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

Videos

See all

Related articles

See all