Software Principal Engineer- AI Compiler

Ampere Computing
Santa Clara, United States
about 1 month ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$195,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Artificial Neural Networks C++ (Programming Language) Data Structures Python (Programming Language) Pattern Recognition Pytorch Deep Learning Information Technology Hardware Acceleration

Job description

In this role, you will optimize computational graphs to unlock the full potential of Ampere’s deep learning accelerator. You’ll work across the full SW/HW stack - from inference serving and framework integration down to compiler, runtime, and compute kernels.

What You’ll Achieve:

  • Optimize deep learning computational graphs for performance, throughput, and latency on Ampere’s accelerator hardware
  • Enable popular models and frameworks (PyTorch, Llama.cpp) and serving platforms (vLLM, SGLang)
  • Identify and implement graph-level optimizations: op fusion, pattern recognition, redundancy elimination
  • Collaborate on HW/SW co-design to push computational efficiency
  • Work with cross-functional teams to integrate AI solutions into Ampere’s AI hardware platforms

Requirements

  • Bachelors degree in Computer Science, Mathematics or a related technical field & 8 years of related experience; or Master’s degree & 6 years
  • Strong CS fundamentals: algorithms, data structures, systems
  • Solid graph algorithm knowledge and reasoning ability
  • Proficiency in Python and C/C++
  • Demonstrated exceptional problem-solving ability - IOI medal, ACM ICPC medal, Codeforces Grandmaster, USACO Platinum, or equivalent competitive programming achievement is a big plus
  • Preferred to have traditional CPU or GPU compiler development background
  • Familiar with LLVM compiler framework as well as MLIR dialects is a plus
  • Fast learner who can pick up new domains quickly and maximize agent-assisted development
  • Familiarity with deep learning concepts and neural network architectures is a plus

Benefits & conditions

At Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $195,000 and $292,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

Benefit highlights include:

  • Premium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.
  • Unlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.
  • A variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.

And there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

LI-Hybrid

About the company

Ampere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.

As a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:49 min

Augmenting junior and principal engineering roles with AI

Neel Sundaresan Neel Sundaresan +1 · World Congress 2026 Europe

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen · Coffee With Developers

4:42 min

Building robust data structures with structs and bound functions

Rainer Stropek Rainer Stropek · World Congress 2021

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

2:17 min

Distinguishing between AI, machine learning, and deep learning

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all