Staff Software Engineer, Deep Learning Acceleration

Aurora Innovation, Inc.
Mountain View, CA, United States
5 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Big Data C++ (Programming Language) Nvidia CUDA Data Centers Linux Python (Programming Language) Performance Tuning Software Architecture Tensorflow Software Engineering High Performance Computing Pytorch
+6 more
Deep Learning Parallel Computation Information Technology Codebase Machine Learning Operations GPT

Job description

Experteer Overview In this role you drive the performance optimization of deep learning networks used in Aurora’s autonomous vehicle systems. You will analyze and optimize software architecture, latency, and deployment both on-vehicle and in data centers. You’ll collaborate with cross-functional teams to scale DL workloads, improving efficiency and reliability of self-driving technology. Your work helps make transportation safer and more accessible at scale. This is a fast-paced, impact-oriented role in a company solving complex, meaningful problems in mobility. Compensation / Benefits * Conduct performance analysis and optimization of Deep Learning networks running on the AV * Improve software architecture, system performance, and latency for deep learning applications * Deploy DL models on the AV and for large-scale data center training * Troubleshoot performance issues using profiling and roofline model techniques * Collaborate with cross-functional teams to enhance self-driving efficiency Tasks * 5+ years of software engineering experience * BS, MS, or PhD in Computer Science or related field * Proficiency in CUDA, C++, and Python * Experience in high-performance computing and parallel programming * Skill with performance analysis tools (NVIDIA Nsight Systems, Nsight Compute) and roofline model * Hands-on DL/ML framework experience (PyTorch, TensorFlow) for model deployment * Understanding of CV and transformer-based architectures * Strong analytical and troubleshooting skills * Ability to learn new technologies quickly in a fast-paced environment * Experience working with large codebases and cross-functional teams * Linux/Unix proficiency Key requirements * annual bonus * equity compensation * benefits * hybrid work environment * in-office 3 days per week

Requirements

codebases Tasks * 5+ years of software engineering experience * BS, MS, or PhD in Computer Science or related field * Proficiency in CUDA, C++, and Python * Experience in high-performance computing and parallel programming * Skill with performance analysis tools (NVIDIA Nsight Systems, Nsight Compute) and roofline model * Hands-on DL/ML framework experience (PyTorch, TensorFlow) for model deployment * Understanding of CV and transformer-based architectures * Strong analytical and troubleshooting skills * Ability to learn new technologies quickly in a fast-paced environment * Experience working with large codebases and cross-functional teams * Linux/Unix proficiency Key requirements * annual bonus * equity compensation * benefits * hybrid work environment * in-office 3 days per week

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

1:24 min

Comprehensive AI infrastructure stacks at the Linux Foundation

Matt White Matt White · WWC 2025

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · WWC 2025

Videos

See all

Related articles

See all