Senior Software Development Engineer

Advanced Micro Devices, Inc.
San Jose, CA, United States
9 days ago
Apply on diversityjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$145,600.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence C++ (Programming Language) Code Generation Profiling Nvidia CUDA Software Debugging General-Purpose Computing on Graphics Processing Units Linux Kernel Machine Learning OpenCL Performance Tuning Tensorflow
+9 more
Software Engineering System Programming Graphics Processing Unit (GPU) High Performance Computing Pytorch Parallel Computation Gpu Programming Information Technology ONNX (Open Neural Network Exchange) Format

Job description

AMD is seeking a Senior Software Development Engineer to develop and optimize software for next-generation AI and GPU computing platforms. You will work across GPU kernel development, performance optimization, AI frameworks, runtime systems, and compiler technologies to deliver industry-leading performance for AI and high-performance computing workloads. THE PERSON

The ideal candidate has strong experience in GPU programming, performance analysis, and systems software development. You understand how AI workloads execute on modern GPU hardware and can identify performance bottlenecks across kernels, runtime, libraries, and compiler-generated code. Experience with MLIR/LLVM-based compiler flows is valuable but not the primary focus of the role. KEY RESPONSIBILITIES

  • Develop and optimize GPU kernels for AI and compute workloads.
  • Analyze and improve GPU performance, including compute utilization, memory bandwidth, cache efficiency, and kernel execution.
  • Collaborate with AI framework, runtime, library, and hardware teams to improve end-to-end workload performance.
  • Profile, benchmark, and debug large-scale AI workloads on AMD GPU platforms.
  • Support compiler-driven optimizations and code generation using LLVM and MLIR-based technologies.
  • Improve performance analysis, testing, benchmarking, and automation infrastructure.
  • Enable and optimize new GPU hardware capabilities through software innovation.
  • Apply AI-assisted development tools to improve productivity, debugging, and optimization workflows., AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

Requirements

  • Strong C/C++ development skills.
  • Experience with HIP, CUDA, OpenCL, or other GPU programming models.
  • Strong understanding of GPU architecture, parallel programming, memory hierarchy, synchronization, and performance optimization.
  • Hands-on experience profiling and optimizing AI, machine learning, HPC, or graphics workloads.
  • Experience with performance analysis tools and benchmarking methodologies.
  • Knowledge of LLVM, MLIR, compiler optimization, or code generation techniques.
  • Experience with AI frameworks such as PyTorch, ONNX Runtime, TensorFlow, or similar ecosystems.
  • Strong problem-solving and debugging skills across software and hardware stacks., * Bachelor’s or Master’s degree in Computer/Software Engineering, Computer Science, or related technical discipline

Benefits & conditions

$145,600.00/Yr.-$218,400.00/Yr.

About the company

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger- technologythat moves the world forward.Join us and, together, we’ll advance your career.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on diversityjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

1:42 min

Navigating emerging hardware standardization in vendor programming ecosystems

Paul Graham Paul Graham · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 · World Congress 2026 Europe

2:17 min

Comparing code profiling with surface level monitoring

Jérôme Vieilledent · LIVE

Videos

See all

Related articles

See all