> Markdown version of [/jobs/ext/1977128-systems-design-engineer-ai-software](https://www.wearedevelopers.com/jobs/ext/1977128-systems-design-engineer-ai-software). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Systems Design Engineer (AI, Software) - **Company:** Advanced Micro Devices, Inc. - **Location:** San Jose, CA, United States - **Contract:** Permanent contract - **Skills:** Adobe Analytics, Board Bringup, Artificial Intelligence, C++ (Programming Language), Profiling, Computer Programming, Software Debugging, Data Flow Control, Python (Programming Language), Linux Kernel, Machine Learning, Performance Tuning, Software Engineering, Graphics Processing Unit (GPU), Pytorch, Parallel Computation, Linux Development, ONNX (Open Neural Network Exchange) Format, Machine Learning Operations, Software Version Control - **Published:** August 7, 2026 - **Apply:** https://us.experteer.com/career/view-jobs/systems-design-engineer-ai-software-san-jose-ca-usa-58831998 ## About the Role model accuracy, and hardware bring-up * Contribute to hardware-software co-design by evaluating architectural tradeoffs * Drive innovation in performance methodologies, benchmarking, tooling, and AI system optimization Tasks * Strong software development experience using C/C++ and Python * Experience with parallel programming and performance optimization * Knowledge of ML inference workloads and common operators (e.g., GEMM, convolution, attention, softmax) * Familiarity with AI frameworks/runtimes (PyTorch, ONNX Runtime, ROCm) * Understanding of computer architecture, memory hierarchies, and accelerator programming models * Experience developing software for GPUs, NPUs, or AI accelerators * Experience with Linux development, debugging, profiling, and source control tools * Familiarity with MLIR, LLVM, compiler technologies, or related stacks * Exposure to quantization techniques (INT8, FP8, FP16, BF16) * Knowledge of dataflow architectures, systolic arrays, or custom accelerators * aa programming patents, or demonstrated contributions in ML systems or computer architecture Key requirements * ## Description Experteer Overview In this role you develop and optimize ML workloads on AMD AI accelerators, bridging hardware and software. You design high-performance ML operator kernels and dataflow libraries, and work with cross-functional teams to bring AI tech from concept to production. You gain full-stack visibility from kernel development to silicon bring-up, contributing to industry-leading AI inference performance. This is a chance to impact products deployed in millions of devices worldwide and help shape accelerator technology. Compensation / Benefits * Develop and optimize ML operator kernels and dataflow libraries for AMD AI accelerators * Profile workloads to identify bottlenecks and drive system-level optimizations * Enable and validate ML models within production inference frameworks and runtimes * Collaborate with compiler, runtime, architecture, and silicon teams to deliver high-performance AI solutions * Debug and resolve issues across kernel implementation, runtime integration, model accuracy, and hardware bring-up * Contribute to hardware-software co-design by evaluating architectural tradeoffs * Drive innovation in performance methodologies, benchmarking, tooling, and AI system optimization Tasks * Strong software development experience using C/C++ and Python * Experience with parallel programming and performance optimization * Knowledge of ML inference workloads and common operators (e.g., GEMM, convolution, attention, softmax) * Familiarity with AI frameworks/runtimes (PyTorch, ONNX Runtime, ROCm) * Understanding of computer architecture, memory hierarchies, and accelerator programming models * Experience developing software for GPUs, NPUs, or AI accelerators * Experience with Linux development, debugging, profiling, and source control tools * Familiarity with MLIR, LLVM, compiler technologies, or related stacks * Exposure to quantization techniques (INT8, FP8, FP16, BF16) * Knowledge of dataflow architectures, systolic arrays, or custom accelerators * Publications, patents, or demonstrated contributions in ML systems or computer architecture Key requirements * ## Related Videos - [Developing an AI.SDK](https://www.wearedevelopers.com/videos/198-developing-an-ai-sdk) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Into the hive of eBPF!](https://www.wearedevelopers.com/videos/1199-into-the-hive-of-ebpf) - [Enhancing Workload Security in Kubernetes](https://www.wearedevelopers.com/videos/356-enhancing-workload-security-in-kubernetes) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)