Principal AI Compiler Engineer

NXP
San Diego, CA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Artificial Neural Networks C++ (Programming Language) Microprocessors Hardware Design Python (Programming Language) Performance Tuning Scrum Methodology Tensorflow Pytorch Generative AI
+4 more
Information Technology Low Latency ONNX (Open Neural Network Exchange) Format Data Analytics

Job description

NXP is searching for a hands-on AI Compiler Engineer who thrives at the convergence of cutting-edge AI, compiler tech, and hardware design. Here, you’ll not only architect and scale a production-class AI compiler toolchain, but also rethink how AI automates, optimizes, and accelerates every step of building and deploying neural networks on NXP’s SoCs. You’ll work shoulder-to-shoulder with visionary engineers-both human and AI-enabling adaptive compilers that learn, evolve, and redefine what’s possible for embedded intelligence. With a relentless focus on hardware-software co-design, you’ll collaborate across teams to translate high-level AI models into blazing-fast, energy-efficient executables, unlocking the full potential of our silicon for real-world impact. Innovation here isn’t a catchphrase-it’s your everyday., * Own the design, implementation, and evolution of an AI compiler toolchain that leverages AI agents to seamlessly map neural networks onto NXP’s SoC platforms.

  • Pioneer new graph transformations, lowering, scheduling, and codegen strategies for CPUs and custom accelerators, driven by insights from AI-powered analytics.
  • Build deep integrations with leading AI frameworks (PyTorch, TensorFlow, ONNX, and more), using AI agents to rapidly onboard new model architectures and ops.
  • Push the envelope on quantization, operator fusion, memory planning, and layout transformations-combining human expertise and AI-guided design for state-of-the-art results.
  • Partner with hardware and software architects, kernel hackers, and AI agents to co-design next-gen compiler and accelerator features, aligning silicon and code for maximum impact.
  • Diagnose and crush performance bottlenecks with AI-enabled profiling and diagnostics, relentlessly tuning for latency, throughput, and power efficiency.
  • Level up validation, benchmarking, and regression pipelines by harnessing AI agents-ensuring compiler correctness and world-class performance, release after release.
  • Uplevel the developer experience by streamlining usability, diagnostics, and documentation-AI agents are your copilots for user support, troubleshooting, and rapid iteration.

Requirements

  • MS/PhD (or equivalent experience) in Computer Science, EE, or related field
  • Deep experience building AI compilers, accelerator backends, or graph optimization frameworks.
  • Strong expertise in graph optimization and performance optimization for NPUs or custom accelerators
  • Experience with MLIR, LLVM, TVM-like systems, or proprietary compiler IRs.
  • Excellent C/C++ and Python skills
  • Solid understanding of AI inference workloads (CNNs, transformers, perception or generative models)
  • Strong communication skills are required, e.g. agile development experience in Scrum team (Product Owner or Scrum Master)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

3:53 min

Architecting machine learning projects with the PAI platform

Qiyang Duan · LIVE

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

Videos

See all

Related articles

See all