Senior AI SoC Modeling Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon.com, Inc.
Austin, TX, United States
24 days ago
Apply on www.amazon.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$193,300.0 - $261,500.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services C++ (Programming Language) Code Review Computer Programming Microarchitecture Software Debugging Microprocessors Hardware Design Machine Learning Distributed Simulation Software Engineering
+8 more
SystemC Multithreading Graphics Processing Unit (GPU) Application Specific Integrated Circuits Model Validation Build Process Software Coding Software Version Control

Job description

Develop and maintain high-fidelity functional model of AI/ML accelerator and its SoC subsystems, including compute engines, memory hierarchies, on-chip interconnects, and data paths - translating architecture specs and RTL behavior into accurate, testable C++ models

  • Validate model behavior against RTL simulations, emulation platforms, or silicon measurements; debug discrepancies and drive model-to-RTL correlation to high fidelity
  • Partner with design verification teams to integrate models into pre-silicon validation environments and catch architectural bugs early in the design cycle
  • Collaborate with architects/micro-architects, RTL design engineers, ML SW engineers, and compiler engineers to evaluate architecture and microarchitecture tradeoffs and help make hardware design decisions
  • Contribute to cycle-approximate performance model effort enabling architectural exploration ahead of RTL availability, early software development
  • Quantify system-level tradeoffs across compute, memory bandwidth, networking, and storage to influence reference architectures and long-term silicon strategy
  • Build and improve modeling infrastructure: simulation frameworks, regression suites, automated correlation checks, and coverage-driven validation flows
  • Develop modeling methodologies and tools that scale across multiple IP blocks and SoC generations, improving team efficiency and model reuse

Why this role is interesting:

  • Your models are used to verify silicon before it’s built - bugs you catch save months of schedule and millions of dollars
  • You’ll work at the intersection of software engineering and chip design, with deep visibility into how custom ML accelerators are architected
  • As the team scales, there’s a clear path into architectural modeling - using your models to influence chip design decisions, not just validate them
  • Small team, high ownership, direct impact on AWS’s most strategic silicon programs

Requirements

Have built functional or performance models of SoCs, ASICs, GPUs, CPUs, or IP blocks

  • Are comfortable working with architectural / design specifications or reference implementations and translating them into C++ or SystemC models
  • Understand verification concepts and have worked with DV teams or in pre-silicon validation environments
  • Care about model fidelity and have experience correlating models against RTL or silicon
  • Are interested in expanding into architectural performance modeling as the team grows
  • Enjoy working on a small, high-impact team where you own significant pieces of the stack

No ML background needed. You’ll learn the ML accelerator domain on the job., 6+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience

  • Experience as a mentor, tech lead or leading an engineering team
  • 6+ years writing functional or performance models of hardware (SoCs, ASICs, GPUs, CPUs, IP blocks)
  • Experience programming in C++, using advanced language features
  • Knowledge of SoC, CPU, GPU, and/or ASIC architecture and micro-architecture

Preferred Qualifications

  • Experience working with DV teams or integrating models into verification flows
  • Experience correlating functional models against RTL simulation, emulation, or silicon
  • Experience developing and calibrating performance models for custom silicon
  • Experience with SystemC, TLM, or cycle-approximate modeling methodologies
  • Experience building regression and CI frameworks for model validation
  • Familiarity with Modern C++ (20 and beyond)
  • Experience with multi-threaded or distributed simulation
  • ML accelerator architecture knowledge (a plus, not required)

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, CA, Cupertino - 193,300.00 - 261,500.00 USD annually USA, TX, Austin - 168,100.00 - 227,400.00 USD annually

About the company

AWS’s Trainium and Inferentia chips power the world’s largest machine learning clusters. Our team builds C++ models of these custom SoCs that RTL designers, verification engineers, and software teams depend on throughout the silicon development lifecycle. We’re looking for a modeling engineer to build and own models that directly impact how our chips are designed, verified, and brought to production.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.amazon.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:59 min

Generating a Swagger JSON file during the build process

Roman Alexis Anastasini · World Congress 2021

3:39 min

Addressing code review surrender and process exploitation

Laura Tacho Laura Tacho · World Congress 2026 Europe

3:18 min

Hardware architectures tailored for specific artificial intelligence computations

Stephan Gillich Stephan Gillich · World Congress 2024

1:48 min

Optimizing container images using multi-step build processes

Adrian Kosmaczewski · LIVE

56 sec

The hidden costs of delayed peer code reviews

Tim Gilboy Tim Gilboy

2:42 min

Dissecting artificial intelligence layers from compute to applications

Christian Nagel Christian Nagel +3 · World Congress 2026 Europe

Videos

See all

Related articles

See all