Director of Cloud, HPC & Sovereign AI Customer Engineering

Advanced Micro Devices, Inc.
Santa Clara, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Cloud Computing Data Centers Software Debugging AI Infrastructure High Performance Computing AI Platforms Information Technology

Job description

At AMD, we push the boundaries of innovation to solve the world’s most important challenges. As the Director of Cloud, HPC & Sovereign AI Customer Engineering within the Compute & Enterprise AI Solutions Customer Engineering organization, you will lead the team responsible for enabling successful deployment, adoption, and lifecycle support of AMD compute and AI solutions across Cloud, HPC, and Sovereign AI customers.

This highly visible leadership role oversees a team of Customer Program Managers (CPMs) responsible for guiding customers through new product introductions, deployment readiness, production ramp, and long-term fleet sustainment. Working closely with customers, Customer Platform Engineering, Product Management, Engineering and Architecture you will drive successful customer outcomes across the full lifecycle of AMD compute and AI platforms., * Lead and scale the Cloud, HPC & Sovereign AI Customer Engineering organization supporting strategic customer deployments and lifecycle management.

  • Lead a team of Customer Program Managers responsible for customer engagement, deployment readiness, new product introduction (NPI) execution, production ramp, and fleet sustainment activities.
  • Drive successful deployment, adoption, and operational readiness of AMD compute and AI solutions across Cloud, HPC, and Sovereign AI customers.
  • Serve as the executive escalation point for strategic customer issues and drive resolution of complex deployment, platform, performance, and operational challenges.
  • Partner closely with Customer Platform Engineering teams, including PAE, BAE, Security Engineering, and Debug Engineering, to ensure successful customer outcomes.
  • Develop deployment methodologies, operational best practices, and customer engagement frameworks that accelerate customer time-to-production.
  • Drive fleet sustainment strategies including observability, telemetry, remote diagnostics, lifecycle management, and operational readiness.
  • Partner with customers, and AMD engineering teams to support successful platform deployment and long-term fleet success.
  • Act as the voice of the customer, ensuring customer deployment experiences and operational insights influence future products, platforms, and solutions.
  • Mentor and develop CPM leaders while building a high-performance, customer-focused culture., AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

Requirements

The ideal candidate is a strong technical and organizational leader with experience supporting large-scale cloud, AI, HPC, or datacenter deployments. You have a proven track record of leading customer-facing technical teams, managing complex customer engagements, and driving successful deployment and sustainment of infrastructure at scale.

You are equally comfortable engaging with customer executives, architects, operations teams, and engineering organizations while driving alignment across AMD. You possess strong technical credibility, customer advocacy skills, and the ability to lead through influence in complex environments., * Experience leading customer-facing engineering, technical program management, cloud infrastructure, AI infrastructure, HPC, datacenter operations, or related technical organizations.

  • Experience leading technical teams responsible for customer deployments, operational readiness, and lifecycle support of complex infrastructure platforms.
  • Experience supporting large-scale cloud, AI, HPC, or enterprise infrastructure deployments from new product introduction through fleet sustainment.
  • Strong understanding of datacenter infrastructure, compute platforms, AI systems, observability, telemetry, and operational readiness.
  • Experience managing customer escalations and driving resolution of complex technical and operational issues.
  • Experience working with hyperscalers, cloud providers, HPC customers, sovereign AI customers.
  • Proven ability to drive alignment across engineering, product, operations, and customer-facing organizations.
  • Strong communication and executive engagement skills.

ACADEMIC CREDENTIALS:

Bachelor’s or Master’s degree in Engineering, Computer Science, or a related technical field.

About the company

At AMD, our mission is to build great products that accelerate next-generation computing experiences-from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges-striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 ¡ LIVE

2:36 min

Choosing between managed AI platforms and custom governance

PÊter Farkas PÊter Farkas ¡ Europe 2026 Virtual

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo ¡ WWC 2024

2:12 min

Navigating technical clarity as a global black belt

Chris Heilmann +2 ¡ LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou ¡ Coffee With Developers

4:03 min

Managing massive power consumption scaling in AI data centers

Stephan Gillich Stephan Gillich +3 ¡ WWC 2024

Videos

See all

Related articles

See all