Technical Support Engineer

Advanced Micro Devices, Inc.
Austin, TX, United States
3 days ago
Apply on diversityjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$71,200.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence JIRA Intelligent Platform Management Interface Bash Shell BIOS Computer Engineering Data Centers Data Center Infrastructure Management (CIM) Software Debugging Microprocessors Disaster Recovery Firmware
+12 more
Python (Programming Language) Linux Commands PCI Express Windows PowerShell CPLD Scripting Computer Network Technologies Hardware Testing Information Technology Hardware Infrastructure Server Operating Systems & Platforms Servicenow

Job description

AMD is seeking a Data Center Operations Lead to oversee daily data center operations, lead a team of technicians, and provide advanced support for pre-production silicon and custom server platforms. This role combines people leadership, hands-on platform debugging, DCIM ownership, and operational excellence to ensure a secure, efficient, and scalable environment supporting AMD engineering and validation teams., Team Leadership & Operations

  • Lead, mentor, and develop a team of data center technicians through coaching, training, and performance management.
  • Manage staffing, schedules, PTO coverage, and escalation support to maintain operational service levels.
  • Partner with hiring managers on recruiting, interviewing, onboarding, and training new employees and contractors.
  • Establish and enforce standards for cabling, labeling, rack builds, ESD handling, safety, and lab cleanliness.
  • Translate engineering priorities into actionable work plans and serve as the primary escalation point for technical and operational issues.

Advanced Platform Bring-Up & Debug

  • Provide senior-level support for bring-up, validation, and debug of pre-production AMD silicon and custom server platforms.
  • Lead root-cause analysis of platform failures, including boot, POST, memory, PCIe, thermal, power, and firmware-related issues.
  • Utilize board-level and platform debugging techniques, including serial consoles, BMC/IPMI/Redfish, POST diagnostics, and lab instrumentation.
  • Perform BIOS, BMC, CPLD, and firmware updates, recovery, and image restoration.
  • Partner with silicon, firmware, validation, and platform engineering teams to drive issues through resolution.
  • Maintain secure handling, tracking, and disposition of pre-release hardware and sensitive assets.
  • Create and maintain runbooks, knowledge articles, and technical documentation.

DCIM & Facilities Management

  • Maintain ownership of DCIM data, including asset tracking, rack elevations, connectivity mapping, and capacity management.
  • Conduct audits and reconciliation activities to ensure asset and infrastructure accuracy.
  • Monitor data center power, cooling, environmental conditions, and redundancy health, proactively addressing risks.
  • Support capacity planning for space, power, cooling, and network resources.
  • Coordinate maintenance activities, vendor support, facility projects, and change-control processes.
  • Maintain operational documentation, diagrams, procedures, and disaster recovery plans.

Continuous Improvement

  • Lead rack-and-stack, cabling, hardware deployment, relocation, and decommissioning projects.
  • Manage inventory, spare parts, asset tracking, and procurement support.
  • Drive process improvements, automation, and operational efficiencies.
  • Ensure compliance with AMD safety, security, environmental, and export control requirements.
  • Track and report operational metrics, team performance, incidents, and capacity trends., AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

Requirements

You are a hands-on technical leader with experience managing teams in data center, server hardware, or lab environments. You excel at troubleshooting complex platform issues, driving operational improvements, and partnering with engineering stakeholders to deliver results. You bring strong organizational skills, a customer-focused mindset, and the ability to balance strategic planning with day-to-day execution., * 7+ years of experience in data center operations, server hardware support, hardware validation, or related technical environments.

  • Prior experience leading or supervising technical teams.
  • Strong knowledge of x86 server architecture, including CPUs, memory, PCIe, storage, networking, power, and thermal subsystems.
  • Hands-on experience with hardware and platform debugging tools, including BMC/IPMI/Redfish, serial consoles, and POST diagnostics.
  • Experience performing BIOS, BMC, and firmware updates and recovery procedures.
  • Experience with DCIM or related asset, capacity, and infrastructure management systems.
  • Understanding of data center power distribution, redundancy, cooling, and environmental monitoring.
  • Linux command-line proficiency and foundational networking knowledge.
  • Strong written and verbal communication skills with demonstrated technical documentation experience.
  • Ability to lift up to 50 lbs., work on ladders or lifts, and perform physical duties in a data center environment.
  • Ability to work onsite in Plano, TX and support occasional after-hours maintenance and escalation activities.

Preferred

  • Experience supporting pre-production or engineering-sample silicon programs.
  • Familiarity with AMD EPYC, AMD Instinct, or comparable server and accelerator platforms.
  • Experience with Sunbird dcTrack or similar enterprise DCIM solutions.
  • Experience with ServiceNow, Jira, ChangeGear, or other ITSM/change management platforms.
  • Scripting experience with Python, Bash, or PowerShell for automation and operational efficiency.
  • Experience supporting colocation facilities or multi-site data center environments.

Education

  • Bachelor’s degree in Information Technology, Computer Engineering, Electrical Engineering, Computer Science, or a related field preferred.
  • Equivalent combination of education, military service, technical certifications, and relevant industry experience will also be considered.

About the company

At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And we’re looking for talent who feel the same: people who want to leave the planet better than they found it, those who don’t shy away from humanity’s challenges but are determined to help solve them.

AMD is powering the next generation of supercomputing, high-performance computing, cloud, and AI. Whether you’re designing next-gen processors, enabling AI breakthroughs, or creating go-to-market plans, every role at AMD contributes to something bigger - technology that moves the world forward.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on diversityjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

41 sec

Massive client data loss and bio-digital storage

Chris Heilmann +1 · LIVE

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

2:48 min

Daily responsibilities and alignment practices for technical engineering leadership

Edoardo Dusi · LIVE

1:37 min

Mapping team collaboration networks using jira metadata

Dmitry Yanter Dmitry Yanter · World Congress 2025

Videos

See all

Related articles

See all