Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms and Servers

Amazon.com, Inc.
Seattle, WA, United States
5 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$157,300.0 - $212,800.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Amazon Elastic Compute Cloud Big Data Cloud Computing Computer Engineering Data Centers Software Debugging Linux Firmware Hardware Design Systems Development Life Cycle Web Services
+4 more
Network Switches Scripting AWS Data Analytics Network Server

Job description

Amazon Web Services (AWS) Hardware Engineering Services (HWEngS) owns the new product development (NPI) and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain, and we’re looking for talented people who want to help. You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers, and you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

Cloud Hardware Development Engineer (CHDE): The Amazon Web Services (AWS) Hardware Engineering Services (HWEngS) Specialized Platforms and Servers team creates Enterprise rack solutions for Amazon’s innovative web services. We are seeking experienced CHDEs to own the fleet health, diagnostics, and automation of the Enterprise rack solutions:

1.) Designing and implementing predictive failure detection systems using telemetry, sensor data, error trends, and log correlation to identify hardware issues before they cause a customer impact.

2.) Driving toward zero-touch operations by building detection, diagnostics, and remediation of faults without human intervention

3.) Debugging complex system failures in time-sensitive settings personally diving deep when the problem demands it.

4.) Completing root cause analysis correlating across firmware, kernel, driver, thermal, power, and physical layers.

What you will do: As a member of the Specialized Platforms and Servers team, you’ll be responsible for collaborating with Elastic Cloud Compute (EC2) service teams and Data Center Operations to maintain fleet health in all the locations we have servers.

You will work closely with internal teams, suppliers, and external partners capturing lessons learned while operating the fleet to ensure next generation designs are of the highest quality, constantly looking for ways to improve your product performance, quality and cost., As a CHDE you will be responsible for scaling how we operate our massive existing & rapidly growing fleet. You will lead the integration and delivery of servers, support the development of automated monitoring, and failure analysis services to operate, debug, and scale our servers. You will work closely with other AWS software teams to tailor and operate servers solutions for the AWS environment. You will support launching our servers into production and operating our fleet of servers.

A day in the life

Your day to day responsibilities will be solving operational challenges to our existing fleet with the goal of improving the current customer experience as well as developing improved systems for future designs.

About the team

The team is comprised of CHDE’s, System Development Engineers and Technical Program Managers, all with the common goal of delivering the best specialized server fleet possible to our customers.

Requirements

  • Bachelor’s degree in electrical engineering, computer engineering, or equivalent, or 3+ years of relevant technical position experience.
  • 3+ years of experience in server level design for compute or other complex product design.
  • Experience in developing design verification plans and functional test procedures.
  • Experience in board and server root cause analysis and resolution., * 5+ years of experience in complex product development such as servers, network switches, or other highly integrated devices with hardware, software, and service aspects.
  • Proficient scripting, debug abilities, and Linux operations commands.
  • Experience deploying and operating hardware and applications across large data centers.
  • Meets/exceeds Amazon’s leadership principles requirements for this role
  • Meets/exceeds Amazon’s functional/technical depth and complexity for this role

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, CA, Cupertino - 157,300.00 - 212,800.00 USD annually

USA, WA, Seattle - 136,000.00 - 184,000.00 USD annually

About the company

Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.

Inclusive Team Culture

Here at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness.

Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:19 min

Orchestrating over-the-air firmware updates for vehicle modules

Denis Grahovac · World Congress 2021

1:24 min

Evaluating formal AWS certifications versus raw practical engineering experience

Jan Giacomelli · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

2:11 min

Building a custom vehicle telemetry platform on AWS

Tonci Zilic · World Congress 2022

Videos

See all

Related articles

See all