System Development Manager, AWS Manufacturing and Repair

Amazon.com, Inc.
Austin, TX, United States
3 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$166,300.0 - $225,000.0
Working hours
Regular working hours
Job source

Tech stack

Testing (Software) Java (Programming Language) Agile Methodology Artificial Intelligence Amazon Web Services Systems Engineering Automation of Tests Bash Shell C Sharp (Programming Language) C++ (Programming Language) Code Review Computer Programming
+27 more
Computer Engineering Continuous Delivery Software Design Patterns Firmware Hardware Design Internet Services Python (Programming Language) Uptime Windows PowerShell Software Architecture Systems Development Life Cycle Cloud Services Ruby Azure Machine Learning Scaled Agile Framework Software Engineering Statistical Process Control (SPC) Test Data Strategies of Testing AI Infrastructure Information Technology Data Analytics AWS Data Analytics Hardware Infrastructure Software Coding Software Version Control Golang

Job description

AWS Manufacturing & Repair operations is focused on automated assembly solutions where innovative approaches are needed to accelerate overall manufacturing time for AI/ML platforms. We work backwards from customers to increase critical capacity for AWS services by improving yield and reliability while reducing cycle time. This highly cross-functional initiative brings together critical expertise from across Amazon teams dedicated to meeting and exceeding customer expectations through innovative validation processes and Amazon operational excellence.

We are seeking a Systems Development Manager to lead and grow a team of System Development Engineers and Manufacturing Test Engineers who own the operational execution of the end-to-end manufacturing test ecosystem for AI server hardware - from early NPI builds through high-volume production and sustaining - across server (L10) and rack-level (L11/L12) integration. This is an opportunity to work with internal partner teams across Software and Hardware Engineering to shape manufacturing test strategy from the ground up, build a world-class engineering team, and directly impact the speed and quality at which AI infrastructure reaches customers.

This role sits at the intersection of hardware, software, production operations, and people leadership. This role will support the development of the technical vision and also hire, mentor, and develop the engineers who bring it to life. You’ll establish the team’s roadmap, drive automation and self-remedy capabilities across the production line, and relentlessly pursue uptime, yield, and operational excellence. You’ll be a hands-on leader who dives deep into the details on business, operations, and engineering while simultaneously growing the next generation of technical leaders on your team.

AWS Manufacturing & Repair is part of the broader AWS Infrastructure Services team, which owns the design, planning, delivery, and operation of all AWS global infrastructure. We are the teams who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain - and we’re looking for talented people who want to help. You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

Come join our team and be a part of history as we deliver results for the largest cloud services company on Earth!

This role directly supports physical manufacturing environments across multiple locations with travel expected up to 25% of the time.

Key job responsibilities

Lead & Develop the Team

Hire, motivate, mentor, and develop a team of System Development and Manufacturing Test Engineers, promoting growth and opportunity at every level.

Define the structure and strategy for the team, determining the right mix of engineers and where they need to be allocated to meet business goals across multiple manufacturing sites.

Develop senior engineers by promoting growth and providing opportunities to demonstrate higher-level scope, impact, complexity, and leadership.

Foster a culture of ownership, curiosity, and operational excellence where engineers take end-to-end responsibility for the systems they build and support.

Establish Roadmap & Technical Vision

Establish and own the manufacturing test engineering roadmap, successfully delivering engineering solutions that execute that vision. Your roadmap influences organizational goals and you are involved in the yearly planning process.

In partnership with internal Software and Hardware Engineering teams, define modular test strategies, pass/fail limits, and statistical process controls (GR&R, Cp/Cpk, correlation analysis) to ensure product quality at scale.

Architect the manufacturing test infrastructure for AI server hardware products, spanning board-level validation through full rack integration and networking.

Drive Automation, Uptime & Yield

Lead the development of automated test processes and self-healing capabilities that reduce manual intervention, increase throughput, and improve first-pass yield across the production line.

Own functional test yield across shifts, lead triage efforts for test-station failures, and drive weekly production quality reviews with cross-functional partners.

Own the Test Ecosystem

Oversee the design and maintenance of production test software and automation frameworks using Python, Bash, and related tooling, including test sequencing, equipment control, data collection, and reporting pipelines.

Guide test fixture design and procurement, ensuring alignment to cost, scalability, and schedule requirements; engage in RFQ processes for optimal end-to-end test solutions.

Provide ongoing support for deployed test systems, implement configuration changes for new product variants, and maintain comprehensive documentation of test procedures, specifications, and infrastructure.

Be an Operational Leader

Take a data-driven approach, consistently utilizing metrics to measure performance gaps and increase productivity by improving processes at the local level while driving improvements on the global scale.

Compile and analyze test data using both automated reporting systems and exploratory analysis techniques to identify trends, root-cause failures, and drive corrective actions.

Support New Product Introduction (NPI) by partnering with hardware engineering and NPI test teams to develop test solutions from concept prototypes through production ramp, ensuring readiness at each phase gate.

Collaborate Across Teams

Work cross-functionally with hardware design, firmware, process engineering, quality, and operations teams to ensure test requirements are aligned and products are designed for efficient, high-quality manufacturing.

Foster constructive dialogue, harmonize discordant views, and lead the resolution of contentious issues across partner organizations.

Deliver actionable design-for-test (DFT) feedback to hardware and firmware development teams to improve testability and manufacturing efficiency for future platforms.

Requirements

  • 5+ years of managing system or software development teams experience
  • Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field
  • Experience in systems engineering and operations leadership for an Internet service or leading-edge IT organization
  • Experience in managing system or software development teams
  • Experience (hands-on) in systems engineering and administrative work in networking, storage systems, and operating systems
  • Experience in agile software development methodology, * Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence
  • Knowledge of systems engineering fundamentals (networking, storage, operating systems)
  • Experience with Agile engineering practices (Kanban, continuous delivery, etc.)
  • Experience with AWS platforms, services, and design patterns
  • Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
  • Experience in automating, deploying, and supporting large-scale infrastructure

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, KY, Florence - 166,300.00 - 225,000.00 USD annually

USA, TX, Austin - 166,300.00 - 225,000.00 USD annually

USA, WA, Bellevue - 166,300.00 - 225,000.00 USD annually

About the company

Here at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences.

Work/Life Harmony

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:22 min

Leveraging unique cultural backgrounds in engineering design

Ixchel Ruiz · LIVE

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

2:27 min

Structuring agile and interdisciplinary engineering teams

Oliver Zimmert · LIVE

Videos

See all

Related articles

See all