Chief Data Center Critical Infrastructure Engineer

Amazon.com, Inc.
Lubbock, TX, United States
6 days ago
Apply on find.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$170,000.0 - $210,000.0
Working hours
Regular working hours
Job source

Tech stack

Computer-Aided Design Amazon Web Services Data Centers Monitoring of Systems Supervisory Control and Data Acquisition (SCADA) Reliability Engineering Data Analytics AWS Data Analytics

Job description

Amazon Web Services, Inc. is seeking a Data Center Chief Engineer to lead reliable, safe, and efficient operation of large-scale AWS data centers. You will own mission-critical electrical, mechanical, and cooling systems, driving 24x7 uptime for global cloud services. Responsibilities include leading maintenance and incident response, optimizing power and cooling capacity, and guiding infrastructure upgrades. You’ll mentor engineering and facilities teams in a high-ownership, data-driven culture, using metrics and automation to improve reliability, efficiency, and safety while partnering with global AWS operations and design teams., * Lead 24x7 reliability and performance of AWS data center critical infrastructure (power, cooling, fire/life safety, BMS/SCADA).

  • Own preventive and predictive maintenance programs; oversee inspections, testing, root-cause analysis, and incident remediation.
  • Manage design, commissioning, and upgrades of electrical and mechanical systems to meet capacity, efficiency, and resiliency goals.
  • Develop and enforce engineering standards, operating procedures, and change management for critical environments.
  • Direct and mentor engineering and facilities teams; coordinate with operations, capacity planning, and construction partners.
  • Monitor building management and monitoring systems; drive data-driven optimization and energy efficiency initiatives.
  • Ensure compliance with safety codes, regulatory requirements, and AWS security and availability standards.
  • Lead on-call and emergency response, incident command, and post-incident reviews for infrastructure events.

Requirements

  • Data center electrical systems (UPS, generators, switchgear)
  • Mechanical/HVAC and cooling systems for critical environments
  • Building Management Systems (BMS) and monitoring tools
  • Reliability engineering, root-cause analysis, and incident management
  • Preventive and predictive maintenance planning
  • Project management for infrastructure upgrades and expansions
  • Regulatory, safety, and code compliance for facilities
  • Capacity planning and energy efficiency optimization
  • Vendor and contractor management
  • Technical leadership and cross-functional communication

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:32 min

Structuring platforms for new services and data analytics

Nevelina Aleksandrova · LIVE

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

4:03 min

Managing massive power consumption scaling in AI data centers

Stephan Gillich Stephan Gillich +3 · World Congress 2024

1:24 min

Evaluating formal AWS certifications versus raw practical engineering experience

Jan Giacomelli · LIVE

1:10 min

Introduction to Microsoft Fabric and data agents

Dr. Alexander Wachtel Dr. Alexander Wachtel +1 · World Congress 2025

Videos

See all

Related articles

See all