Chief Data Center Critical Infrastructure Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
Amazon Web Services, Inc. is seeking a Data Center Chief Engineer to lead reliable, safe, and efficient operation of large-scale AWS data centers. You will own mission-critical electrical, mechanical, and cooling systems, driving 24x7 uptime for global cloud services. Responsibilities include leading maintenance and incident response, optimizing power and cooling capacity, and guiding infrastructure upgrades. You’ll mentor engineering and facilities teams in a high-ownership, data-driven culture, using metrics and automation to improve reliability, efficiency, and safety while partnering with global AWS operations and design teams., * Lead 24x7 reliability and performance of AWS data center critical infrastructure (power, cooling, fire/life safety, BMS/SCADA).
- Own preventive and predictive maintenance programs; oversee inspections, testing, root-cause analysis, and incident remediation.
- Manage design, commissioning, and upgrades of electrical and mechanical systems to meet capacity, efficiency, and resiliency goals.
- Develop and enforce engineering standards, operating procedures, and change management for critical environments.
- Direct and mentor engineering and facilities teams; coordinate with operations, capacity planning, and construction partners.
- Monitor building management and monitoring systems; drive data-driven optimization and energy efficiency initiatives.
- Ensure compliance with safety codes, regulatory requirements, and AWS security and availability standards.
- Lead on-call and emergency response, incident command, and post-incident reviews for infrastructure events.
Requirements
- Data center electrical systems (UPS, generators, switchgear)
- Mechanical/HVAC and cooling systems for critical environments
- Building Management Systems (BMS) and monitoring tools
- Reliability engineering, root-cause analysis, and incident management
- Preventive and predictive maintenance planning
- Project management for infrastructure upgrades and expansions
- Regulatory, safety, and code compliance for facilities
- Capacity planning and energy efficiency optimization
- Vendor and contractor management
- Technical leadership and cross-functional communication
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
Making Data Warehouses Fast: A Developer’s Story
Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters
Top Big Data Technologies That You Need to Know