Chief Data Center Facilities Engineering Manager
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
Amazon Web Services, Inc. is seeking a Data Center Chief Engineer to lead critical infrastructure operations for large-scale AWS data centers. In this role, you will own the reliability, performance, and safety of power, cooling, and life-safety systems that support millions of customers worldwide. You will direct maintenance programs, vendor activities, incident response, and continuous improvement initiatives in a fast-paced, high-ownership environment. Partnering with cross-functional AWS teams, you will drive standards, capacity planning, and engineering best practices while mentoring technicians and engineers in a customer-obsessed, data-driven culture., * Lead operation and maintenance of AWS data center critical infrastructure (power, cooling, life-safety systems).
- Develop and implement preventive and predictive maintenance programs to maximize uptime.
- Oversee vendor management, inspections, and repair activities to meet AWS reliability standards.
- Monitor facility performance, analyze data, and drive root-cause investigations and corrective actions.
- Ensure compliance with safety, environmental, and regulatory requirements across the facility.
- Collaborate with operations, capacity planning, and product teams on new deployments and upgrades.
- Design and improve procedures, runbooks, and standards for incident response and change management.
- Provide technical leadership, mentoring, and training for engineering and facilities staff.
- Support capacity planning, risk assessments, and redundancy strategies for large-scale infrastructure.
- Participate in on-call rotation and lead response to critical incidents and emergency situations.
Requirements
- Data center infrastructure management (power, cooling, UPS, generators)
- Electrical and mechanical systems troubleshooting
- Building Management Systems (BMS) and SCADAReliability engineering and root cause analysis (RCA)
- Preventive and predictive maintenance planning
- Incident, change, and risk management
- Reading and interpreting technical drawings and schematics
- Project management for infrastructure upgrades
- Capacity planning for power and cooling
- Regulatory, safety, and environmental compliance
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters
Making Data Warehouses Fast: A Developer’s Story
What Are The Top Skills Required For Azure Developers?