US Head of Data Center Network Operations

Bridgesource Solutions
United States
26 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Data Analysis Cloud Computing Data Centers Network Architecture Network Administration AI Infrastructure Computer Network Operations HybridCloud Data Center Networking Kubernetes Data Management

Job description

Our client provides cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. They enable AI-focused companies to achieve superior results by reducing the complexity of AI development. Their GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.

We are seeking a US Head of Data Center Network Operations to lead the end-to-end network management of the datacenter portfolio across the United States. You’ll be responsible for ensuring operational excellence, safety, compliance, and reliability across all network infrastructure, while driving continuous improvement and scaling operations to support rapid business growth. Regular travel to the Data Centers (DC) is essential to the success of this position. This is a high-impact leadership role where you’ll own the strategic direction of DC network operations, manage cross-functional teams, and serve as a critical partner to senior leadership in delivering world-class infrastructure that powers their Hyperscale platform., * Own the strategic vision and execution of datacenter network infrastructure operations across the region, ensuring alignment with business objectives and growth plans.

  • Establish and maintain operational standards, processes, and procedures that drive reliability, safety, and efficiency across all sites.
  • Lead the development and implementation of operational roadmaps that support capacity planning, infrastructure scaling, and service delivery milestones.
  • Drive continuous improvement initiatives to reduce downtime, enhance operational maturity, and optimize costs.
  • Build, mentor, and lead high-performing teams across multiple Data Center sites, specifically network operations staff.
  • Establish clear accountability structures, performance metrics, and development pathways for direct reports and broader teams.
  • Foster a culture of ownership, safety, and excellence where team members are empowered to make decisions and drive impact.
  • Conduct regular performance reviews, provide constructive feedback, and support career progression.
  • Oversee Data Center Site Managers in their execution of day-to-day Network Operational procedures, from routine inspections to the handling of IT Service Management (ITSM) tickets ensuring all Service Level Agreements (SLA)s are met.
  • Maintain accurate asset inventory for all AI Infrastructure and supporting hardware and tooling.
  • Establish and maintain Service Level Objectives (SLO)/Service Level Indicators (SLI) for Data Center availability, performance, and incident response.
  • Lead incident response and root-cause analysis for network operational failures; own remediation and prevention strategies.
  • Ensure full compliance with health and safety regulations, industry standards and best practices.
  • Support ongoing certifications and audits (familiar with ISO 27001, ISO 22237, SOC 2, Cyber Essentials Plus, ISO 22301).
  • Maintain comprehensive documentation for compliance, audit readiness, and regulatory requirements.
  • Partner closely with US Head of Mechanical, Electrical, and Plumbing (MEP), Infrastructure Engineering, Network Engineering, and Security teams to ensure operational readiness and alignment.
  • Work with the Deployment Supply Chain team to support hardware intake, staging, and deployment timelines.
  • Collaborate with Finance and Commercial teams on capacity planning, cost optimization, and customer commitments.
  • Support project delivery teams in commissioning new sites and scaling existing facilities.
  • Engage with senior leadership on operational metrics, risk management, and strategic initiatives.
  • Establish Key Performance Indicators (KPI) and Key Risk Indicators (KRI) for operational health (uptime, energy efficiency, cost per rack, incident rates, etc.).
  • Manage monitoring and alerting systems to track infrastructure performance and environmental conditions.
  • Produce regular operational reports for senior leadership, including performance metrics, risks, and improvement initiatives. Use data-driven insights to identify optimization opportunities and inform decision-making.

Requirements

  • 10+ years of experience in Data Center operations or Network infrastructure management at scale. Background in hyperscale, cloud, or HPC Data Center environments (preferred).
  • Proven track record leading regional or multi-site operations in a high-growth, fast-paced environment.
  • Experience managing teams across multiple locations and coordinating complex operational initiatives.
  • Demonstrated success in maintaining high reliability standards, scaling operations, and improving efficiency.
  • Deep understanding of Data Center network infrastructure, networking, and security, and working knowledge of including power systems and cooling.
  • Familiarity with ISO 22237 (Datacenter design and operations) and ISO 27001 Annex A.11 (physical security).
  • Understanding of GPU/HPC infrastructure and the unique operational requirements of AI cloud platforms.
  • Proven ability to establish processes, standards, and controls that scale with business growth.

Nice to Have

  • Experience with Palantir Foundry or similar data platforms for operational analytics.
  • Familiarity with infrastructure telemetry and usage-based billing data.
  • Familiarity with operations infrastructure including Mechanical, Electric, and Plumbing.
  • Background in sustainability and energy efficiency optimization.
  • Experience supporting customer-facing SLAs and service delivery commitments.
  • Knowledge of Kubernetes, container orchestration, or hybrid cloud architecture.
  • Security certifications or deep familiarity with GRC tooling.

About the company

Bridge Source Utilities Solutions is a Smart Grid Advisory & Strategic Staffing Solutions firm focused on connecting multinational Utility corporation, Power Generation, and Data Center markets. Founded in 2012 with the vision to assist companies and resources to develop a clear path for achieving their goals, BridgeSource provides advice and a connect workforce in the areas of grid modernization, technology transformations, project management, change adoption, distribution operations, and capital investment planning.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · WWC 2022

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:03 min

Managing massive power consumption scaling in AI data centers

Stephan Gillich Stephan Gillich +3 · WWC 2024

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · WWC 2022

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all