Data Center Server Operations Manager (2nd Shift)

Google LLC
Stillwater, OK, United States
3 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$122,000.0 - $173,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computer Clusters Data Centers Systems Theories Machine Learning Network Protocols Network Routers Google Cloud Hardware Infrastructure

Job description

Google isn’t just a software company. The Hardware Operations team is responsible for monitoring the state-of-the-art physical infrastructure behind Google’s powerful search technology. As a Hardware Operations Manager, you will manage a team of Data Center Technicians. You will oversee the quality installation of server hardware and components and take charge of complicated installations/troubleshooting. Your team will install, configure, test, troubleshoot and maintain hardware (like servers and its components) and server software (like Google’s Linux cluster). They will also take on the configuration of more complex components such as networks, routers, hubs, bridges, switches and networking protocols. They may lead small project teams on larger installations and develop project contingency plans. The AI and Infrastructure team is redefining what’s possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide. We’re the driving force behind Google’s groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more. Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $122000 - $173000 (USD) + 15% bonus target + equity + benefits Learn more about . Responsibilities

  • Lead a team of individuals, communicate individual and team priorities that support organizational goals to repair, fix, and perform preventative maintenance on equipment, servers, machines, or infrastructure based on issues.
  • Partner with teams to meet goals and stakeholders to manage facility activities and set/implement strategies.
  • Maintain, monitor, and execute security and operational procedures and analyze trends to identify opportunities for improvements ensuring alignment with organizational policies.
  • Support and contribute to the implementation of Environmental Health and Safety (EHS) and other compliance programs and initiatives in collaboration with other teams to ensure environmental and safety incidents are investigated, resolved, and reported.
  • Manage a team of Machine Learning (ML) travelers remotely and contribute and support 24/7 initiatives.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also and If you have a disability or special need that requires accommodation, please let us know by completing our .

Requirements

  • Bachelor’s degree or equivalent practical experience.
  • 5 years of experience in computing infrastructure, networking, operating systems, or hardware.
  • 3 years of experience managing a technical team.
  • Experience leading and managing projects from initiation to completion, and directing operations management initiatives.

Ability to work non-standard hours and differing rotations/shifts. Preferred qualifications:

  • Experience with initiating and executing initiatives in a global environment.
  • Experience working in data center environments, including building and operating infrastructure, and network and compute architecture and lifecycle, and Linux/Unix system administration.
  • Ability to implement and drive the safety culture at the site, while fostering a collaborative team environment.
  • Excellent problem-solving and presentation skills.
  • Track record of leading and improving Environmental Health and Safety initiatives.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:05 min

Adapting workplace policies to support distributed technical teams

Tejas Chopra Tejas Chopra · LIVE

5:02 min

Manual port forwarding configuration using network address translation

Oliver Seitz Oliver Seitz · World Congress 2025

6:10 min

Unlocking free learning credits via Google Cloud Innovators

Asrar Asrar · World Congress 2024

51 sec

Repurposing hardware and operating underwater data centers

Chris Heilmann +1 · LIVE

4:48 min

Provisioning and modifying Google Cloud compute resources

Devlin Duldulao · LIVE

4:38 min

How Skupper routes traffic across clusters

Alex Soto Alex Soto · World Congress 2024

Videos

See all

Related articles

See all