Senior Product Manager, Host Networking Software

Google LLC
Sunnyvale, CA, United States
26 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$192,000.0 - $278,000.0
Working hours
Regular working hours
Job source

Tech stack

X86 Architecture Intelligent Platform Management Interface BIOS Computer Networks Network Congestion Data Centers Device Drivers Distributed Systems Ethernet InfiniBand Linux Kernel Network Architecture
+7 more
Network Protocols PCI Express Azure Machine Learning Server Administration Computer Networking Systems Computer Network Technologies Data Center Networking

Job description

  • Partner with engineering leads to define the strategy for ML networking reliability, delivering host networking libraries for training and inference (e.g., NCCL, NIXL and other collective communication libraries), telemetry and observability. This applies to Google’s standard compute, TPU and GPU systems, aligning with company goals and needs.
  • Collaborate with Google’s Host Networking Software teams and Google internal PAs to identify capabilities and facilitate the project proposal and selection process, ensuring quality, feasibility, and alignment with objectives.
  • Work with customers and internal engineering teams to define requirements and performance of system level host networking software stacks.
  • Define success metrics and track progress throughout the program. Analyze results, identify areas of improvement, and communicate insights to stakeholders.
  • Develop an understanding of different segments of Google Host Networking infrastructure and align roadmaps with business objectives and customer needs.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

Requirements

  • Bachelor’s degree or equivalent practical experience.
  • 8 years of experience in product management or a related technical role.
  • 3 years of experience taking technical products from conception to launch (e.g., ideation to execution, end-to-end, 0 to 1, etc.).
  • Experience with networking protocols, congestion control, distributed computing or network infrastructure., * Experience with scale-up and scale-out networking solutions for ML Platforms using high performance networking technologies such as Ethernet, Infiniband, RoCE, NVLINK, etc.
  • Experience defining product requirements for data centers, particularly at a cloud provider, ISP, or telco.
  • Advanced knowledge of the concepts and architectures of data center networking, Linux networking, PCIe, Ethernet, x86/Arm, GPU/TPU server platforms.
  • Understanding of server system design, server management (baseboard management controller (BMC)), and Linux kernel (basic input/output system (BIOS), device drivers).

About the company

At Google, we put our users first. The world is always changing, so we need Product Managers who are continuously adapting and excited to work on products that affect millions of people every day.

In this role, you will work cross-functionally to guide products from conception to launch by connecting the technical and business worlds. You can break down complex problems into steps that drive product development.

One of the many reasons Google consistently brings innovative, world-changing products to market is because of the collaborative work we do in Product Management. Our team works closely with creative engineers, designers, marketers, etc. to help design and develop technologies that improve access to the world’s information. We’re responsible for guiding products throughout the execution cycle, focusing specifically on analyzing, positioning, packaging, promoting, and tailoring our solutions to our users.

We strive to build the world’s best domain specific and standard computing networking infrastructure by developing the host networking software stack and infrastructure for TPU, GPU and standard compute supercomputers for both the Google Cloud and Google first-party (1P) technical infrastructure. Our mission centers on building high-performance host networking software-ranging from drivers, libraries, congestion management, reliability, observability and telemetry-that empowers the next generation of AI training and inference.

The AI and Infrastructure team is redefining what’s possible. We empower Google customers with breakthrough capabilities and insights by delivering AI and Infrastructure at unparalleled scale, efficiency, reliability and velocity. Our customers include Googlers, Google Cloud customers, and billions of Google users worldwide.

We’re the driving force behind Google’s groundbreaking innovations, empowering the development of our cutting-edge AI models, delivering unparalleled computing power to global services, and providing the essential platforms that enable developers to build the future. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more. Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $192000 - $278000 (USD) + 20% bonus target + equity + benefits

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:57 min

Routing cross-rack traffic seamlessly with NCCL

Kevin Klues Kevin Klues · World Congress 2025

4:52 min

Connecting namespaces with local virtual ethernet pairs

Oliver Seitz Oliver Seitz · World Congress 2025

41 sec

Massive client data loss and bio-digital storage

Chris Heilmann +1 · LIVE

1:51 min

Bypassing the CPU stack with remote direct memory access

Lerna Ekmekcioglu Lerna Ekmekcioglu · Europe 2026 Virtual

3:23 min

The AI workload technology stack and its components

Lerna Ekmekcioglu Lerna Ekmekcioglu · Europe 2026 Virtual

Videos

See all

Related articles

See all