Cloud Platforms and Infrastructure Engineer, TPU/GPU

Google LLC
San Francisco, CA, United States
26 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$152,000.0 - $221,000.0
Working hours
Regular working hours
Job source

Tech stack

C (Programming Language) Java (Programming Language) Artificial Intelligence Border Gateway Protocol C++ (Programming Language) Cloud Computing Cyber Security Continuous Integration Data Structures DevOps Hypertext Transfer Protocols (HTTP) Identity and Access Management
+20 more
Python (Programming Language) Key Management Linux System Administration Machine Learning Network Protocols Tensorflow Service Discovery Software Engineering TCP/IP Google Cloud Load Balancing Cloud Platform System Pytorch Kubernetes Information Technology Fortinet Slurm Serverless Computing Golang Vmware

Job description

As a Cloud Platform and Infrastructure Engineer, you will provide technical guidance to customers adopting Google Cloud Platform (GCP) services, including providing best practices on secure foundational cloud implementations, automated provisioning of infrastructure and applications, cloud-ready application architectures, and more. You will also provide guidance in ensuring that customers receive the best of what GCP can offer and have the best experience in migrating, building, modernizing, and maintaining applications in GCP. Additionally, you will work with Product Management and Product Engineering to drive excellence in Google Cloud products and features., * Propose solution architectures and manage the deployment of cloud networking solutions according to customer requirements and implementation best practices.

  • Design and implement cloud-based technical architectures, migration approaches, and application optimizations that enable business objectives in collaboration with customers.
  • Work with internal Specialists, Product, and Engineering teams to package approaches, best practices, and lessons learned into thought leadership, methodologies, and published assets.
  • Interact with sales, partners, and customer technical stakeholders to manage project scope, priorities, deliverables, risks and issues, and timelines.
  • Travel up to 30% for in-region for meetings, technical reviews, and onsite delivery activities.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google’s EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

Requirements

  • Bachelor’s degree in Computer Science or equivalent practical experience.
  • 6 years of experience automating infrastructure provisioning, Developer Operations (DevOps), continuous integration, or delivery utilizing Kubernetes and Linux-based systems.
  • 3 years of experience in project management and technical solution delivery.
  • Experience coding in one or more general purpose languages (e.g., Python, Java, Go, C or C++) including data structures, algorithms, software design, Linux environments and Kubernetes orchestration.
  • Experience working with Cloud Providers such as Google Cloud Platform (GCP).
  • Ability to travel 30% of the time, as needed, for client engagements., * Experience with third-party networking (e.g., PANW, Fortinet, VMWare) and design, including redundancy and load balancing.
  • Experience troubleshooting networking protocols including TCP/IP, Hypertext Transfer Protocol, and Border Gateway Protocol (BGP).
  • Experience in customer-facing migration, including service discovery, assessment, planning, execution, and operations.
  • Experience with standard IT security practices, including IAM, data protection, encryption, and certificate/key management.
  • Experience running AI/ML training and inference workloads on GPU/TPU using frameworks such as PyTorch, JAX, TensorFlow, or Slurm.
  • Knowledge of containerization and orchestration technologies, including Google Kubernetes Engine (GKE) and related cloud-native services., Google Cloud accelerates every organization’s ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google’s cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:06 min

High-paying roles driven by Google Cloud certifications

Asrar Asrar · World Congress 2024

1:40 min

Managing containerized infrastructure with Podman Desktop

Cedric Clyburn Cedric Clyburn +1 · World Congress 2025

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:41 min

Parallels between cloud and legacy infrastructure lock-ins

Björn Stahl Björn Stahl · World Congress 2024

Videos

See all

Related articles

See all