Site Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+1 more
Job description
We are looking for a talented Site Reliability Engineer (SRE) with a strong background in Google Cloud Platform (Google Cloud Platform) and Kubernetes.
The ideal candidate will be responsible for ensuring the reliability, performance, and scalability of on premise and cloud-based systems, with a focus on reducing costs for Google Cloud.
Responsibilities
System Reliability: Ensure the reliability and uptime of critical services and infrastructure.
Google Cloud Expertise: Design, implement, and manage cloud infrastructure using Google Cloud services.
Automation: Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention.
Monitoring and Incident Response: Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery.
Capacity Planning: Conduct capacity planning and performance tuning to ensure systems can handle future growth.
Requirements
Proficiency in Google Cloud services, including:
Compute Engine
Kubernetes Engine (GKE)
Cloud Storage
Pub/Sub
Experience with database technologies, including:
SQL and NoSQL
Databricks
Familiarity with Google BI and AI/ML tools, including
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
7 Cloud Computing Trends Coming in 2025 for Developers
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
Dev Digest 137 - AI'm not sure about this