> Markdown version of [/jobs/ext/3027281-sre-senior-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/3027281-sre-senior-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # SRE / Senior Infrastructure Engineer - **Company:** Clickhouse Inc - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Computer Programming, Continuous Integration, DevOps, PostgreSQL, Performance Tuning, Cloud Services, Prometheus, Google Cloud, Grafana, Multi-Cloud, Kubernetes, Vertica, Terraform - **Published:** September 22, 2026 - **Apply:** https://www.careerbuilder.com/job-details/senior-infrastructure-engineer-postgres-san-francisco-ca--7b62591c-f50e-4dd7-8490-aa2e9d5eea96 ## About the Role * Experience: 7+ years in SRE, DevOps, or infrastructure engineering, with a track record of running distributed, production-grade systems. * Database Operations: Solid understanding of Postgres operations, scaling, and performance tuning. * Cloud Expertise: Deep hands-on experience across AWS, with exposure to GCP and Azure; comfortable navigating multi-cloud topologies. * Automation Skills: Proficient with Terraform, Kubernetes, and container-based infrastructure. * Programming: Strong Go development skills (or willingness to write and own production Go code). * Observability: Familiar with tools like Prometheus, Grafana, Loki, OpenTelemetry, or equivalents. * Reliability Focus: Deep understanding of SLOs, incident response, and continuous improvement in service reliability. * Mindset: You operate with a founder's mentality - hands-on, resourceful, and willing to dive deep to get things done. You take pride in hard work, autonomy, and shipping impactful systems. ## Description You'll be at the center of how we run and evolve our next-generation data platform-building the automation, observability, and operational rigor that ensure a fast, secure, and dependable customer experience. This is a hands-on, high-impact role where you'll write code, shape architecture, and enable the broader engineering team to deliver with confidence and velocity. What You'll Do * Lead reliability and operations for ClickHouse's Postgres integration - upgrades, patching, maintenance, and scaling. * Design and implement automation for provisioning, deployments, and service lifecycle management across AWS, GCP, and Azure. * Develop infrastructure-as-code using Terraform and modern CI/CD tooling to ensure consistent, repeatable deployments. * Contribute Go-based tooling and services that improve automation, observability, and developer experience. * Own observability and monitoring, ensuring robust alerting, metrics, and tracing across environments. * Drive incident management and postmortem practices that strengthen reliability and learning loops. * Collaborate cross-functionally with platform, networking, and product teams to improve service operability. * Mentor and enable engineers, helping the team scale effectively as customer adoption grows. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers)