> Markdown version of [/jobs/ext/202587-senior-infrastructure-engineer-observe-by-snowflake](https://www.wearedevelopers.com/jobs/ext/202587-senior-infrastructure-engineer-observe-by-snowflake). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Infrastructure Engineer, Observe by Snowflake - **Company:** Snowflake Inc. - **Location:** Menlo Park, CA, United States - **Experience:** Expert - **Salary:** $200,000.0 - $287,500.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Apache HTTP Server, Microsoft Azure, Cloud Computing, Cloud Engineering, Code Review, Computer Programming, DevOps, Programming Tools, Distributed Systems, Fault Tolerance, Python (Programming Language), Reliability Engineering, Ansible, System Availability, Snowflake, Kubernetes, Infrastructure Automation Frameworks, Data Management, Terraform - **Published:** May 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=97b7d1e0b1475545 ## About the Role * 5+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), DevOps, or related roles. * Demonstrated experience designing and operating production systems at scale, with deep ownership of reliability and operational excellence. * Strong experience with container orchestration platforms such as Kubernetes or Nomad, including architectural decision-making and operational tuning. * Hands-on experience managing cloud infrastructure using Infrastructure-as-Code tools such as Terraform, Ansible, or similar, with a focus on scalable system design. * Strong programming skills in Go, Python, or similar languages, with a track record of building automation and infrastructure systems. * Experience driving cross-team technical initiatives and influencing infrastructure best practices. * Ability to balance immediate operational demands with long-term architectural vision. NICE TO HAVE * Deep experience operating large-scale distributed systems. * Familiarity with observability platforms, telemetry pipelines, or monitoring infrastructure. * Experience building or evolving internal developer platforms. * Experience working in high-growth, rapidly evolving engineering environments. * Experience with GCP and Azure ## Description At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don't just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset - who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is an AI-powered observability platform built on the Snowflake AI Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lakehouse, using open formats like Apache Iceberg, at dramatically lower cost. A dynamic Context Graph and chat-based AI SRE provide rich context and automated workflows so teams can move from detection to root cause and resolution 10x faster. Leading engineering teams at companies like Capital One, Topgolf, and Dialpad rely on Observe to troubleshoot hundreds of terabytes of telemetry daily while maintaining reliability at enterprise scale. As part of Snowflake, Observe combines startup-style ownership and velocity with the global reach, operational excellence, and ecosystem of one of the world's leading data platforms. The Infrastructure team at Observe by Snowflake is responsible for architecting, scaling, and operating the development and production environments that power our observability platform. We are a small, highly collaborative team with broad scope and high ownership, focused on delivering reliable, secure, and scalable infrastructure while continuously evolving the systems that support our engineers and customers. As a Senior Infrastructure Engineer, you will play a key role in shaping our infrastructure strategy, driving architectural decisions, and elevating operational excellence across the organization. WHAT YOU'LL DO * Lead the design, build, and operation of scalable, secure cloud infrastructure in AWS supporting a high-scale observability platform. * Drive architectural improvements that enhance reliability, performance, scalability, and operational visibility across development and production environments. * Own and evolve CI/CD pipelines, developer tooling, and platform automation to improve productivity and deployment safety at scale. * Proactively identify reliability, performance, and security risks, and lead efforts to mitigate them. * Design and implement infrastructure patterns that ensure high availability, fault tolerance, and operational resilience. * Play a key role in incident response, root cause analysis, and post-incident improvements, driving systemic reliability enhancements. * Partner cross-functionally with Product and Engineering teams to ensure infrastructure strategy supports long-term platform evolution. * Mentor and support other engineers through design reviews, code reviews, and operational best practices. ## Related Videos - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [How building an industry DBMS differs from building a research one](https://www.wearedevelopers.com/videos/768-how-building-an-industry-dbms-differs-from-building-a-research-one) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Terraform for Developers](https://www.wearedevelopers.com/videos/3-terraform-for-developers) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)