Infrastructure Engineer, Observe by Snowflake

Snowflake Inc.
Menlo Park, CA, United States
3 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$160,000.0 - $230,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Apache HTTP Server Microsoft Azure Cloud Computing Cloud Engineering Computer Programming DevOps Distributed Systems Graph Database Python (Programming Language) Systems Development Life Cycle
+10 more
Reliability Engineering Ansible Google Cloud System Availability Snowflake Reliability of Systems Data Lakes Kubernetes Infrastructure Automation Frameworks Terraform

Job description

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset - who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn’t just to execute a function, but to help redefine the future of how work gets done.

Observe by Snowflake is an AI-powered observability platform built on the Snowflake Data Cloud and engineered for scale. We ingest and store logs, metrics, traces, and events on an open, scalable data lake using open formats like Apache Iceberg, delivering deep correlation and long-term analytics at dramatically lower cost. A dynamic Knowledge Graph and chat-based AI SRE provide rich context and guided workflows so teams can move from detection to root cause and resolution significantly faster.

The Infrastructure team at Observe by Snowflake is responsible for building, scaling, and operating the development and production environments that power our observability platform. We are a small, highly collaborative team with a broad scope, focused on delivering reliable infrastructure while continuously improving the systems that support our engineers and customers.

WHAT YOU’LL DO

  • Design, build, and operate scalable cloud infrastructure in AWS supporting a high-scale observability platform.
  • Improve system reliability, performance, and operational visibility across development and production environments.
  • Develop and maintain CI/CD pipelines and internal tooling to improve developer productivity and deployment safety.
  • Identify and mitigate security risks, and help maintain internal security standards and compliance requirements.
  • Build infrastructure that supports high availability, scalability, and operational resilience
  • Participate in an on-call rotation, contributing to incident response and post-incident improvements.
  • Partner closely with engineering teams to ensure infrastructure supports evolving product and platform needs.

Requirements

  • 2+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), DevOps, or related roles.
  • Experience operating container orchestration platforms such as Kubernetes
  • Hands-on experience managing cloud infrastructure using Infrastructure-as-Code tools such as Terraform, Ansible, or similar.
  • Strong programming skills in Go, Python, or similar languages, with a focus on automation and systems development.
  • Experience supporting production systems at scale, with a focus on reliability and operational excellence.
  • Strong problem-solving skills and the ability to balance short-term operational needs with long-term infrastructure design.
  • Experience with AWS, GCP, and Azure

NICE TO HAVE

  • Experience operating large-scale distributed systems.
  • Familiarity with observability platforms, telemetry pipelines, or monitoring infrastructure.
  • Experience improving developer platform tooling or internal infrastructure platforms.
  • Experience working in high-growth or rapidly evolving engineering environments.

Benefits & conditions

Parental leave, 401(k), Health insurance, Paid time off, Vision insurance, Health savings account, Dental insurance, Employee assistance program, Every Snowflake employee is expected to follow the company’s confidentiality and security standards for handling sensitive data. Snowflake employees must abide by the company’s data security plan as an essential part of their duties. It is every employee’s duty to keep customer information secure and confidential.

Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

How do you want to make your impact?

For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com

The following represents the expected range of compensation for this role:

  • The estimated base salary range for this role is $160,000 - $230,000.
  • Additionally, this role is eligible to participate in Snowflake’s bonus and equity plan.

The successful candidate’s starting salary will be determined based on permissible, non-discriminatory factors such as skills, experience, and geographic location. This role is also eligible for a competitive benefits package that includes: medical, dental, vision, life, and disability insurance; 401(k) retirement plan; flexible spending & health savings account; at least 12 paid holidays; paid time off; parental leave; employee assistance program; and other company benefits.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

1:33 min

Integrating internal APIs and maintaining data sovereignty

Mahran Meißner Mahran Meißner · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all