Site Reliability Engineer

ADS, inc
San Francisco, CA, United States
4 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$190,800.0 - $267,100.0
Working hours
Regular working hours
Job source

Tech stack

Cloud Engineering Distributed Systems Monitoring of Systems Reliability Engineering Software Engineering Data Logging

Job description

The Ads Reliability team partners closely with Ads Engineering to improve reliability, scalability, operational excellence, and developer productivity across Reddit’s advertising ecosystem. We help build and operate highly available services that drive revenue and maintain advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate, and scale the critical systems behind Reddit Ads.

What you’ll do:

  • Partner with Ads Engineering teams to improve reliability, scalability, and operational excellence of ad-serving, auction, targeting, measurement, and billing systems.
  • Design, build, and maintain infrastructure, tooling, and automation that improve service reliability and engineering productivity.
  • Improve observability through monitoring, alerting, tracing, logging, and dashboards.
  • Participate in on-call rotations and lead incident response efforts for critical production systems.
  • Run root cause analysis and drive corrective actions following incidents.
  • Collaborate with software engineers throughout the service lifecycle, from design reviews through production operations.
  • Drive adoption of SRE best practices including SLIs, SLOs, error budgets, capacity planning, and operational readiness reviews.
  • Reduce operational toil through automation and self-service tooling.
  • Help define and measure advertiser-critical user journeys such as campaign creation, ad delivery, reporting, and billing.
  • Scale Ads systems to support continued traffic growth, increased advertiser demand, and evolving business requirements.

Requirements

  • 5+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or related roles operating large scale distributed systems.
  • Strong experience supporting high traffic, user facing production environments.
  • Strong cross-functional collaborations skills to lead and influence operational excellence.
  • Good understanding of modern distributed systems, scale engineering, and cloud-native architectures.
  • Strong software engineering skills in languages like general-purpose backend languages like Go.
  • Demonstrated ability to troubleshoot complex issues across applications, infrastructure, networking, and services.
  • Experience with observability platforms, monitoring systems, alerting, and incident response.
  • Experience driving automation and operational improvements.

Benefits & conditions

  • Comprehensive Health benefits
  • 401k Matching
  • Workspace benefits for your home office
  • Personal & Professional development funds
  • Family Planning Support
  • Flexible Vacation & Reddit Global Days Off
  • 4+ months paid Parental Leave
  • Paid Volunteer time off

Pay Transparency:

This job posting may span more than one career level.

In addition to base salary, this job is eligible to receive equity in the form of restricted stock units, and depending on the position offered, it may also be eligible to receive a commission. Additionally, Reddit offers a wide range of benefits to U.S.-based employees, including medical, dental, and vision insurance, 401(k) program with employer match, generous time off for vacation, and parental leave. To learn more, please visit https://www.redditinc.com/careers/.

To provide greater transparency to candidates, we share base salary ranges for all US-based job postings regardless of state. We set standard base pay ranges for all roles based on function, level, and country location, benchmarked against similar stage growth companies. Final offer amounts are determined by multiple factors including, skills, depth of work experience and relevant licenses/credentials, and may vary from the amounts listed below. The base salary range for this position is: $190,800-$267,100 USD

About the company

Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com.

The Ads organization powers Reddit’s advertising platform, enabling advertisers to reach highly engaged communities while helping Reddit grow its business. The reliability of our Ads systems directly impacts advertiser success, revenue generation, and user experience.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:10 min

Exposing sensitive information through partial search logs

Dennis Schulz Dennis Schulz +1 ¡ World Congress 2026 Europe

1:23 min

Closing thoughts and educational resources for edge network engineering

Austin Gil ¡ LIVE

1:29 min

Overcoming challenges in AI-assisted distributed system development

Przemysław Ładyński Przemysław Ładyński · World Congress 2026 Europe

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley ¡ World Congress 2021

1:56 min

Discovering incidents using logs, metrics, and traces

Nele Uhlemann ¡ World Congress 2023

8:02 min

Integrating service level objectives into incident management

Diana Todea ¡ LIVE

Videos

See all

Related articles

See all