Senior Site Reliability Engineer

Fiserv, Inc.
Sunnyvale, CA, United States
17 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$160,000.0 - $240,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Cloud Computing Configuration Management DevOps Github HAProxy HTTP Secure Python (Programming Language) Routing Reliability Engineering Cloud Services Ansible
+12 more
Prometheus Shell Script Datadog Data Logging Load Balancing Delivery Pipeline Grafana Containerization Kubernetes Puppet Terraform Software Version Control

Job description

You will join our global team in Sunnyvale and help operate financial platforms at scale. You will partner with cross-functional teams to improve reliability, automate operations, and drive continuous improvement across our cloud-native environments.

What you will do:

  • Design, build and maintain automation to eliminate manual, repetitive operational tasks (runbooks, deployment pipelines, remediation scripts).
  • Operate and enhance monitoring, logging and alerting systems to ensure strong observability across services.
  • Participate in on-call rotations and lead incident response activities; run and document post-incident RCA and follow-up actions.
  • Collaborate with stakeholders to define SLIs and SLOs, manage error budgets and translate reliability goals into measurable actions.
  • Forecast capacity needs and contribute to resource planning to ensure performance and cost-efficiency.
  • Troubleshoot production issues: deep-dive analysis, isolate root causes and implement durable fixes.
  • Drive continuous improvement and platform hardening through runbook improvements, automation, and best-practice adoption.
  • Work closely as part of an international team to deliver project outcomes and operational excellence.

Requirements

  • Solid practical experience in site reliability, operations or DevOps at a mid-to-senior level.
  • Strong shell scripting skills and a foundation in programming concepts.
  • Hands-on experience with cloud workloads-specifically Google Cloud Platform (GCP) and GKE.
  • Proven experience with containerisation and orchestration (Kubernetes).
  • Working knowledge of Infrastructure as Code and configuration management (Terraform, Ansible, Puppet).
  • Familiarity with monitoring and observability tooling such as Prometheus, Grafana and Datadog.
  • In-depth understanding of HTTP(s) traffic, routing and load-balancing, with practical experience observing and operating HAProxy.
  • Comfortable using GitHub and GitHub Actions for code management, automation and IaC pipelines.
  • Strong troubleshooting skills, a pragmatic problem-solving approach and effective communication for cross-team collaboration.

What would be great to have:

  • Experience programming in Python, Go or Java.
  • Exposure to large-scale financial services platforms or highly regulated environments.
  • Experience defining SLIs/SLOs and managing error budgets in production environments.

About the company

We’re Fiserv, a global leader in Fintech and payments, and we move money and information in a way that moves the world. We connect financial institutions, corporations, merchants and consumers to one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit card, pay through a mobile app, or withdraw money from the bank, we’re involved. If you want to make an impact on a global scale, come make a difference at Fiserv.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

2:26 min

Understanding Puppeteer and its underlying architectural design

Miki Lombardi · JS Congress

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all