SW Engineer- Developer Systems Reliability Engineering

Visa Inc.
Austin, TX, United States
16 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$88,000.0 - $136,900.0
Working hours
Regular working hours

Tech stack

HTML Java (Programming Language) JavaScript (Programming Language) Microsoft Windows Artificial Intelligence Microsoft Azure Cloud Computing Computer Programming Continuous Integration Linux Distributed Systems Github
+41 more
Hyper-V Infrastructure as a Service (IaaS) Information Technology Operations JSON Python (Programming Language) PostgreSQL MongoDB MySQL Octopus Deploy OpenShift Platform as a Service (PAAS) Windows PowerShell Release Management Reliability Engineering Ansible Prometheus Runbook Software Engineering Virtualization Technology Extensible Markup Language (XML) YAML Datadog Scripting Cloud Platform System GitHub Copilot System Availability Grafana Reliability of Systems Cloudformation Sentry Non-relational Database Bitbucket Terraform Splunk New Relic (SaaS) GPT Dynatrace Jenkins Artifactory Golang Vmware

Job description

TheSite Reliability Engineering - SRE - role is a critical part of Visa’s Cloud Platform strategy. In this role, you will help ensure Visa’s development platform and tooling enable key stakeholders, including development engineers across the technology organization, to focus more on innovation and less on infrastructure management. You will drive the adoption of observability best practices, automation, and AI-AIOps capabilities to improve reliability, reduce toil, and resolve recurring operational issues.

You must be comfortable partnering with software engineering teams and supporting their needs to ensure the security, availability, reliability, and performance of the platform. These software engineering teams are both peer engineering teams supporting the platform and customers within Visa Engineering consuming the platform. This engineer will be expected to triage complex issues, collaborate with peer infrastructure and operations teams, enhance monitoring and alerting coverage, and operationalize automation and AI-driven solutions that improve incident response, service health, and platform efficiency.

This is a hands-on engineering role with a strong focus on building and advancing reliability engineering practices for theVisa Cloud Platform.

Essential Functions:

  • Help maintain the platform’s defined SLAs and SLOs by driving operational excellence, delivering value-added process and procedure improvements, and partnering with engineering and operations teams to eliminate manual touchpoints through automation and standardization.
  • Own and operationalize end-to-end observability, alerting, and monitoring for the Visa Cloud Platform across IaaS, PaaS, and Container-as-a-Service environments, ensuring telemetry, dashboards, alerts, SLIs, and operational workflows are meaningful, actionable, and effective in supporting production reliability.
  • Own and deliver automation and AI-AIOps initiatives that reduce toil, improve reliability, accelerate incident response, and enhance operational intelligence across the SRE organization.
  • Partner with development teams during release and service transition reviews to define and validate operational requirements, including SLIs-SLOs, monitoring, alerting, dashboards, runbooks, incident response procedures, capacity expectations, and production readiness criteria.
  • Partner with Operations & Infrastructure peers to support ongoing platform maintenance, enhancement, and reliability.
  • Support multiple internal stakeholders across a variety of technical challenges by analyzing recurring issues, identifying patterns, and proposing effective solutions.
  • Support the Visa Cloud SRE team’s 24-7-365 operating model, including shift-based and on-call coverage, with weekend support as required.

Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.

Requirements

  • Bachelor’s degree, OR 3+ years of relevant work experience, * Bachelor’s degree, OR 3+ years of relevant work experience
  • 2 or more years working in a Platform, SRE or Production Engineering group for high availability-critical platforms-applications
  • Bachelor’s degree in IT, CS or related field and-or 3+ Years Working Experience in IT Operations and Delivery.
  • 2 years experience with CI-CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Bitbucket, Azure DevOps in a large-scale environment
  • 2 years experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry, etc. in a large-scale environment
  • 2 years experience supporting relational and non-relational databases [MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scaling
  • Basic understanding of YAML, JSON, HTML, XML.
  • Hands on experience in Linux and -or Windows systems and good understanding of distributed computing environments.
  • Beginner level programming and-or scripting in 3 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation.
  • Experience with AI-enabled solutions and tools such as Claude, ChatGPT, GitHub Copilot, or similar technologies, with practical application in automation, observability, incident response, or operational workflows.
  • Experience managing container infrastructure and supporting development transformation to a container first model
  • Exposure to Virtualization (Hyper-V, VMware, Openshift Virtualization etc.)
  • Experience managing a distributed container platform including but not limited to deployment-release management, provisioning, capacity management, workload management

Benefits & conditions

$88,000.00 to $ 136,900.00 life insurance, paid time off, 401(k) United States, Texas, Austin Jul 25, 2026, The estimated salary range for this positionis $88,000.00 to $ 136,900.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program.

About the company

Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.

At Visa, you’ll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.

Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on diversityjobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

2:21 min

Projecting external HTML content using default and named slots

Rowdy Rabouw Rowdy Rabouw · WWC 2022

1:38 min

Managing and versioning system prompts as YAML files

Kevin Lewis Kevin Lewis +1 · WWC 2025

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · WWC 2025

Videos

See all

Related articles

See all