Software Engineer (SRE)

Visa
Basingstoke, UK
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

HTML Java (Programming Language) JavaScript (Programming Language) Microsoft Windows JIRA Microsoft Azure Cloud Computing Computer Programming Continuous Integration Linux Distributed Systems Github
+35 more
Information Technology Operations JSON Python (Programming Language) PostgreSQL MongoDB MySQL Octopus Deploy Windows PowerShell Release Management Reliability Engineering Ansible Prometheus Software Engineering Extensible Markup Language (XML) YAML Datadog Data Logging Scripting Cloud Platform System System Availability Grafana Cloudformation Containerization Kubernetes Sentry Non-relational Database Bitbucket Terraform Splunk New Relic (SaaS) Dynatrace Docker Jenkins Artifactory Golang

Job description

Site Reliability Engineering (SRE) is essential to Visa’s Cloud platform strategy. In this role, you’ll ensure our development platform and tools let engineers focus on innovation instead of infrastructure. You’ll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands-on expertise is required, especially with major DevTools like GitHub, Jenkins, Jira, and Artifactory., * DevTools Support You will be the primary point of contact for developers using tools like GitHub, Jenkins, Jira, or Artifactory.

  • Troubleshoot and resolve tool-related issues promptly to minimize developer downtime.
  • Maintain and optimize CI-CD pipelines and integrations for reliability and scalability.
  • Collaborate with development teams to improve workflows and automation.
  • Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance.
  • Monitor and improve application reliability through proactive measures and incident response.
  • Develop and maintain observability solutions (metrics, logging, tracing).
  • Participate in on-call rotations and drive root cause analysis for incidents.
  • Collaboration & Continuous Improvement Partner with engineering teams to identify reliability risks and implement best practices.
  • Document processes, troubleshooting guides, and reliability of playbooks.
  • Advocate for automation and self-service solutions to reduce operational overhead.

This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager. Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.

Requirements

Do you have experience in XML?, Do you have a Bachelor’s degree?, * Bachelor’s degree, OR 3+ years of relevant work experience, * Bachelor’s degree, OR 3+ years of relevant work experience

  • Bachelor’s degree in IT, CS or related field and-or 3+ Years Working Experience IT Operations and Delivery.
  • Experience: 3 years in SRE and-or DevTools support roles.
  • Beginner level programming and-or scripting in 2 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation.
  • Basic understanding of YAML, JSON, HTML, XML.
  • Hands on experience in Linux and -or Windows systems and good understanding of distributed computing environments.
  • 2 years experience with CI-CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Azure DevOps in a large-scale environment
  • 2 years experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry, etc. in a large-scale environment
  • 2 years experience supporting relational and non-relational databases (MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scaling
  • 2 or more years working in a Platform, SRE or Production Engineering group for high availability-critical platforms-applications
  • Experience managing a distributed container platform including but not limited to deployment-release management, provisioning, capacity management, workload management
  • Experience managing container infrastructure and supporting development transformation to a container first model.
  • This role requires oncall support as the team provides 24-7 operational support.
  • Technical Expertise: Proficiency in at least one DevTool (GitHub, Jenkins, ArgoCD, Jira, Artifactory, ).
  • Strong understanding of CI-CD principles and pipelines.
  • Solid knowledge of Linux systems, networking, and containerization (Docker-Kubernetes).
  • Hands-on experience with cloud platforms.
  • Programming-Scripting: Proficiency in Python, Ansible, or similar languages.
  • Mindset: Strong problem-solving skills, systems thinking, self-starter, and a passion for reliability.

About the company

Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.

At Visa, you’ll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.

Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

2:21 min

Projecting external HTML content using default and named slots

Rowdy Rabouw Rowdy Rabouw · WWC 2022

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

1:38 min

Managing and versioning system prompts as YAML files

Kevin Lewis Kevin Lewis +1 · WWC 2025

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

6:12 min

Streaming HTML content natively using declarative processing instructions

Chris Heilmann +2 · LIVE

Videos

See all

Related articles

See all