Software Engineer

Visa Inc.
Austin, TX, United States
about 1 month ago

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
0 years minimum
Compensation
$88,000.0 - $136,900.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Systems Engineering Microsoft Azure Bash Shell Cloud Computing Information Systems Computer Programming Continuous Integration Software Debugging
+23 more
DevOps Distributed Systems Github Monitoring of Systems Issue Tracking Systems Python (Programming Language) Reliability Engineering Site Reliability Engineering Practices Prometheus Software Engineering Datadog Scripting Cloud Platform System System Availability Grafana Reliability of Systems Kubernetes Information Technology Machine Learning Operations Splunk Docker Jenkins Golang

Job description

Join a high-impact team focused on improving reliability, operational excellence, and developer productivity through automation, AI-enabled tooling, and secure-by-design engineering practices. We build and support scalable platforms, intelligent workflows, and modern operational solutions that help teams design, deploy, monitor, and maintain resilient systems.

Our mission is to simplify operations, reduce manual effort, improve system health, and enable engineers to move faster and safer-using automation, observability, and emerging AI technologies.

Role Overview

As a Site Reliability Engineer (SRE), you will help support and improve the reliability, scalability, and operational efficiency of critical systems and engineering platforms. You will work closely with senior engineers and cross-functional teams to monitor services, automate repetitive tasks, improve incident response, and contribute to AI-assisted operational workflows.

This role is ideal for someone who is passionate about system reliability, automation, cloud/platform operations, observability, and the growing role of AI in engineering and operations. You will have the opportunity to learn modern SRE practices while contributing to platforms and tools that improve both developer experience and production resilience., Reliability Engineering & Operations

  • Support the reliability, availability, and performance of production systems and internal engineering platforms.
  • Monitor system health, investigate alerts, and assist in troubleshooting incidents and service disruptions.
  • Participate in incident response, issue tracking, root cause analysis, and post-incident reviews.
  • Help improve operational readiness by maintaining runbooks, dashboards, alerts, and support documentation.

Automation & Platform Support

  • Contribute to automation efforts that reduce manual operational work and improve consistency across environments.
  • Assist in building and maintaining scripts, tools, and workflows for deployment, monitoring, remediation, and system maintenance.
  • Support CI/CD pipelines, environment stability, and platform reliability initiatives.
  • Help improve observability through logs, metrics, dashboards, and tracing.

AI-Enabled Operations

  • Work with senior engineers to adopt AI-assisted tools and workflows that improve troubleshooting, alert analysis, documentation, and operational efficiency.
  • Support the use of AI for incident summarization, knowledge discovery, automation recommendations, and operational insights.
  • Help validate and safely use AI-generated suggestions in engineering and operational tasks.
  • Learn how AI can enhance reliability engineering while maintaining strong security, governance, and quality standards.

Collaboration & Learning

  • Partner with software engineers, platform teams, security teams, and operations teams to support reliable service delivery.
  • Learn SRE best practices for system design, monitoring, incident management, and automation.
  • Contribute to team documentation, operational reviews, and continuous improvement efforts.
  • Seek mentorship from senior engineers and actively grow technical skills in cloud, automation, observability, and AI-powered operations.

What You Will Learn

  • Fundamentals of Site Reliability Engineering
  • Production monitoring, alerting, and incident response
  • Automation and scripting for operational efficiency
  • CI/CD and platform reliability practices
  • Observability using metrics, logs, traces, and dashboards
  • AI-assisted engineering and operations workflows
  • Reliability, security, and compliance best practices in enterprise environments

This is a hybrid position. Expectation of days in the office will be confirmed by your Hiring Manager. Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.

Requirements

Basic Qualifications * Bachelor’s degree, OR 3+ years of relevant work experience Preferred Qualifications * 2 or more years of work experience * Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field, or equivalent practical experience * 0-2 years of experience in software engineering, systems engineering, cloud operations, DevOps, SRE, or technical support * Basic understanding of Linux/Unix systems, networking, APIs, and distributed systems concepts * Familiarity with at least one programming or scripting language such as Python, Java, Go, or Bash * Interest in automation, cloud platforms, monitoring tools, and operational excellence * Curiosity about AI/ML tools and how they can improve engineering workflows and reliability operations * Strong problem-solving, communication, and teamwork skills * Exposure to cloud platforms such as AWS, Azure, or GCP * Familiarity with monitoring and observability tools such as Prometheus, Grafana, Splunk, Datadog, or similar * Understanding of CI/CD concepts and tools such as GitHub Actions, Jenkins, or similar * Exposure to containers and orchestration technologies such as Docker and Kubernetes * Internship, academic, or project experience related to automation, infrastructure, reliability, or platform engineering * Interest in using AI tools to improve debugging, knowledge sharing, and operational workflows

Benefits & conditions

12401 Research Blvd Bldg Ii, Austin, TX 78759 Hybrid work $88,000 - $136,900 a year - Full-time, Pulled from the full job description

  • 401(k)
  • Health insurance
  • Paid time off
  • Vision insurance
  • Health savings account
  • Dental insurance
  • Flexible spending account, The estimated salary range for this position is $88,000.00 to $ 136,900.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program.

About the company

Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.

At Visa, you’ll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.

Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all