Site Reliability Engineer- Product Reliability Engineering
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+9 more
Job description
This role is focused on maintaining, improving, and scaling the reliability of production systems. You will work across engineering and operations teams to monitor system health, respond to incidents, automate operational workflows, and proactively identify areas for improvement.
This position is ideal for engineers who enjoy hands-on production support, troubleshooting complex systems, and building automation to improve reliability and efficiency.
The Work Itself:
Production Reliability & Incident Management
Support the reliability and availability of production applications, products, and services. Actively participate in incident response, troubleshooting, and SWAT calls to restore service quickly and minimize impact.
Proactive Monitoring and Problem Prevention
Perform ongoing analysis of system performance and reliability to identify potential issues before they impact users. Drive discussions and solutions to improve system stability and resilience.
Automation and Operational Efficiency
Build and maintain automation to reduce manual work, improve response times, and increase system reliability. Focus on eliminating repetitive operational tasks through scripting and tooling.
System Reliability Improvements
Partner with engineering and product teams to implement solutions that improve scalability, performance, and system efficiency. Support deployments, upgrades, and production changes with a reliability-first mindset.
Observability and Monitoring
Contribute to improving monitoring, alerting, and observability across systems to ensure visibility into performance and failures.
Documentation and Knowledge Sharing
Develop and maintain clear documentation, runbooks, and knowledge repositories to support operational excellence and team collaboration.
Emerging Technology & AI Enablement
Contribute to the adoption of AI/automation tools to improve troubleshooting, incident response, and system reliability.
*This is a hybrid position. Expectation of days in office will be confirmed by your hiring manager.
- This role is not eligible for employment-based be authorized to work in the United States without current or future sponsorship.
Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.
Requirements
- 2+ years of relevant work experience and a Bachelor’s degree, OR 5+ years of relevant work experience., * 2-5 years of experience in Site Reliability Engineering, DevOps, or production operations roles
- Experience supporting production systems and participating in incident response/on-call rotations
- Working knowledge of one or more programming or scripting languages (Python, Java, C#, PowerShell, Bash)
- Experience with automation of operational tasks and processes
- Understanding of Linux/Unix systems
- Understanding of networking fundamentals (protocols, architecture)
- Familiarity with monitoring and observability tools
- Strong problem-solving skills in real-time production environments
- Familiarity with AI/ML tooling for operational efficiency
- Experience with CI/CD pipelines
- Experience with cloud environments (AWS, Google Cloud Platform, or Azure)
- Exposure to infrastructure as code or configuration management tools
- Experience using Git/GitHub for version control
Benefits & conditions
The estimated salary range for this position is $110,700.00 to $ 171,800.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program.
About the company
Visa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.
At Visa, you’ll have the opportunity to create impact at scale - tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.
Join Visa and do work that matters - to you, to your community, and to the world. Progress starts with you.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Is Software Engineering Over-Saturated?
Highest Paying Tech Companies for Developers
The 12 Best Jobs for Software Engineers
Best Countries for Software Engineers