Lead Software Engineer - Cloud Native Platforms (GCP)

Visa Inc.
Foster City, CA, United States
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Clean Code Principles Java (Programming Language) Automation of Tests Cloud Computing Cloud Engineering Profiling Software Quality Continuous Integration DevOps Distributed Systems Java Platform Enterprise Edition (J2EE) Fault Tolerance
+20 more
Java Virtual Machine (JVM) Reliability Engineering Secure Coding Software Engineering Software Systems Data Logging Google Cloud Cloud Platform System Performance Testing Cloud Monitoring System Availability Caching Kubernetes Low Latency Deployment Automation Api Design Restful APIs Terraform Docker Microservices

Job description

Experteer Overview As a Lead Software Engineer, you will design, build, and operate secure, scalable cloud-native platforms on Google Cloud Platform (GCP) to support high-volume, globally distributed payments. You will lead engineering excellence across application development, distributed systems, and production operations while mentoring others. You’ll collaborate with cross-functional teams to deliver enterprise-grade, reliable solutions that meet strict latency and availability targets. This role offers impact at scale and an opportunity to shape platform reliability and performance in a dynamic fintech environment. Compensation / Benefits * Design, develop, test, deploy, and operate scalable, secure, and highly available software solutions. * Build cloud-native apps on GCP using Kubernetes (GKE), containers, CI/CD, observability, and automation. * Design distributed systems supporting high transaction volumes with strict latency/availability goals. * Lead performance initiatives: throughput, latency, capacity planning, profiling, and scalability. * Implement modern deployment strategies: Canary, Blue-Green, feature flags, automated rollback. * Translate business requirements into resilient, maintainable technical solutions. * Lead architecture discussions, design reviews, and technical decision-making. * Develop high-quality code with secure coding, automated testing, reviews, and operational readiness. * Troubleshoot production issues, perform root-cause analysis, and drive reliability improvements. * Improve platform reliability, scalability, security, observability, and efficiency through automation. * Collaborate with Product, Architecture, Security, SRE, DevOps, and cross-functional teams. * Mentor engineers and contribute to a culture of ownership, innovation, and continuous learning. * Leverage agentic workflows to boost developer productivity, code quality, and testing efficiency. Tasks * 10+ years of relevant work experience with a Bachelor’s Degree, or 7+ years with advanced degree, or 4 years with PhD, or 13+ years total experience * Strong hands-on experience with Google Cloud Platform (GCP), including GKE, Cloud Monitoring, Cloud Logging, networking, and security services * Deep expertise in Java/J2EE, RESTful APIs, microservices, and distributed systems * Experience with Kubernetes, Docker, CI/CD, and modern software delivery practices * Proven ability to design high-throughput, low-latency applications for large transaction volumes with stringent SLAs/SLOs * Track record of performance improvements (throughput, latency, JVM tuning, caching, db performance) * Experience with performance testing, load/stress testing, benchmarking, and capacity planning * Experience with Canary/Blue-Green deployments, feature flags, progressive delivery, and automated rollback * Experience building fault-tolerant, active-active systems with resiliency and automated recovery * Strong knowledge of Infrastructure as Code (Terraform), deployment automation, and DevOps practices * Experience with observability (monitoring, logging, tracing, alerting) and operational analytics * Experience in production operations, incident management, root cause analysis, and operational excellence * API-first architecture design, secure service communications, and enterprise integration patterns * Ability to mentor engineers and influence technical direction across multiple teams * Experience leveraging AI-assisted development tools and automation to improve efficiency Key requirements * Medical, Dental, Vision * 401(k) * FSA/HSA * Life Insurance * Paid Time Off * Wellness Program

Requirements

throughput, latency, capacity planning, profiling, and scalability. * Implement modern deployment strategies: Canary, Blue-Green, feature flags, automated rollback. * Translate business requirements into resilient, maintainable technical solutions. * Lead architecture discussions, design reviews, and technical decision-making. * Develop high-quality code with secure coding, automated testing, reviews, and operational readiness. * Troubleshoot production issues, perform root-cause analysis, and drive reliability improvements. * Improve platform reliability, scalability, security, observability, and efficiency through automation. * Collaborate with Product, Architecture, Security, SRE, DevOps, and cross-functional teams. * Mentor engineers and contribute to a culture of ownership, innovation, and continuous learning. * Leverage agentic workflows to boost developer productivity, code quality, and testing efficiency. Tasks * 10+ years of relevant work experience with a Bachelor’s Degree, or 7+ years with advanced degree, or 4 years with PhD, or 13+ years total experience * Strong hands-on experience with Google Cloud Platform (GCP), including GKE, Cloud Monitoring, Cloud Logging, networking, and security services * Deep expertise in Java/J2EE, RESTful APIs, microservices, and distributed systems * Experience with Kubernetes, Docker, CI/CD, and modern software delivery practices * Proven ability to design high-throughput, low-latency applications for large transaction volumes with stringent SLAs/SLOs * Track record of performance improvements (throughput, latency, JVM tuning, caching, db performance) * Experience with performance testing, load/stress testing, benchmarking, and capacity planning * Experience with Canary/Blue-Green deployments, feature flags, progressive delivery, and automated rollback * Experience building fault-tolerant, active-active systems with resiliency and automated recovery * Strong knowledge of Infrastructure as Code (Terraform), deployment aaaaaaaaa

  • and DevOps practices * Experience with observability (monitoring, logging, tracing, alerting) and operational analytics * Experience in production operations, incident management, root cause analysis, and operational excellence * API-first architecture design, secure service communications, and enterprise integration patterns * Ability to mentor engineers and influence technical direction across multiple teams * Experience leveraging AI-assisted development tools and automation to improve efficiency Key requirements * Medical, Dental, Vision * 401(k) * FSA/HSA * Life Insurance * Paid Time Off * Wellness Program

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:15 min

Reversing the caching model for artifact delivery

Thijs Feryn Thijs Feryn · WWC Europe 2026

4:18 min

Prioritizing communication and structural awareness over strict tool mastery

Liam Hurrel +1 · WWC 2021

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all