Systems Operations Senior Manager - Cloud Platform Operations

Wells Fargo
Charlotte, NC, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Kubernetes Security Systems Engineering Cloud Computing Cloud Engineering Configuration Management Continuous Integration Information Engineering Data Infrastructure Disaster Recovery OpenShift Data Processing Google Cloud
+10 more
Data Storage Technologies Cloud Platform System Apache Spark Containerization Infrastructure Automation Frameworks Storage Technologies Performance Monitor Data Management Terraform Serverless Computing

Job description

  • Manage and develop teams of managers, engineers, and analysts providing operational support for Google Cloud Platform (GCP) platforms and capabilities, and On-premise cloud-native services such as Spark-based analytics workloads and Iceberg storage architectures on OpenShift container platforms.
  • Own the operational health, availability, scalability, and performance of GCP and OCP platforms.
  • Engage and influence application, data engineering, architecture, security, and infrastructure partners to design, onboard, and operate resilient cloud-native solutions
  • Define and execute operational strategies to support moderate to high-risk initiatives, including platform modernization, container adoption, and large-scale Spark workload enablement
  • Lead administration, lifecycle management, and capacity planning for OpenShift clusters, Spark platforms, and underlying GCP infrastructure
  • Drive infrastructure automation and standardization using Terraform, including environment provisioning, configuration management, and policy enforcement
  • Ensure effective incident, problem, and change management practices, including stakeholder communication for incidents, planned maintenance, and platform changes
  • Perform operational risk assessments, resiliency reviews, and security evaluations in partnership with cybersecurity and risk teams
  • Interpret, develop, and enforce cloud and platform operations policies, procedures, and standards aligned with enterprise risk and compliance requirements
  • Provide implementation support for key resiliency, disaster recovery, and data protection initiatives related to cloud-native and data platforms
  • Collaborate with and influence professionals at all levels to improve operational maturity, automation adoption, and platform reliability
  • Manage allocation of people, vendor services, and financial resources to meet platform service-level objectives and business demand
  • Build and sustain a strong culture of operational excellence, automation-first mindset, accountability, and talent development aligned with cloud-native and data platform strategy, Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit’s risk appetite and all risk and compliance program requirements.

Requirements

  • 7+ years of experience in cloud platform operations, systems engineering, or technology architecture, with a strong focus on Google Cloud Platform (GCP)
  • 7+ years of Systems Engineering and Technology Architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
  • 3+ years of management or leadership experience
  • Hands-on experience managing OpenShift container platforms and operating cloud-native workloads at scale
  • Strong experience driving automation using Terraform for infrastructure provisioning and platform operations

Desired Qualifications:

  • Experience operating large, enterprise-scale OpenShift clusters on-prem and on GCP
  • Strong background in cloud-native operations, SRE principles, and operational risk management
  • Experience with container security, data platform resiliency, and disaster recovery
  • Proven ability to improve operational efficiency through Terraform, CI/CD integration, and automation frameworks
  • Experience supporting Spark-based data processing workloads and modern data storage technologies such as Iceberg

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

4:47 min

Exploring OpenShift and Red Hat Developer Sandbox resources

Markus Eisele Markus Eisele · WWC 2024

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

2:32 min

Overview of Terraform and Terraform Cloud features

Devlin Duldulao · LIVE

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all