Senior Director, Platform Software Engineering

Oracle
United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computing Platforms Cloud Computing DevOps Network Control Oracle (Applications) Performance Tuning Cloud Services Service-Oriented Architecture Software Deployment Software Engineering System Software
+1 more
Oracle Cloud Infrastructure

Job description

As the Director of Software Engineering in the IoT and Connected Devices team, you will provide technical leadership to the teams responsible for building and operating Oracle Cloud Infrastructure services that support connected devices across many industries. This role requires deep experience building cloud services within OCI, not only deploying applications on top of OCI. You should be familiar with OCI service architecture patterns, operational standards, control plane and data plane considerations, service ownership expectations, and OCI release and deployment processes.

By fostering a high-performance culture, you will mentor and develop team members, creating an environment that encourages continuous learning and professional growth. You are an OCI service owner and will work closely with product management, architects, DevOps, security, and operations teams to ensure customer success.

As Director of Engineering, you will own and drive the operations, architecture, and delivery of critical OCI services supporting IoT and Connected Devices. You will oversee customer experience, system stability, reliability, availability, performance, and scalability. You will lead incident response, root-cause analysis, and adoption of best practices for operating production-grade OCI services.

You will provide feature and architectural leadership, set strategic direction, and ensure the team builds services using OCI-native service architecture patterns. You will guide teams through OCI release processes, production readiness requirements, and safe deployment practices. You will define metrics and KPIs, lead community engagement, and be responsible for talent development, organizational growth, and execution excellence.

Key Responsibilities

Service Ownership & Operations

Own and drive the service’s operations, overseeing customer experience, system stability, reliability, availability, performance, and scalability. Lead incident response, root-cause analysis, operational reviews, and adoption of best practices for running production-grade OCI services. Ensure the team meets OCI service ownership expectations, including monitoring, alerting, operational readiness, security posture, compliance, and continuous improvement.

OCI-Native Service Architecture & Engineering Leadership

Lead the design, development, and evolution of services built within OCI. Guide teams in applying OCI service architecture patterns, including service decomposition, tenancy and identity integration, control plane and data plane design, regional deployment models, resilience patterns, observability, security, and operational automation. Partner with architects and senior engineers to drive architectural improvements, technical roadmaps, platform enhancements, and long-term service strategy.

Release, Deployment & Operational Readiness

Ensure engineering teams follow OCI release processes, deployment standards, change management practices, and operational readiness requirements. Drive disciplined execution around staged rollouts, regional releases, service validation, rollback planning, production readiness reviews, and post-release monitoring. Promote safe, repeatable, and auditable release practices across multiple sprint teams.

Feature & Roadmap Delivery

Guide and coordinate new feature development across multiple sprint teams. Align feature delivery with product strategy, customer needs, architectural direction, and operational requirements. Remove roadblocks, manage dependencies, and ensure teams deliver high-quality capabilities while maintaining service reliability and operational excellence.

Metrics & KPIs

Define and track KPIs and qualitative goals that measure the team’s impact and align with Connected Devices business objectives. Use data-driven approaches to prioritize work, monitor progress, improve service health, and track tactical and strategic deliverables. Establish metrics for reliability, release quality, operational performance, customer experience, engineering productivity, and team execution.

Talent Development & Leadership

Inspire, mentor, and retain top engineering talent by setting clear goals, providing actionable feedback, and supporting career growth. Guide frontline managers in establishing objectives, tracking performance, and developing team members in alignment with business and technical needs. Recruit, develop, and lead senior individual contributors and managers capable of building and operating critical OCI services.

Project Management & Execution

Ensure strong execution discipline across engineering initiatives. Own and communicate project timelines, priorities, risks, and dependencies. Drive continuous progress on service improvements, architectural investments, feature delivery, and operational programs. Ensure teams balance innovation, execution speed, reliability, and long-term maintainability.

Requirements

  • BS or MS degree in a relevant field, or equivalent experience.
  • Minimum of 10 years of people management experience, with a proven track record managing large-scale, critical cloud services.
  • Demonstrated experience building services within OCI or a comparable hyperscale cloud infrastructure environment, not only consuming cloud services as an application platform.
  • Familiarity with OCI service architecture patterns, including service ownership models, control plane and data plane architecture, identity and access integration, regional deployment models, reliability patterns, observability, and operational automation.
  • Experience with OCI release processes, production readiness expectations, change management, staged deployments, regional rollouts, rollback planning, and post-release validation.
  • Demonstrated expertise in strategic planning, resource management, service reliability, and performance optimization within a technology-driven environment.
  • Exceptional decision-making abilities, with experience influencing senior leadership, architects, product management, and cross-functional engineering teams.
  • Strong background in recruiting, developing, and leading senior individual contributors and frontline managers.
  • Ability to make decisions that significantly impact organizational infrastructure, customer experience, revenue, and long-term operational success.
  • Experience in networking and cloud infrastructure operations.
  • Experience managing large-scale critical services in a cloud computing environment.
  • Strong understanding of operational excellence practices, including incident management, root-cause analysis, service health reviews, capacity planning, monitoring, alerting, and continuous improvement.

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on eeho.fa.us2.oraclecloud.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:41 min

Transitioning artificial intelligence infrastructure into scalable commodity cloud services

juarezjunior juarezjunior · World Congress 2024

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · World Congress 2024

2:14 min

Solving complex platform architecture challenges at an enterprise scale

Maria Apazoglou · Coffee With Developers

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all