Manager, Cloud Engineering & Operations

SS8 Networks, Inc.
United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$175,000.0 - $210,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Big Data Cloud Computing Cloud Engineering Databases Continuous Integration Customer Data Management Linux DevOps Disaster Recovery
+24 more
Distributed Data Store Distributed Systems Fault Tolerance Identity and Access Management Reliability Engineering Site Reliability Engineering Practices Software Engineering Software Vulnerability Management AI Infrastructure Private Cloud Environment SSL Certificate Management Data Logging Google Cloud Cloud Platform System System Availability Large Language Models Mttr Multi-Cloud HybridCloud Containerization Kubernetes Infrastructure Automation Frameworks Stream Processing Oracle Cloud Infrastructure

Job description

SS8 is seeking an experienced Manager, Cloud Engineering & Operations to lead the architecture, engineering, deployment, and continuous improvement of SS8’s cloud and cloud-native infrastructure. This is a hands-on technical leadership and people-management role with end-to-end accountability for both Cloud Engineering and Cloud Operations. You will lead a team responsible for building and operating highly available, secure, scalable, and resilient infrastructure supporting SS8’s mission-critical solutions across public, private, hybrid, sovereign, customer-managed, and restricted or air-gapped environments.

The ideal candidate is a player-coach who combines strong people leadership with deep technical credibility and operational accountability - equally comfortable conducting an architecture review with senior engineers, leading a major production incident, reviewing operational KPIs, and discussing cloud strategy with executive leadership. You should be able to create a culture in which reliability, security, automation, and operational ownership are engineering responsibilities from day one, not activities that begin after development is complete.

What You’ll Work On

  • Owning the architecture, engineering strategy, and technical roadmap for SS8’s cloud and cloud-native infrastructure, including public/private/hybrid cloud, Kubernetes, distributed infrastructure, and disaster recovery.
  • Establishing a disciplined Cloud Operations function covering 24x7 availability, incident response, root-cause analysis, change management, and capacity/performance management.
  • Driving DevOps and SRE practices across the organization, including Infrastructure as Code, CI/CD, and automation of provisioning, monitoring, and remediation.
  • Building comprehensive observability across infrastructure, platforms, and applications - metrics, logging, tracing, and security telemetry
  • Developing infrastructure and operational models that let SS8 products run consistently across AWS, Azure, GCP, OCI, private cloud, customer data centers, sovereign, and air-gapped environments.
  • Identifying opportunities to bring AI and automation into Cloud Operations, including intelligent alert correlation, automated triage, and anomaly detection., * Set technical direction and remain technically engaged through architecture reviews, design reviews, proof-of-concept work, and critical engineering initiatives - following the full lifecycle from architecture through build, automation, deployment, operation, monitoring, recovery, and continuous improvement.
  • Define operational ownership across Cloud Engineering, Software Engineering, Support, Professional Services, and Security; establish SLIs, SLOs, and KPIs, and own measurable outcomes including platform availability and reliability, MTTD/MTTA/MTTR improvement, change failure rate, deployment consistency across customer environments, and reduction in customer-impacting incidents.
  • Build, lead, mentor, and develop a high-performing Cloud Engineering & Operations team, including hiring, performance management, resource planning, and succession planning - with success measured in part by the team’s growth, retention, and ability to independently own major elements of the platform.
  • Partner with Security and Product teams to embed encryption, IAM, secrets/certificate management, hardening, and vulnerability management into engineering and operational processes, tracking security and vulnerability remediation as a core outcome.
  • Establish proactive capacity planning and forecasting and optimize cloud consumption and cost without compromising reliability or performance.
  • Partner with Product Management, Software Engineering, Security, QA, Professional Services, and Technical Program Management, translating requirements into cloud engineering priorities and communicating infrastructure strategy, operational risk, and major incidents to leadership - while ensuring new product releases launch with strong operational readiness and the overall cloud technology roadmap executes successfully.

Requirements

  • 10+ years of experience in cloud infrastructure, platform engineering, software engineering, distributed systems, DevOps, SRE, or Cloud Operations.
  • 5+ years of engineering management or significant technical leadership experience.
  • Demonstrated ability to lead engineers while remaining technically engaged and hands-on when needed.
  • .Experience owning production infrastructure and mission-critical services.
  • Strong architecture and design experience with large-scale distributed systems, including Kubernetes and container platforms.
  • Experience with Infrastructure as Code, automated provisioning, and CI/CD pipelines.
  • Strong knowledge of Linux, networking, compute, storage, databases, and distributed infrastructure.
  • Experience implementing observability, monitoring, logging, alerting, and telemetry platforms.
  • Strong production incident-management and troubleshooting experience, including high availability, fault tolerance, disaster recovery, and capacity planning.
  • Experience with at least one major cloud platform (AWS, Azure, GCP, or OCI).
  • Strong understanding of cloud and infrastructure security principles and best practices.
  • Ability to operate effectively across Engineering, Security, Product, Operations, and customer-facing teams.

Nice to Have

  • Multi-cloud, private/hybrid, or sovereign cloud architecture experience.
  • Experience supporting government or national-security environments and air-gapped deployments.
  • Large-scale data platforms, distributed databases, and real-time data processing
  • FedRAMP or comparable government compliance experience.
  • AI infrastructure, LLM/RAG infrastructure, AI Ops, or intelligent operations experience.

Benefits & conditions

The expected base salary range for this position is $175,000 - $210,000. Actual compensation will be determined based on the candidate’s skills, experience, and qualifications. This role is also eligible to participate in SS8’s corporate bonus program, subject to the terms of the applicable plan. We offer a comprehensive benefits package including medical, dental, vision, 401(k) with company match, and paid time off.

At SS8 Networks, our approach to flexible work is centered on trust and optimized for culture, connection, clarity, and the evolving needs of our business. The work location of this role is Milpitas (Hybrid), meaning it will be performed [both from home and from an SS8 office on select days / from home], as determined by the business needs of the team.

SS8 does not accept unsolicited resumes from staffing agencies, search firms, or third parties. Any resumes submitted without a signed agreement in place will be considered the property of Company, and no fees will be paid if a candidate is hired as a result.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · World Congress 2021

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all