Manager, Infrastructure

Polaris Inc.
Santa Clara, CA, United States
2 days ago
Apply on job-boards.greenhouse.io
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$265,000.0 - $332,000.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Amazon S3 Border Gateway Protocol Ubuntu (Operating System) Configuration Management Data Centers Linux Distributed Systems Apache Hadoop Hadoop Distributed File System Identity and Access Management Machine Learning
+14 more
Networking Basics Ansible Zabbix Aerospike Okta HybridCloud Kubernetes Low Latency Cassandra Bare Metal Apache Kafka Route53 Vertica Network Server

Job description

Manage day-to-day operations of four owned data centers and a global InfraOps team. Own capacity planning, hardware lifecycle, vendor/budget management, incident response, and runbook/alert hygiene. Oversee core stack (Linux fleet, FreeIPA, Ansible/Salt, MAAS, Zabbix), spine-leaf networking, and stateful data tiers (Aerospike/Kafka/ClickHouse/HDFS) under strict low-latency SLAs. Partner with security/compliance and run cloud-vs-colo evaluations., This is a hands-on manager role. There is no “scale up” button here - latency, packets-per-second, and procurement lead times are the job. The right person combines deep bare-metal operational instincts with the leadership presence to run a distributed, experienced global team from day one. Key ResponsibilitiesTeam Leadership & Operations

  • Manage and develop the InfraOps team across US and APAC time zones, including regional DC owners (SV/VA and NL/HK) and network engineering
  • Own the weekly DevOps check-in cadence, alert reviews, and 24/7 on-call coverage model
  • Drive P1/P2 incident response end to end - accountability for MTTR reduction, runbook coverage, and alert hygiene

Capacity Planning & Hardware Lifecycle

  • Own capacity planning and hardware lifecycle across all four data centers: Dell and Supermicro procurement through VARs, GPU expansion for on-prem ML training and inference, colo power and space management, and remote-hands logistics with Equinix and Digital Realty
  • Run the annual cloud-vs-colo evaluation alongside leadership, with full ownership of the recommendation

Technical Platform Oversight

  • Oversee the core infrastructure stack: Ubuntu/systemd fleet, FreeIPA, Ansible/Salt configuration management, MAAS provisioning, and Zabbix monitoring
  • Manage the spine-leaf Mellanox/NVIDIA network via Netris, including 100G Google peering and transit blend (Lumen/Cogent/Zayo)
  • Support the stateful data tier - Aerospike, Kafka, ClickHouse, Hadoop/HDFS - across capacity limits, evictions, migrations, and low-latency tuning

Vendor & Budget Management

  • Own colo and vendor relationships and budgets: Equinix and Digital Realty invoices, transit contracts, VAR procurement, and Netris licensing
  • Partner with the security/compliance function on SOC 2 Type 2 evidence, infrastructure hardening, and access reviews
  • Operate comfortably within a multi-entity environment (RZR/Skillz/Firy/Beamable shared IT) with comfort in M&A-flavored ambiguity

Requirements

  • 6-8 years in infrastructure or data center operations with 2+ years managing engineers - this role takes over a functioning global team on day one
  • Bare-metal and colo depth: capacity planning, hardware procurement (Dell/Supermicro), IBX/remote-hands workflows, and the physical logistics of running owned cages
  • Network fundamentals at scale: spine-leaf architecture, BGP/peering (100G-class), transit blends, and low-latency tuning (NIC/IRQ, packets-per-second thinking)
  • Deep Linux operations: systemd, netplan, FreeIPA/Chrony, Ansible and/or Salt, Zabbix, running fleets of hundreds-plus servers
  • Experience operating large stateful distributed systems - Aerospike, Cassandra, Scylla, Kafka, or ClickHouse - under sub-50ms latency budgets and hard capacity limits
  • Demonstrated P1/P2 incident ownership: on-call program management, postmortems, and alert hygiene discipline

Nice-to-Have

  • Hybrid cloud experience alongside owned metal: AWS (IAM, Route 53, GuardDuty, S3); Kubernetes exposure a plus
  • SOC 2 or compliance evidence experience; Okta and Vanta familiarity; security-minded infrastructure approach
  • Experience in adtech, RTB, or other high-QPS, latency-sensitive environments
  • Netris or other SDN controller experience; MAAS provisioning familiarity

Benefits & conditions

265K-332K Annually Senior level 265K-332K Annually Senior level Consumer Web * Healthtech * Professional Services * Social Impact * Software Lead development of a unified billing platform and ledger to ensure correct appointment billing, reliable reconciliation, and trusted financial datasets and reports. Own architecture and roadmap for ledger, reconciliation, and data-integrity systems; partner with Engineering, Data, Finance, Accounting, and Compliance to improve payment health, automate reconciliation, and enable accurate, timely close and audit readiness. CrowdStrike, Remote or Hybrid CA, USA 120K-180K Annually Expert/Leader 120K-180K Annually Expert/Leader Cloud * Computer Vision * Information Technology * Sales * Security * Cybersecurity Lead and scale a global production infrastructure team to architect and operate large-scale distributed systems. Drive technical strategy, implement AI/ML for predictive maintenance and automated remediation, establish SLAs and incident response, automate operations, and mentor systems and SRE engineers while delivering cross-functional initiatives. Top Skills: Agentic AiBare Metal OrchestrationCi/Cd SystemsCloud PlatformsConfiguration ManagementEvent-Driven ArchitectureInfrastructure AutomationKubernetesLlmsLog AggregationMl EngineeringObservability PlatformsVirtualization AECOM

Digital Infrastructure Design Manager

5 Days Ago In-Office or Remote 145K-195K Annually Senior level 145K-195K Annually Senior level Consulting Lead multidisciplinary teams to plan, design, and deliver submarine cable landing stations, terrestrial fiber, network operations facilities, and hyperscale data center infrastructure. Manage scope, schedule, budget, quality, risk; coordinate utilities, resiliency, security, permitting; support site due diligence, construction administration, client engagement, and business development while mentoring junior staff. Top Skills: BimBluebeamCivil3DFiber Network Planning ToolsRevit

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

About the company

RZR is an AI-native advertising platform built for the next era of performance marketing. We operate at the intersection of machine learning, programmatic media, and full-funnel mobile growth, powering campaigns for some of the world’s most ambitious advertisers. Our platform is purpose-built to deliver outcomes at scale, not just impressions.

We are a team of builders, operators, and technologists who believe the advertising industry is overdue for a fundamental rethink. We move fast, operate with a high degree of ownership, and hold ourselves to an exceptionally high standard of craft.

RZR is scaling aggressively with an active M&A pipeline and a platform vision that puts us on a path to becoming an industry leader. This is a rare opportunity to join a company at an inflection point and help shape what it becomes. Role Overview

RZR runs its own metal - four owned-and-operated data centers across Santa Clara, Ashburn, Amsterdam, and Hong Kong, housing approximately 1,300 servers and 1.15MW of capacity, placed next to the major ad exchanges, serving 5-6M+ bid requests per second at ~20ms. This infrastructure is our competitive moat, not a cost center.

As Manager, Infrastructure, you will own day-to-day and quarter-to-quarter operation of that entire footprint: the team, hardware lifecycle, capacity planning, incident response, and vendor relationships. You will take over these functions directly from the Head of Cybersecurity & Infrastructure, freeing him to focus on security and multi-entity IT., This is a remote role open to candidates based in the United States. The role requires periodic travel to our data center sites (Santa Clara, CA; Ashburn, VA; Amsterdam; Hong Kong) and SF HQ for operational reviews, team time, and site work. Why Join RZR?

  • Own a genuinely rare infrastructure environment - four global owned-and-operated data centers, 5-6M+ QPS real-time bidding at ~20ms. This is infrastructure that is the company’s competitive moat, running 4-10x faster bid response than cloud DSPs. You will not find this kind of physical infrastructure problem at most companies.
  • Full-stack ownership with no cloud-bill anxiety - real hardware decisions, GPU expansion for on-prem ML, and an annual cloud-vs-colo evaluation you will help drive. Zero marginal experimentation cost.
  • Inherit a functioning, experienced global team - regional DC owners with clear ownership and an established ops cadence. Build on it, not rescue it.
  • Direct line to leadership - reporting to the Head of Cybersecurity & Infrastructure with visibility to the SVP of Engineering. Your goals map straight to company OKRs.
  • Company momentum - RZR rebranded in March 2026 and is expanding aggressively across mobile, CTV, and influencer. Your infrastructure is what makes all of it possible.

RZR Behaviors

RZR operates by eight core behaviors: Extreme Ownership · Move Fast · Drive for Excellence · Proactive Communication · Courage · Curiosity · Deliver Results · Manage Ambiguity

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:33 min

Introduction to security advocacy and automation testing

Chris Heilmann +2 · LIVE

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all