Lead DevOps Engineer

Paramount Pictures
New York, NY, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours

Tech stack

A/B Testing Application Programming Interfaces (APIs) Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Continuous Integration Software Debugging DevOps Github Python (Programming Language) Machine Learning
+25 more
Memcached Node.Js Octopus Deploy Queueing Systems Redis Reliability Engineering Prometheus Data Streaming Scripting Google Cloud Load Balancing System Availability Grafana Caching Backend Event Driven Architecture Kubernetes Low Latency Real Time Data Apache Kafka Api Gateway Terraform New Relic (SaaS) Jenkins Microservices

Job description

We are looking for a Lead DevOps Engineer to join our Applied Intelligence Personalization team. You’ll build and maintain scalable, low-latency infrastructure that powers personalization and engagement across Paramount’s streaming platforms. Our workloads include real-time machine-learning inference, event-driven messaging, and high-traffic backend services - so deep Kubernetes, CI/CD, and cloud-infrastructure expertise is the core of the role, with ML-serving experience as a solid plus., Design, implement, and manage scalable and reliable Kubernetes-based infrastructure for personalization services.

  • Build and own CI/CD pipelines that ship services (and ML models) to production safely - canary, rollback, progressive delivery.
  • Stand up observability and monitoring with Prometheus, New Relic, OpenTelemetry, and Grafana; define SLIs / SLOs and drive error-budget discipline.
  • Ensure high availability, security, and performance of production APIs and streaming data pipelines.
  • Partner with application, data, and ML engineers to integrate their workloads smoothly into the platform.
  • Implement autoscaling strategies (HPA, KEDA, traffic-driven) for variable, bursty traffic patterns.
  • Manage Pub/Sub and event-driven architectures for real-time messaging, engagement analytics, and inter-service communication.
  • Optimize hot-path services using Redis, Memcached, and other caching strategies.
  • Debug and tackle production issues around latency, scaling, and reliability.

Key Projects

  • Build and optimize real-time serving infrastructure for personalization and engagement (including ML-inference workloads).
  • Develop scalable, secure CI/CD pipelines for deploying services and ML models.
  • Implement log aggregation and monitoring solutions for platform-wide observability.
  • Optimize Kubernetes-based serving for minimal latency and efficient resource utilization.
  • Improve A/B testing infrastructure to track the impact of personalized experiences.
  • Enhance streaming data pipelines to support real-time data consumers.

Requirements

The ideal candidate has 4+ years of production DevOps / SRE experience, is fluent in Kubernetes and infrastructure as code, and cares deeply about reliability, latency, and developer velocity., 4+ years of experience in DevOps, Site Reliability Engineering (SRE), or Cloud Infrastructure Engineering.

  • Solid experience with Kubernetes and container orchestration.
  • Hands-on experience with CI/CD tools such as GitHub Actions, Jenkins, and ArgoCD.
  • Deep knowledge of Google Cloud Platform (GCP), AWS, or Azure.
  • Expertise in infrastructure as code (IaC) using Terraform and Helm.
  • Experience with message queues and event-driven architectures (Pub/Sub, Kafka, etc.).
  • Proficiency in monitoring and logging solutions (New Relic, Prometheus, OpenTelemetry, etc.).
  • Strong scripting skills in Python, Bash, or Go for automation.
  • Track record of owning production reliability - SLIs / SLOs, incident response, postmortems.

Preferred / Additional Qualifications

  • Experience deploying and operating ML models in production (TF Serving, Triton, TorchServe, Ray Serve).

  • GPU / accelerator scheduling and node-pool management on Kubernetes.
  • Knowledge with load balancing, API gateways, and caching strategies at scale.
  • Experience with A/B testing frameworks and experimentation infrastructure.
  • Experience optimizing low-latency microservices for personalization or recommendation workloads.
  • Passion for building and maintaining high-performance infrastructure for real-time applications.

Benefits & conditions

  • Attractive compensation and comprehensive benefits packages. Check out our full list of benefits here: https://www.paramount.com/careers/benefits
  • Generous paid time off.
  • An exciting and fulfilling opportunity to be part of one of Paramount’s most dynamic teams.
  • Opportunities for both on-site and virtual engagement events.
  • Unique opportunities to make meaningful connections and build a vibrant community, both inside and outside the workplace.
  • Explore life at Paramount: https://www.paramount.com/careers/life-at-paramount

Paramount is an equal opportunity employer (EOE) including disability/vet.

At Paramount, the spirit of inclusion feeds into everything that we do, on-screen and off. From the programming and movies we create to employee benefits/programs and social impact outreach initiatives, we believe that opportunity, access, resources and rewards should be available to and for the benefit of all. Paramount is proud to be an equal opportunity workplace. We are committed to equal employment opportunity regardless of race, color, ethnicity, ancestry, religion, creed, sex, national origin, sexual orientation, age, citizenship status, marital status, disability, gender identity, gender expression, and Veteran status.

About the company

We’ve got the brands, we’ve got the stars, we’ve got the power to achieve our mission to entertain the planet - now all we’re missing is… YOU! Becoming a part of Paramount means joining a team of passionate people who not only recognize the power of content but also enjoy a touch of fun and uniqueness. Together, we co-create moments that matter - both for our audiences and our employees - and aim to leave a positive mark on culture., About Paramount Streaming Paramount Streaming, a division within Paramount Global, is the home to the company’s direct-to-consumer services spanning free and paid in the form of Pluto TV and Paramount+. Pluto TV is the global leader in free ad-supported TV, delivering more than 1,400 global channels and an extensive library of streaming content, including live and original channels. Paramount+, digital subscription video-on-demand and live streaming service, combines live sports, breaking news, and A Mountain of Entertainment . Paramount+ features an expansive library of original series, hit shows and popular movies across every genre from world-renowned brands and production studios, including SHOWTIME®.

Paramount Streaming, a division within Paramount Global, is the home to the company’s direct-to-consumer services spanning free and paid in the form of Pluto TV and Paramount+. Pluto TV is the global leader in free ad-supported TV, delivering more than 1,400 global channels and an extensive library of streaming content, including live and original channels. Paramount+, digital subscription video-on-demand and live streaming service, combines live sports, breaking news, and A Mountain of Entertainment . Paramount+ features an expansive library of original series, hit shows and popular movies across every genre from world-renowned brands and production studios, including SHOWTIME®.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careers.paramount.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

3:32 min

Shifting to a DevOps career from non-technical backgrounds

Megha Kadur · LIVE

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · WWC 2022

Videos

See all

Related articles

See all