> Markdown version of [/jobs/ext/2549938-staff-site-reliability-engineer-playout](https://www.wearedevelopers.com/jobs/ext/2549938-staff-site-reliability-engineer-playout). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Site Reliability Engineer, Playout - **Company:** Nbcuniversal Media, LLC - **Location:** Stamford, CT, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Cloud Computing, Cloud Engineering, Codecs, H.264/MPEG-4 AVC, Internet Protocol, Linux System Administration, Performance Tuning, Reliability Engineering, Runbook, Data Logging, Grafana, Mttr, Containerization, Information Technology, Live Streaming, Splunk, Docker, Servicenow - **Published:** August 10, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/27925197/Staff-Site-Reliability-Engineer-Playout-Connecticut-Stamford-1011 ## About the Role This role requires the ability to operate in a fast-paced environment. For systems in production, you will lead an on-call team and drive L1 and L2 troubleshooting, incident management and continuous improvement to maintain reliable distribution., * Bachelor's degree in computer science or related degree / experience * Eight years' hands-on-keyboard Engineering experience working with broadcast automation playout environments e.g. Snell, Harris, Imagine, Amagi * Requires on-call 24/7 availability for escalations * Hands-on-keyboard experience administrating Linux environments * Experience with monitoring/logging tools e.g. Splunk and Grafana * Experience with streaming protocols and codecs (e.g. TS, HEVC, H.264, HLS, CMAF, SCTE-35, SCTE-224, ESAM, SRT/RIST) * Experience with IP networking and interfacing with cloud-based networks * Experience with containerization (Docker & Kubernetes) * Excellent communicator and able to clearly articulate complex issues and technologies * Expert with broadcast playout systems (master control) technologies * Expert with public cloud environments using AWS services * Comfortable working in a fast-paced agile environment. Requirements change quickly and our team needs to constantly adapt to meet objectives * An automate-first and automate everything attitude Desired Characteristics * Experience with cloud native playout vendor solutions (Amagi, Evertz, GrassValley, Harmonic, Imagine, CoralBay, Veset, etc.) * Experience building strong operational readiness practices (runbooks, alert tuning, on-call health, incident reviews) * Ability to create user interface designs based on client workflows ## Description NBCUniversal Operations & Technology is looking for a Staff SRE, Playout Engineering to provide technical leadership to a team of Site Reliability Engineers. This team drives reliability, observability, and operational excellence for cloud-based master control playout systems supporting all NBCUniversal live linear channels - NBC, Telemundo, Peacock Virtual Channels, etc. In this position, you will shape the reliability strategy for live linear playout-defining service levels, improving resiliency, and strengthening monitoring and incident response-so the platform can meet evolving business needs with predictable performance and availability., * Team lead for SRE engineers on playout Engineering team * Define and manage reliability targets (SLIs/SLOs) and operational readiness criteria for playout services * Drive incident response: establish on-call practices, lead major incident management, and ensure post-incident reviews result in measurable improvements * Partner with engineering, product, and operations teams to improve reliability through capacity planning, performance tuning, and resilience testing * Provide high-level conceptual drawings and operational runbooks to support architecture reviews, support readiness, and project planning * Leadership in driving automation to reduce toil and improve reliability and mean time to recovery (MTTR) * L1 & L2 support to maintain playout infrastructure/services for NBCUniversal including providing after hours on-call support select week * Leadership in creating monitoring dashboards (Grafana) and proper alerts(teams/slack/ServiceNow) ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025) - [How to Write a CV and Interview if You Don't Fully Qualify For The Job](https://www.wearedevelopers.com/magazine/183-how-to-write-a-cv-and-interview-if-you-don-t-fully-qualify-for-the-job)