TELECOMMUTE Operational Support Engineer (L2)

Kani Solutions
United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Content Delivery Networks Cloud Computing DevOps Pattern Recognition Reliability Engineering Prometheus Data Streaming WebRTC Grafana Infrastructure as Code (IaC)
+7 more
Backend Containerization Kubernetes Video Streaming Kibana Terraform Microservices

Job description

We are building a dedicated L2 Operational Support team responsible for the stability and availability of our 24/7 live video streaming, ad insertion, and real-time delivery platforms. This is a high-ownership, production-level role acting as the technical bridge between Support, Engineering, DevOps, and enterprise customers during major live events., * Incident Ownership & Production Operations: Own and resolve high-impact production incidents. Safely operate directly on production infrastructure, execute live emergency changes, and make CDN adjustments under tight SLAs.

  • Infrastructure as Code (IaC): Understand and safely modify infrastructure using Terraform, Helm, Kubernetes manifests, and GitOps workflows.
  • AI-Driven Operations: Leverage AI tooling for automated runbook execution, intelligent alert correlation, pattern detection, and automated troubleshooting workflows.
  • Observability & Triaging: Correlate backend streaming metrics, player telemetry, and CDN signals using Grafana, Kibana/ELK, Prometheus, and Loki to isolate root causes.

Requirements

Experience in the video streaming domain is mandatory.

  • 5+ Years of relevant experience in a technical support, operations, or reliability engineering role.
  • Streaming Protocol Mastery: Deep understanding of HLS, DASH, CMAF, WebRTC, DRM, and global CDN architectures.
  • Infrastructure Familiarity: Solid troubleshooting experience across distributed cloud systems, APIs, microservices, and containerized deployments.
  • Incident Management: Exceptional ability to lead live incident bridges, remain highly structured under pressure, and deliver clear customer-facing communications.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:21 min

Deploying a primary Elasticsearch and Kibana cluster configuration

Philipp Krenn · WWC 2022

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:07 min

Simplifying peer-to-peer connections using WebRTC abstraction libraries

André Dietrich André Dietrich · WWC 2024

2:35 min

Exploring advanced video stream enhancements and API features

Phil Cluff · LIVE

4:51 min

Executing simple full-text search queries using the Kibana interface

Derek Binkley · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all