WeAreDevelopers LIVE Dec 3, 2020

5 steps for running a Kubernetes environment at scale

Stijn Polfliet

Hitting memory limits will instantly terminate your production pods. Master a five-step observability strategy to troubleshoot crash loops, monitor telemetry, and confidently run Kubernetes at scale.

Pause
Mute Enter Fullscreen
#1 about 6 min

Introduction to the five-layer Kubernetes observability model

A high-level five-layer framework helps teams monitor and scale containerized environments reliably.

#2 about 6 min

Identifying failing pods visually in the cluster explorer

Visualizing node and pod health provides immediate context on pending or crash-looping workloads.

#3 about 8 min

Managing node capacity using resource requests and limits

Tracking resource consumption alongside configured request limits prevents node starvation and unexpected out-of-memory errors.

#4 about 3 min

Using smaller container images for performance and security

Switching to lightweight distribution bases like Alpine Linux improves build times and minimizes potential security vulnerabilities.

#5 about 5 min

Controlling pod routing with readiness and liveness probes

Implementing readiness and liveness probes allows the scheduler to control traffic routing and automated pod restarts accurately.

#6 about 3 min

Centralizing Kubernetes microservice log collection with Fluent Bit

Forwarding microservice log streams via lightweight log shippers centralizes log search and debugging.

#7 about 10 min

Tracing distributed communication pathways across Kubernetes microservices

Injecting trace identifiers into HTTP headers reveals latency bottlenecks across complex inter-service communication paths.

#8 about 11 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Ingesting open-source metric endpoints into a centralized platform unifies query workflows alongside distributed traces and events.

#9 about 9 min

Deploying the Kubernetes observability stack using Helm charts

Installing specialized monitoring components via package managers ensures continuous visibility alongside dynamic application deployments.

Matching moments

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

1:59 min

Orchestrating with Kubernetes against Docker and accessing slides

Philipp Krenn · World Congress 2022

1:39 min

Introduction to Kubernetes security challenges and opportunities

Marc Nimmerrichter · World Congress 2022

1:10 min

Optimizing Kubernetes clusters for resource and cost efficiency

Christian Grieger Christian Grieger · Europe 2026 Virtual

3:32 min

Transitioning from monolith architectures to microservices and Kubernetes

Liviu Costea · LIVE

5:04 min

Visualizing Kubernetes for alternative and creative learning styles

Aurélie Vache Aurélie Vache · World Congress 2026 Europe

Upcoming sessions on this topic

Open session

World Congress 2026 North America

September 24, 2026 · 15:30–16:00

Stage 9

Run your agents in Kubernetes: Build once, deploy anywhere. But really?

Michal Salanci

Senior Systems Engineer at ESET Cybersecurity

Michal Salanci
Open session

World Congress 2026 North America

September 25, 2026 · 15:30–16:00

Stage 7

Trust, But Verify: Continuous GPU Validation at Scale

Kyle Bell

VP of AI @ TensorWave

Kyle Bell
Open session

World Congress 2026 North America

September 24, 2026 · 11:40–12:10

Stage 2

Stop Running Mystery Meat in Production

Jeroen van Erp

Technical Advocate @ SUSE

Jeroen van Erp
Open session

World Congress 2026 North America

September 24, 2026 · 10:20–10:50

Stage 2

From Static Rules to Reasoning Platforms: Scaling Intelligent Canary Delivery in 2026

Daniel Oh

Senior Principal Developer Advocate

Daniel Oh
Open session

World Congress 2026 North America

September 24, 2026 · 17:30–18:00

Stage 4

Boring Failover: Predictable Region Recovery Across 5,000 Microservices

Garvit Kataria, Sahil Sabharwal

Garvit Kataria
Sahil Sabharwal
Open session

World Congress 2026 North America

September 24, 2026 · 15:30–16:00

Stage 4

Why Infrastructure Forecasting Fails – Building a Self-Serve Forecasting Platform

Ankur Gupta

LinkedIn, Senior Staff Technical Program Manager

Ankur Gupta