World Congress 2024 Aug 29, 2024 Session details

Operating etcd for Managed Kubernetes

Mario Valderrama , Tinashe Mundangepfupfu

Struggling to scale etcd for multi-tenant Kubernetes? Discover how Ionos rebuilt their architecture to eliminate network latency and execute zero-downtime control plane migrations.

Pause
Mute Enter Fullscreen
#1 about 4 min

Adopting and scaling managed Kubernetes environments

Tracking the operational metrics of scaling managed infrastructure reveals the centralized importance of the primary datastore.

#2 about 2 min

Overcoming cost constraints with shared etcd instances

Sharing control planes reduces infrastructure overhead and lowers operational hardware costs for multi-tenant environments.

#3 about 3 min

Evaluating etcd deployment operators and helm charts

Migrating from community operators to a simplified custom Helm chart avoids hanging deployments and scripting conflicts.

#4 about 3 min

Addressing performance impacts in multi-tenant etcd setups

High event churn across shared application environments creates compounding technical strain that requires active client bounds.

#5 about 2 min

Tuning database compaction and system defragmentation

Enabling automatic compaction and expanding parallel operation capacity maintains reliable database memory footprint limits.

#6 about 3 min

Mitigating geographical latency in redundant cluster layouts

Resolving internal raft consensus drift requires adjusting heartbeat intervals or reverting to localized network architectures.

#7 about 3 min

Executing zero-downtime cluster control plane migrations

Utilizing dedicated mirror implementations enables continuous data synchronization without disrupting live database client processes.

#8 about 4 min

Manipulating BoltDB snapshot revisions for seamless transitions

Forcing a higher revision key into underlying database snapshots guarantees uninterrupted connections for existing client watches.

#9 about 2 min

Exploring dedicated clusters and alternative deployment models

Transitioning away from manual cluster migrations prompts exploration into dedicated environments and specialized proxy controllers.

#10 about 1 min

Contrasting shared deployment strategies with external tools

Assessing alternative fleet management setups highlights crucial differences in how component statefulness and configurations are implemented.

Matching moments

4:43 min

Managing the complexity of bare metal Kubernetes deployments

Josip Stuhli Josip Stuhli · World Congress 2026 Europe

3:26 min

Running Kubernetes clusters efficiently on enterprise cloud platforms

Niklas Heidloff · LIVE

5:40 min

Adapting Kubernetes deployment patterns for heterogeneous edge device fleets

Thomas Weinschenk Thomas Weinschenk · World Congress 2026 Europe

6:02 min

Customizing operator inputs and overriding default cluster configurations

Philipp Krenn · World Congress 2022

2:11 min

Executing massive cloud network migrations while maintaining live systems

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

1:10 min

Optimizing Kubernetes clusters for resource and cost efficiency

Christian Grieger Christian Grieger · Europe 2026 Virtual

Upcoming sessions on this topic

Open session

World Congress 2026 North America

September 24, 2026 · 17:30–18:00

Stage 4

Boring Failover: Predictable Region Recovery Across 5,000 Microservices

Garvit Kataria, Sahil Sabharwal

Garvit Kataria
Sahil Sabharwal
Open session

World Congress 2026 North America

September 24, 2026 · 15:30–16:00

Stage 9

Run your agents in Kubernetes: Build once, deploy anywhere. But really?

Michal Salanci

Senior Systems Engineer at ESET Cybersecurity

Michal Salanci
Open session

World Congress 2026 North America

September 25, 2026 · 15:30–16:00

Stage 7

Trust, But Verify: Continuous GPU Validation at Scale

Kyle Bell

VP of AI at TensorWave

Kyle Bell
Open session

World Congress 2026 North America

September 25, 2026 · 11:00–11:30

Stage 5

Managing GPUs by Just Asking, Infrastructure in the Age of MCP

Jessica Garson Beauchemin

Developer Relations Lead, Community at Runpod

Jessica Garson Beauchemin
Open session

World Congress 2026 North America

September 25, 2026 · 09:00–09:30

Stage 2

Autonomous Infrastructure: Building AI Agents for Global-Scale Capacity Efficiency

Tommy Tran

Software Engineer at Meta

Tommy Tran
Open session

World Congress 2026 North America

September 24, 2026 · 15:30–16:00

Stage 4

Why Infrastructure Forecasting Fails – Building a Self-Serve Forecasting Platform

Ankur Gupta

LinkedIn, Senior Staff Technical Program Manager

Ankur Gupta