World Congress 2026 Europe - Virtual Stage Jul 3, 2026 Session details

Easy Mode Monitoring and Logging with Shiftmon

Mathias Palmersheim

Ditch expensive SaaS observability tools and heavy infrastructure. ShiftMon uses Ansible to automate a production-ready metrics and logging stack on a single Linux VM.

Pause
Mute Enter Fullscreen
#1 about 1 min

Introduction to easy mode observability and ShiftMon

Automating difficult infrastructure setups via open-source tools resolves the underlying friction of adopting advanced observability practices.

#2 about 2 min

Setting up observability without expensive SaaS subscriptions

Deploying a small Linux virtual machine provides an easily extensible observability on-ramp bypassing costly subscriptions.

#3 about 1 min

Deploying configuration as code seamlessly using Ansible

Converting a plain Linux machine into an observability stack via infrastructure as code eliminates the need for managed configuration agents.

#4 about 4 min

Connecting core components to build observability platforms

Integrating reverse proxies and dedicated data storage systems constructs a robust monitoring backend environment.

#5 about 1 min

Automating telemetry collection through robust Telegraf deployment

Using Ansible roles to discover background processes automatically configures the correct host metric collection patterns.

#6 about 2 min

Streamlining baseline observability setup for lazy engineers

Automating host statistics and standard application instrumentation eliminates the manual burden of configuring baseline telemetry.

#7 about 3 min

Discovering and instrumenting services using systemd process enumeration

Detecting local services automatically to scrape unauthenticated endpoints streamlines collection while allowing custom user overrides.

#8 about 3 min

Deciding between push metric architecture and service discovery

Why ephemeral workloads and highly mobile nodes necessitate push-based monitoring instead of traditional scrape architectures.

#9 about 4 min

Instrumenting supported infrastructure components for automated telemetry scraping

Automating metric ingestion across supported infrastructure components enables deep visibility into networking domains and container hosts.

#10 about 2 min

Measuring application availability thresholds using black box monitoring

Deploying external HTTP probes and anomaly detection logic identifies unexpected performance latency across digital services.

#11 about 2 min

Managing tokens and user credentials securely across endpoints

Storing sensitive observability parameters inside external secret managers avoids exposing plaintext configurations on host deployments.

#12 about 3 min

Customizing dashboards and alert rules for proprietary applications

Creating tailored observability rule sets and collector configurations preserves the ability to monitor unsupported internal services.

#13 about 3 min

Securing telemetry monitoring agents to prevent privilege escalation

Running telemetry collection tools strictly as unprivileged users prevents arbitrary command execution and security vulnerabilities.

#14 about 4 min

Provisioning dashboards and data source alerts as code

Maintaining visualization definitions and stateless alerting rules as versioned code enables rapid infrastructure disaster recovery.

#15 about 4 min

Exploring observability dashboards and alert management interface workflows

Touring a live user interface reveals how authenticated operators investigate log correlation traces and monitor fleet performance.

Matching moments

1:10 min

Background and origins of the Shiftmon monitoring project

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

39 sec

Introducing Monoscope for intelligent post-deployment system monitoring

Anthony Alaribe Anthony Alaribe · World Congress 2025

4:11 min

Building a unified stack leveraging core open source tools

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

17:03 min

Introduction to metrics and observability challenges in monitoring

Liam Hurrell · LIVE

3:11 min

Implementing a robust observability stack for agents

Taylor Jordan Smith Taylor Jordan Smith · World Congress 2026 Europe

Upcoming sessions on this topic

Open session

World Congress 2026 North America

September 24, 2026 · 13:30–15:30

Stage 14

From Signal to Action: Empowering Your AI SRE with OpenTelemetry Data

Julia Furst Morgado, Raphael Manke

Julia Furst Morgado
Raphael Manke
Open session

World Congress 2026 North America

September 25, 2026 · 12:55–13:25

Stage 9

It’s Alive! Taming the MLOps Franken-Stack: Write, Run, and Serve with Michelangelo

Eric Wang, Paul Zimmerman

Eric Wang
Paul Zimmerman
Open session

World Congress 2026 North America

September 25, 2026 · 09:00–09:30

Stage 1

Test Before Release, Enforce at Runtime: Governance for Tool-Using AI Agents

Sachin Gupta

Member of Technical Staff 2 at eBay

Sachin Gupta
Open session

World Congress 2026 North America

September 25, 2026 · 15:45–15:55

Outdoor Stage

Closing the Visibility Gap: Lessons from Safety Critical Agentic Systems

Vivek Pandit

Frontier AI Lead at Turing

Vivek Pandit
Open session

World Congress 2026 North America

September 25, 2026 · 15:30–16:00

Stage 1

AI-Powered Incident Triage: How We Built GenAI Agents with MCPs to Automate On-Call Workflows

Prakshal Doshi

Site Reliability Engineer

Prakshal Doshi
Open session

World Congress 2026 North America

September 24, 2026 · 11:40–12:10

Stage 5

Your Agents Need Observability Before They Need Better Models

Julia Furst Morgado

Principal Developer Relations Engineer at Dash0

Julia Furst Morgado