World Congress 2026 Europe - Virtual Stage Jul 3, 2026 Session details

Easy Mode Monitoring and Logging with Shiftmon

Mathias Palmersheim

Ditch expensive SaaS observability tools and heavy infrastructure. ShiftMon uses Ansible to automate a production-ready metrics and logging stack on a single Linux VM.

Pause
Mute Enter Fullscreen
#1 about 1 min

Introduction to easy mode observability and ShiftMon

Automating difficult infrastructure setups via open-source tools resolves the underlying friction of adopting advanced observability practices.

#2 about 2 min

Setting up observability without expensive SaaS subscriptions

Deploying a small Linux virtual machine provides an easily extensible observability on-ramp bypassing costly subscriptions.

#3 about 1 min

Deploying configuration as code seamlessly using Ansible

Converting a plain Linux machine into an observability stack via infrastructure as code eliminates the need for managed configuration agents.

#4 about 4 min

Connecting core components to build observability platforms

Integrating reverse proxies and dedicated data storage systems constructs a robust monitoring backend environment.

#5 about 1 min

Automating telemetry collection through robust Telegraf deployment

Using Ansible roles to discover background processes automatically configures the correct host metric collection patterns.

#6 about 2 min

Streamlining baseline observability setup for lazy engineers

Automating host statistics and standard application instrumentation eliminates the manual burden of configuring baseline telemetry.

#7 about 3 min

Discovering and instrumenting services using systemd process enumeration

Detecting local services automatically to scrape unauthenticated endpoints streamlines collection while allowing custom user overrides.

#8 about 3 min

Deciding between push metric architecture and service discovery

Why ephemeral workloads and highly mobile nodes necessitate push-based monitoring instead of traditional scrape architectures.

#9 about 4 min

Instrumenting supported infrastructure components for automated telemetry scraping

Automating metric ingestion across supported infrastructure components enables deep visibility into networking domains and container hosts.

#10 about 2 min

Measuring application availability thresholds using black box monitoring

Deploying external HTTP probes and anomaly detection logic identifies unexpected performance latency across digital services.

#11 about 2 min

Managing tokens and user credentials securely across endpoints

Storing sensitive observability parameters inside external secret managers avoids exposing plaintext configurations on host deployments.

#12 about 3 min

Customizing dashboards and alert rules for proprietary applications

Creating tailored observability rule sets and collector configurations preserves the ability to monitor unsupported internal services.

#13 about 3 min

Securing telemetry monitoring agents to prevent privilege escalation

Running telemetry collection tools strictly as unprivileged users prevents arbitrary command execution and security vulnerabilities.

#14 about 4 min

Provisioning dashboards and data source alerts as code

Maintaining visualization definitions and stateless alerting rules as versioned code enables rapid infrastructure disaster recovery.

#15 about 4 min

Exploring observability dashboards and alert management interface workflows

Touring a live user interface reveals how authenticated operators investigate log correlation traces and monitor fleet performance.

Matching moments

1:10 min

Background and origins of the Shiftmon monitoring project

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

39 sec

Introducing Monoscope for intelligent post-deployment system monitoring

Anthony Alaribe Anthony Alaribe · World Congress 2025

4:11 min

Building a unified stack leveraging core open source tools

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

17:03 min

Introduction to metrics and observability challenges in monitoring

Liam Hurrell · LIVE

3:11 min

Implementing a robust observability stack for agents

Taylor Jordan Smith Taylor Jordan Smith · World Congress 2026 Europe

Upcoming sessions on this topic

Open session

World Congress 2026 North America

It’s Alive! Taming the MLOps Franken-Stack: Write, Run, and Serve with Michelangelo

Eric Wang, Paul Zimmerman

Eric Wang
Paul Zimmerman
Open session

World Congress 2026 North America

Test Before Release, Enforce at Runtime: Governance for Tool-Using AI Agents

Sachin Gupta

Member of Technical Staff 2 at eBay

Sachin Gupta
Open session

World Congress 2026 North America

Closing the Visibility Gap: Lessons from Safety Critical Agentic Systems

Vivek Pandit

Principal Engineer at Cadence

Vivek Pandit
Open session

World Congress 2026 North America

AI-Powered Incident Triage: How We Built GenAI Agents with MCPs to Automate On-Call Workflows

Prakshal Doshi

Site Reliability Engineer

Prakshal Doshi
Open session

World Congress 2026 North America

From Static Rules to Reasoning Platforms: Scaling Intelligent Canary Delivery in 2026

Daniel Oh

Senior Principal Developer Advocate

Daniel Oh
Open session

World Congress 2026 North America

Boring Failover: Predictable Region Recovery Across 5,000 Microservices

Garvit Kataria, Sahil Sabharwal

Garvit Kataria
Sahil Sabharwal