Topic mix

Incident response

16 moments from 16 videos · 45:09 min total

Learn how engineering teams handle high-pressure outages and systematically resolve production issues. These curated conference moments focus on blameless post-mortems and coordination.

Unlocking the AI Black Box: Building Trust in the Era of Agentic Production
Play section Tracking operational challenges and incident response metrics
Tracking operational challenges and incident response metrics thumbnail

Tracking operational challenges and incident response metrics

Connecting metric tracking directly to enterprise revenue by reducing incident investigation and mitigation times.

Shipping with Confidence: Observability and Quality at Scale
Play section Closing the development loop with automated AI incident responses
Closing the development loop with automated AI incident responses thumbnail

Closing the development loop with automated AI incident responses

Managing the escalating velocity of coding pipelines requires deploying analytical agents in production to quickly diagnose anomalies and submit automated repairs.

Empathy: The secret sauce of Resilience
Play section Analyzing incident responses to uncover organizational culture
Analyzing incident responses to uncover organizational culture thumbnail

Analyzing incident responses to uncover organizational culture

Evaluating how a company investigates and learns from production incidents reveals its true capacity for self-correction.

3 Key Steps for Optimizing DevOps Workflows
Play section Prioritizing incident detection and recovery over prevention
Prioritizing incident detection and recovery over prevention thumbnail

Prioritizing incident detection and recovery over prevention

Minimizing the duration of incidents reduces overall disruption more effectively than pursuing expensive prevention strategies.

Applying Agile Principles to Incident Management
Play section Applying software development methodologies to incident response
Applying software development methodologies to incident response thumbnail

Applying software development methodologies to incident response

Adapting core frameworks like agile iterations, devops culture, and sre automation to manage technical crises.

Software That Fixes Itself
Play section Shifting software delivery bottlenecks to operations and incident response
Shifting software delivery bottlenecks to operations and incident response thumbnail

Shifting software delivery bottlenecks to operations and incident response

Accelerating upstream coding shifts system pressure toward downstream deployment verification and stressful on-call environments.

Agentic DevOps: How AI-Powered Automation Transforms Software Delivery on GitHub and Azure
Play section Streamlining incident response and root cause analysis automatically
Streamlining incident response and root cause analysis automatically thumbnail

Streamlining incident response and root cause analysis automatically

How automated log aggregation and meeting transcriptions improve incident documentation and systemic mitigation.

SRE Methods In an Agency Environment
Play section Assigning strict roles for incident response teams
Assigning strict roles for incident response teams thumbnail

Assigning strict roles for incident response teams

Clear escalation boundaries isolate the specific duties of an incident commander, communications lead, and operations lead during an active issue.

Enabling automated 1-click customer deployments with built-in quality and security
Play section Introduction to network security and endpoint monitoring architectures
Introduction to network security and endpoint monitoring architectures thumbnail

Introduction to network security and endpoint monitoring architectures

An overview of deploying dedicated security zones and connecting them to central incident response tools.

Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps
Play section Maximizing global incident coverage through asynchronous remote team distribution
Maximizing global incident coverage through asynchronous remote team distribution thumbnail

Maximizing global incident coverage through asynchronous remote team distribution

Enhancing worldwide incident capabilities and meeting efficiency by mandating strictly aligned agendas within a remote-first culture.

Blameless Postmortems That Change Nothing: Cultural Anti-Patterns in Incident Learning
Play section The hidden reality of unresolved engineering incidents
The hidden reality of unresolved engineering incidents thumbnail

The hidden reality of unresolved engineering incidents

Superficial fixes and poor incident documentation after outages prevent meaningful organizational learning.

APItoolkit: Using Merkle Trees and LLMs to Detect the UnDetectable in Software Monitoring
Play section Applying large language models to evaluate real-time system changes
Applying large language models to evaluate real-time system changes thumbnail

Applying large language models to evaluate real-time system changes

Combining language models with observability data enables context-aware impact analysis and conversational interactions during incident responses.

DevSecOps: Security in DevOps
Play section Defining application, pipeline, and security operations roles
Defining application, pipeline, and security operations roles thumbnail

Defining application, pipeline, and security operations roles

Breaking down the security journey into application security, pipeline deployment security, and incident response operations.

Walking into the era of Supply Chain Risks
Play section Creating secure baselines by tracking container environment configurations
Creating secure baselines by tracking container environment configurations thumbnail

Creating secure baselines by tracking container environment configurations

Empowering software engineering teams with precise inventory visibility reduces the cognitive load during incident response.

Leave it to the Dutch: How to save on your Azure bill!
Play section Configuring automated cost alerts for budget thresholds
Configuring automated cost alerts for budget thresholds thumbnail

Configuring automated cost alerts for budget thresholds

Setting alert limits slightly above anticipated spending enables rapid incident response to sudden billing spikes.

How to make a 4 day week work for everyone
Play section Measuring productivity and business impact of reduced hours
Measuring productivity and business impact of reduced hours thumbnail

Measuring productivity and business impact of reduced hours

Analyzing ticket resolution velocity and incident response metrics quantitatively validates that a four-day week improves developer productivity.

Your mix. Instantly.

More mixes