Uploaded May 2026 | Updated September 2026, 2 weeks ago
Kevin Panahi, Senior Site Reliability Engineer at Favor Delivery, walks through the Operational Readiness Review (ORR) ceremony his 4-person SRE team runs across 100 services and 12 squads. Built on a single Grafana Cloud dashboard — pulling Istio metrics through Grafana Alloy, error budgets through Grafana SLO, and incident follow-ups from Jira via the Infinity data source — the biweekly ceremony has cut alert fatigue, dropped Ask-SRE interrupts 30%, and flipped Favor's incident detection from 1-in-3 internal to 2-in-3 internal.
Key chapters:
• 00:00 Introduction and Favor Delivery background
• 01:00 Inside the SRE and DevOps teams
• 01:30 Where Favor's monitoring stood when Kevin joined
• 02:20 Istio service mesh metrics into Grafana Cloud via Alloy
• 03:30 Building SLIs and SLOs on availability and latency
• 05:00 Tiered SLO model — squads own the tier and the target
• 06:00 Error budgets and burn rate alerts
• 07:00 The ORR ceremony and why it exists
• 08:00 Naming the meeting (and missing the acronym)
• 09:00 Mini ORR walkthrough — the incident follow-ups panel
• 10:00 The active SLOs panel and the 80% budget rule
• 13:00 Active alerts and the alert fatigue review
• 15:00 SRE adoption across 100 services
• 17:30 Outcomes: detection ratio, alert volume, interrupts
• 19:30 What's next: per-endpoint SLOs, FARO, k6 private load zones, Kafka SLOs
Links/resources:
Get started with the Grafana Cloud forever-free tier: grafana.com/g/cloud
Have a question? Ask Grot, your AI helper: grafana.com/grot
Reach out in our community forums: https://gra.fan/communityyf
---
Thanks for watching!
👍 Was this video helpful? Like and subscribe to our channel for more videos.
Connect with Grafana Labs:
X: (twitter.com/grafana)
LinkedIn: (linkedin.com/company/grafana-labs/)
Facebook: (facebook.com/grafana)
#Grafana #SRE #Observability #SLO #DevOps #FavorDelivery
Kevin Panahi, Senior Site Reliability Engineer at Favor Delivery, walks through the Operational Readiness Review (ORR) ceremony his 4-person SRE team runs across 100 services and 12 squads. Built on a single Grafana Cloud dashboard — pulling Istio metrics through Grafana Alloy, error budgets through Grafana SLO, and incident follow-ups from Jira via the Infinity data source — the biweekly ceremony has cut alert fatigue, dropped Ask-SRE interrupts 30%, and flipped Favor's incident detection from 1-in-3 internal to 2-in-3 internal.
Key chapters:
• 00:00 Introduction and Favor Delivery background
• 01:00 Inside the SRE and DevOps teams
• 01:30 Where Favor's monitoring stood when Kevin joined
• 02:20 Istio service mesh metrics into Grafana Cloud via Alloy
• 03:30 Building SLIs and SLOs on availability and latency
• 05:00 Tiered SLO model — squads own the tier and the target
• 06:00 Error budgets and burn rate alerts
• 07:00 The ORR ceremony and why it exists
• 08:00 Naming the meeting (and missing the acronym)
• 09:00 Mini ORR walkthrough — the incident follow-ups panel
• 10:00 The active SLOs panel and the 80% budget rule
• 13:00 Active alerts and the alert fatigue review
• 15:00 SRE adoption across 100 services
• 17:30 Outcomes: detection ratio, alert volume, interrupts
• 19:30 What's next: per-endpoint SLOs, FARO, k6 private load zones, Kafka SLOs
Links/resources:
Get started with the Grafana Cloud forever-free tier: grafana.com/g/cloud
Have a question? Ask Grot, your AI helper: grafana.com/grot
Reach out in our community forums: https://gra.fan/communityyf
---
Thanks for watching!
👍 Was this video helpful? Like and subscribe to our channel for more videos.
Connect with Grafana Labs:
X: (twitter.com/grafana)
LinkedIn: (linkedin.com/company/grafana-labs/)
Facebook: (facebook.com/grafana)
#Grafana #SRE #Observability #SLO #DevOps #FavorDelivery










