What problem does it solve? Production incidents often devolve into chaotic, uncoordinated firefighting with unclear ownership, poor stakeholder communication, and no follow-through on root causes. This Skill provides a structured incident command framework that turns outages into organized response efforts with defined roles, severity classification, and blameless post-mortems. ## Core Features & Use Cases - Structured Incident Response: Establishes SEV1–SEV4 severity classification, assigns roles (Incident Commander, Communications Lead, Technical Lead, Scribe), and enforces time-boxed troubleshooting with fixed communication cadences. - Post-Mortem Facilitation: Generates blameless post-mortem documents with timelines, 5 Whys root cause analysis, and tracked action items with owners and deadlines. - SLO/SLI & On-Call Design: Defines error budgets, burn-rate alerts, on-call rotation schedules, escalation policies, and runbook templates with tested remediation steps. - Use Case: Your payment API starts returning 5xx errors at 2 AM. The Skill guides you through declaring a SEV1, assigning roles, executing a rollback runbook, communicating with stakeholders every 15 minutes, and producing a complete post-mortem within 48 hours. ## Quick Start Ask the agent to help you declare and coordinate a response for a production outage affecting your checkout service, including severity classification and stakeholder communication.