Incident Response

On-call / production incident triage: correlate logs + recent deploys + config changes into a most-likely-cause, read a console screenshot and give exact fix commands, or write+run a diagnostic query over metrics/logs and interpret it. TRIGGER when the user says 'API latency doubled in the last hour, check the logs, recent deploys, and config changes, then tell me the most likely cause', 'here is a screenshot of the AWS console, walk me through why the RDS instance is failing and give me the exact commands to fix it', 'show me all 5xx events for /payments over the last 24h, write the query, run it, and tell me what stands out', or reports an ongoing outage/latency spike/error surge. Hebrew: 'תקרית פרודקשן', 'ה-latency קפץ, תבדוק לוגים ודיפלויים', 'למה ה-RDS נופל', 'תכתוב שאילתת אבחון על הלוגים', 'טריאז׳ לתקרית'. For a single reproducible bug (not a live incident), use a root-cause debugging skill like debug instead, if you have one.

hiyabh Updated

File contents

hiyabh/claude-skills/tree/main/src/incident-response commit 21d3f55618

Frequently asked questions

npx skillmds@latest add hiyabh/incident-response