Trace incidents to evidence and leave systems easier to debug.
site reliability engineering
Investigates production incidents and bugs through systematic, hypothesis-driven triage and evidence-backed analysis.
When to use it
Use for incident response, root cause analysis, production debugging, or log investigation.
Give it an incident or debugging problem; it works with you to return verified, evidence-backed findings.
This skill
Slack
Uploads diagrams to Slack (irreversible)
production deployment
Rolls back production deployments
feature flags
Reverts feature flags
traffic controls
Blocks or rate-limits traffic
The documented prerequisites require access to one or more Axiom deployments.
The documented prerequisites include jq.
A timeout command is required; on macOS the README directs users to install coreutils.