Works with the AI tools you already use

    Claude CodeClaude CodeCursorCursorCodex CLICodex CLIMuseMuseDotsDots+22 more

    Production Incident Triage Starter

    1

    Turn a rough production incident report into a scoped first-look brief with facts, unknowns, severity signals, next checks, and a clean handoff.

    Free

    2 installsSecurity scanned
    Production Incident Triage Starter

    Production Incident Triage Starter

    Example session with this skill installed

    At 14:07 UTC checkout latency rose from 400 ms to 8 s. A deployment finished at 13:58. Error rate increased only in eu-west, and some customers report duplicate refresh attempts. We have one timeout graph and 30 sanitized log lines, but no confirmed root cause. Build a first-look triage brief and safe next checks. Do not access production or invent a fix.

    • Read your context and instructions
    • Compiled the production incident triage

    First-look triage: the regional latency increase and deployment timing are facts; deployment causality and duplicate-attempt impact remain unproven. Preserve the timeline, compare changed versus unchanged regions, check timeout/retry ownership, quantify affected requests, and hand off to the appropriate reliability or debugging skill once evidence identifies the failure class.

    Connects securely to your tools. The creator never sees your data.

    About this skill

    The problem

    Initial incident reports are often a chaotic mix of conflicting logs, team theories, and fragmented timestamps. Without a clear triage, engineers waste time chasing red herrings or confusing correlation with actual causality.

    What it does

    • Separates observed facts from unverified hypotheses and team assumptions.
    • Normalizes incident timelines and identifies critical missing data intervals.
    • Classifies impact signals based strictly on provided telemetry and reports.
    • Generates a list of safe, low-risk checks to disprove leading theories.
    • Maps findings to specific specialist roles for efficient handoff.

    Why this beats prompting it yourself

    General-purpose models often hallucinate root causes or conflate timing with proof. This skill enforces a strict diagnostic discipline that prevents premature conclusions and ensures your next investigation step is based on evidence rather than intuition.

    Use cases

    • Synthesizing fragmented reports from multiple teams during an active outage.
    • Standardizing the first-look brief before escalating to senior SREs.
    • Mapping out safe diagnostic steps when deployment timing is known but impact is not.
    • Converting raw log dumps and alerts into a structured status report for stakeholders.

    Known limitations

    This tool performs no external actions and cannot access production environments. It relies entirely on sanitized material provided by the user.

    How to install

    Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.

    ~30 seconds
    1. 1

      Download the ZIP

      Free skills download straight away. Paid skills unlock right after purchase.

    2. 2

      Unzip into your skills folder

      Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.

    3. 3

      Ask your agent to use it

      Restart the agent if it was already running. It picks the skill up automatically - no config needed.

    Skills folder by agent

    Click the path to copy it. Create the folder if it does not exist yet.

    Reviews

    No reviews yet

    Be one of the first to try it. Every listed skill passes our trust checks below.

    Security scanned

    Passed our 8-point scan before listing

    2 installs

    Downloaded by developers to date

    Free forever

    No account required to browse

    Trust & safety

    Security scanned

    Verified clean 1 month ago

    • Free to download with an account

    Listed1 month ago

    What's inside

    Frequently Asked Questions