Works with the AI tools you already use

    Claude CodeClaude CodeCursorCursorCodex CLICodex CLIMuseMuseDotsDots+22 more

    Product Experiment Decision Gate

    1

    The problem Product experiments can produce clean-looking results without producing trustworthy decisions.

    $6.99

    Secure checkout via Stripe

    30-day refund guarantee

    Converts to your local currency at checkout

    Security scanned
    Product Experiment Decision Gate

    Product Experiment Decision Gate

    Example session with this skill installed

    Review my A/B test plan for the new checkout flow. We are changing the button color and removing two form fields to increase conversion. We'll ship if conversion goes up by 2%.

    • Read your context and instructions
    • Compiled the product experiment decision

    Gate verdict

    REVISE BEFORE RUNNING. The intervention boundary is contaminated; changing color and form fields simultaneously makes it impossible to attribute success.

    Primary outcome

    Checkout conversion (USEFUL BUT INCOMPLETE PROXY).

    Experiment integrity gaps

    P0: Bundle intervention (color vs fields).
    P1: No guardrail for completion quality or error rates.

    Decision rule

    DECISION RULE INCOMPLETE. A 2% lift may not outweigh the implementation cost if the baseline is low.

    Connects securely to your tools. The creator never sees your data.

    What you get

    Validate experiment designs to ensure they support clear ship or stop decisions.Identify hidden confounds like novelty effects or population contamination.Establish guardrails to prevent shipping features that damage retention.Audit completed test results to prevent post-hoc metric cherry-picking.

    About this skill

    The problem

    Product experiments can produce clean-looking results without producing trustworthy decisions. A metric goes up, the team calls the test a win, and important questions about proxy quality, guardrails, contamination, measurement, and false positives are left unresolved.

    What it does

    • Checks whether an experiment is actually capable of supporting the product decision attached to it.
    • Separates the hypothesis, proposed mechanism, and current evidence so assumptions do not silently become proof.
    • Audits primary metrics for proxy risk and identifies meaningful guardrails against hidden product or operational harm.
    • Finds instrumentation gaps, contamination risks, population mismatches, novelty effects, and post-hoc interpretation.
    • Checks whether sample-size logic, statistical method, thresholds, and stopping rules are actually established rather than assumed.
    • Tests the cost of false positives and false negatives before recommending SHIP, ITERATE, STOP, or RE-TEST.

    Why this beats prompting it yourself

    Generic experiment prompts are usually good at generating test ideas, but weaker at challenging whether the test can actually justify the decision that follows. This skill pressure-tests the decision logic underneath the experiment and refuses to invent baselines, sample sizes, thresholds, or statistical certainty just to make the plan look complete.

    Use cases

    • Review A/B tests before launch to find weak hypotheses, bad success metrics, or missing guardrails.
    • Audit onboarding, pricing, growth, feature, or UX experiments before rollout decisions.
    • Check whether a positive result is strong enough to support shipping or scaling.
    • Find false-pass pathways where a test can look successful while creating downstream harm.
    • Review completed experiment results for post-hoc interpretation, confounds, or decision-rule drift.

    Known limitations

    This skill does not calculate statistical power without the required inputs, certify causal validity, independently validate analytics instrumentation, or replace a statistician or experiment analyst when formal statistical review is required. Strong conclusions depend on the quality of the experiment design, data, and context provided.

    How to install

    Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.

    ~30 seconds
    1. 1

      Download the ZIP

      Free skills download straight away. Paid skills unlock right after purchase.

    2. 2

      Unzip into your skills folder

      Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.

    3. 3

      Ask your agent to use it

      Restart the agent if it was already running. It picks the skill up automatically - no config needed.

    Skills folder by agent

    Click the path to copy it. Create the folder if it does not exist yet.

    Reviews

    No reviews yet

    Be one of the first to try it. Every listed skill passes our trust checks below.

    Security scanned

    Passed our 8-point scan before listing

    Fresh listing

    Recently published to Agensi

    30-day refund

    Not a fit? Get your money back

    Trust & safety

    Security scanned

    Verified clean 1 month ago

    • Passed all security checks, Safe to install

    Listed1 month ago
    Updated16 days ago

    What's inside

    Frequently Asked Questions