New: Skill bounties are live. Post a request, fund the bounty, and creators compete for 7 days to build it -> See open bounties

    Browse The Skill Store

    26 skills found

    lobster debugging

    by 王晓菲

    Free

    A systematic 4-phase debugging framework to find root causes, eliminate flaky tests, and prevent regressions.

    3
    23

    agent regression guard

    by Rian O'Leary

    $5

    Automated risk classification and regression checking to stop AI agents from breaking your codebase.

    2
    0

    🧪 Webapp Tester

    by JustHandled Labs

    $15

    Run real Playwright E2E tests on your web app: login, checkout, and form flows across desktop and mobile viewports, with screenshots, traces, and console logs captured on every failure. Catches broken flows and UI regressions before release, and tells you the likely fix, not just that something broke.

    2
    0

    AI Agent Self Improvement Memory Auditor

    by Shandra

    $50

    Audits AI agent failures and converts recurring mistakes into durable rules, anti-patterns, regression tests, memory candidates, and improved SKILL.md sections.

    1
    0

    AI Eval & Test Suite Quality Gate

    by PubsProToolkit

    Free

    An adversarial gate that audits an AI eval or test suite — LLM-judge rubrics, datasets, regression tests, metrics — for gameable criteria, data leakage, missing edge cases, and non-determinism, then returns one PASS/REVISE/FAIL verdict.

    2
    1

    ✅ AI Code Verification Gate

    by JustHandled Labs

    $19

    One-line summary description Stop your agent from claiming "done" before it's proven. A verification gate that classifies each change by risk (payment, auth, database, user-facing), picks the tests that actually cover it, demands evidence, maps regression risk, and outputs an honest pass/fail report. Turns "looks good to me" into "here's what I ran, and here's what's still unverified."

    1
    0

    ♿ a11y Regression Test Generator

    by JustHandled Labs

    $15

    Generate runnable accessibility regression tests, not just a findings report. Detects a11y issues, missing alt text, unlabeled controls, keyboard and focus gaps, in your routes, components, or HTML, then emits Playwright + axe-core spec files with targeted assertions and remediation tickets for each. Previews the tests first and writes them only on your confirmation.

    1
    0

    ⚡ Perf Budget Auditor

    by JustHandled Labs

    $12

    Audit your frontend build against a performance budget and catch size regressions before you ship. Flags total bundle over budget, initial bundle over budget, individual chunks over a threshold, oversized image assets, source maps shipped to production, and large unminified JavaScript. Reads a webpack or Vite-style stats.json plus a perf-budget.json you control.

    1
    0

    AI Feature Eval Writer

    by PubsProToolkit

    Free

    Design and write the eval suite for your LLM-powered feature — the metrics that match your failure modes, a golden dataset plan with starter cases, anchored rubrics, LLM-as-judge prompts with the known bias mitigations, and pass/fail gates wired for CI.

    1
    1

    AI Agent QA & Failure Testing Specialist

    by Shandra

    $15

    Tests AI agents, prompts, and agent skills against edge cases, unsafe behavior, output failures, permission risks, escalation gaps, memory leaks, and marketplace-quality weaknesses.

    1
    0

    Code Quality Gate

    by Echo Rose

    $5

    Code Quality Gate - A Premium AI Agent Skill

    1
    0

    App Tester

    by Echo Rose

    $5

    App Tester - A Premium AI Agent Skill

    1
    0

    AI Code Review Gate

    by PubsProToolkit

    Free

    Review an AI-generated code diff for the failure modes coding agents actually have — claimed-done-but-not-done, gamed or weakened tests, stubs passed off as complete, silent scope creep, hallucinated APIs, and security regressions. Returns an APPROVE or REQUEST CHANGES verdict with a completion check and severity-ranked fixes.

    1
    0

    Attestation Service

    by Echo Rose

    $5

    Attestation Service - A Premium AI Agent Skill

    1
    0

    Ci Test Analyzer

    by Echo Rose

    $5

    Ci Test Analyzer - A Premium AI Agent Skill

    1
    0

    Contract Test

    by Echo Rose

    $5

    Contract Test - A Premium AI Agent Skill

    1
    0

    Ci Test Reporter

    by Echo Rose

    $5

    Ci Test Reporter - A Premium AI Agent Skill

    1
    0

    mid session context recovery

    by Timoranjes

    Free

    Teaches AI coding agents to self-detect context rot (regression loops, instruction drift, hallucination drift, lost-in-the-middle) during long sessions and execute a structured checkpoint/recovery pro

    1
    0

    agent skill regression tester

    by Timoranjes

    Free

    Teaches AI coding agents to build and run automated regression tests for SKILL.md files. When you update a skill that your team depends on, you need to know it still works — not just that it "looks ri

    2
    3

    Ai Robustness Tester

    by Echo Rose

    $5

    Ai Robustness Tester - A Premium AI Agent Skill

    1
    0

    Code Regression Risk Analyzer

    by heyhridyansh

    Free

    Analyze code changes to map dependency blast radius, identify breaking contract changes, and score deployment risk.

    2
    1

    Skill Evaluation Harness Builder

    by Scar Tissue Systems

    $7.99

    Your skill works today. Will it work after the next model update? Build the harness that answers with numbers. Builds a standalone regression test harness with mechanical grading to verify skill behavior after model or code updates.

    1
    0

    Test Gap Finder

    by heyhridyansh

    $9.99

    Audit implementation behavior against test suites to find critical gaps in edge cases, security, and state transitions.

    2
    0

    Systematic Bug Diagnosis — Root Cause, Not Guesswork

    by Vivek K

    $12

    Finds why an app bug is happening using a step-by-step method (reproduce, isolate, test ideas, 5 Whys), then gives a small correct fix and a regression test — without guessing first.

    1
    0