Ship better AI in 30 seconds. Browse 2,000+ expert-built and security scanned skills -> Browse skills

    Browse The Skill Store

    11 skills found

    Peer Review Stress Test

    by PubsProToolkit

    $12

    An adversarial self-review gate that hunts your agent's weakest claim, overclaims, and missing limitations before a human sees the output.

    2
    0No reviews

    Evidence Guard

    by PubsProToolkit

    $14

    Audit any AI-generated output for unsupported claims, then verify every factual and technical assertion against its real source before it ships.

    2
    0No reviews

    LLM Eval Framework Builder

    by Arnstein Larsen

    $17

    You changed the prompt, tried four inputs, it looked better, you shipped — and three days later support tickets say outputs are worse for an entire class of inputs you didn't test

    1
    0No reviews

    LLM Prompt Stabilizer — 6 Layer Pattern for Consistent Agent Output

    by Shogun Labs

    $15

    Battle-tested prompting patterns to eliminate LLM output drift. Sandwich structure, few-shot examples, history limits, retry, and token caps — 6 composable layers for production-grade agent reliability.

    2
    0No reviews

    Prompt Injection & Agent Security Gate

    by PubsProToolkit

    $14

    An adversarial security gate that audits untrusted content — web pages, tool outputs, documents, emails — for embedded instructions, exfiltration, and authority spoofing, then returns a SAFE/REVIEW/BLOCK verdict.

    2
    0No reviews

    Subagent Workflow Patterns To Boost Output Quality

    by zoninmane

    $9

    Deploy 6 battle-tested multi-agent orchestration patterns to eliminate agent laziness and boost output quality.

    1
    0No reviews

    AI Agent QA & Failure Testing Specialist

    by Shandra

    $15

    Tests AI agents, prompts, and agent skills against edge cases, unsafe behavior, output failures, permission risks, escalation gaps, memory leaks, and marketplace-quality weaknesses.

    1
    0No reviews

    Agent Harness Architect

    by PubsProToolkit

    $14

    Model quality is table stakes — the harness is where agents win or fail. This designs yours: it writes a structured, testable system prompt (role, tools, boundaries, method, output contract, failure handling) and maps every concern to the right layer — prompt, tool, guardrail, or evaluation — so the pieces reinforce each other instead of fighting.

    1
    0No reviews

    API Spec Researcher — Systematic External API Investigation

    by Shogun Labs

    Free

    Battle-tested prompting patterns to eliminate LLM output drift. Sandwich structure, few-shot examples, history limits, retry, and token caps — 6 composable layers for production-grade agent reliability.

    2
    1No reviews

    MCP Server Builder — Model Context Protocol Development Guide

    by Shogun Labs

    Free

    Battle-tested prompting patterns to eliminate LLM output drift. Sandwich structure, few-shot examples, history limits, retry, and token caps — 6 composable layers for production-grade agent reliability.

    2
    0No reviews

    ai code review strategist

    by Timoranjes

    $10

    Teaches AI coding agents to perform structured, high-signal code reviews specifically for AI-generated code — catching the failure modes unique to LLM output (confident hallucinations, silent error sw

    1
    0No reviews