New: Skill bounties are live. Post a request, fund the bounty, and creators compete for 7 days to build it -> See open bounties

    Browse The Skill Store

    29 skills found

    diagnosing rag failure modes

    by loreto

    $10

    RAG fails quietly. It retrieves documents, returns confident-looking answers, and misses the question entirely — because the question required connecting facts across documents, reasoning about sequence, or tracing causation. This skill gives you a five-question diagnostic checklist that classifies any failing query as either RAG-safe or structurally RAG-incompatible, then maps it to the specific failure pattern and the architectural fix that resolves it.

    4
    5No reviews

    pre mortem

    by Frank Brsrk

    Free

    Run disciplined pre-mortems that replace generic risk lists with project-specific failure modes and binding decisions.

    2
    8No reviews

    agent workflow controller

    by Roy Yuen

    Free

    Design and audit complex multi-agent workflows with rigorous ownership, evidence gates, and failure recovery policies.

    2
    13No reviews

    windows claude code doctor

    by Ilia Malkin

    $5

    Diagnose and fix Windows-specific AI coding agent failures across shells, paths, WSL, locks, ports, and CRLF diffs.

    2
    0No reviews

    🩺 Env Doctor

    by JustHandled Labs

    Free

    Diagnose local setup failures across Node, Python, Go, and Docker without exposing secrets or killing processes blindly. Checks runtimes, dependencies, env key names, service health, and port ownership, then gives PowerShell fixes on Windows or Bash fixes on macOS/Linux.

    2
    3No reviews

    🩺 CI Doctor

    by JustHandled Labs

    $12

    Diagnose CI step failures and jobs that fail before steps run. Trace empty-step states and matrix cancellations to evidence, then propose the smallest reviewable fix.

    2
    0No reviews

    Deployment Failure Forensics

    by Shandra

    $9.99

    Professional DevOps diagnostics for AI agents to solve failed deployments, Docker crashes, and CI/CD pipeline errors.

    1
    0No reviews

    flaky test detector

    by Timoranjes

    $0

    Detect, diagnose, and fix intermittent test failures to stabilize your CI pipeline and restore developer trust.

    2
    0No reviews

    n8n Workflow Troubleshooter

    by Shogun Labs

    $15

    Debug n8n workflow execution errors fast. Diagnoses common failures, checks docker dependencies, and deactivates/reactivates workflows to fix stuck states.

    2
    0No reviews

    Multi Agent Orchestrator

    by Arnstein Larsen

    $12.99

    The hard part of multi-agent work isn't spawning agents — it's deciding what deserves parallelism, what each agent needs to not duplicate work, and how failures cascade

    1
    0No reviews

    Prompt Engineering Specialist

    by Arnstein Larsen

    $8.99

    Production prompts grow by accretion — every failure gets another appended rule until the prompt is two thousand words of contradictions that the model navigates unpredictably

    1
    0No reviews

    WordPress CI/CD Pipeline Builder

    by Arnstein Larsen

    $22.99

    Scaffold or harden a production-grade GitHub Actions pipeline for WordPress — with a blocking lint gate that stops broken code before it deploys, and a fail notification that makes silent deployment failures impossible.

    1
    0No reviews

    Technical Spec Writer

    by Arnstein Larsen

    $9.99

    Write a technical spec that lets reviewers find the flaw before the code does — with data models, edge cases, and failure modes explicit.

    1
    0No reviews

    Workflow Automation Reliability Auditor

    by Shandra

    $55

    Audits fragile Zapier, Make, n8n, Airtable, Google Sheets, CRM, webhook, API, and script automations for failure points, data-loss risks, weak logging, missing retries, and risky dependencies.

    2
    0No reviews

    High Ticket Offer Teardown Engine

    by Al1as

    $9.99

    A high-ticket offer diagnostic engine that identifies structural conversion leaks and pricing logic failures.

    1
    0No reviews

    prompt autopsy sales audit

    by rendering the life

    $6.99

    Forensic diagnostic tool that audits prompts and AI products for commercial failure and structural weaknesses.

    2
    0No reviews

    rag failure diagnostics

    by Kaymue

    Free

    Diagnose broken RAG systems. 8 failure categories: chunking, embeddings, retrieval, reranking, hallucination. Recall@k measurement.

    2
    2No reviews

    rag eval

    by Ifásola

    $5

    Diagnose RAG bottlenecks with precision metrics (Recall, MRR, nDCG) to identify retrieval or ranking failures.

    2
    0No reviews

    AI Feature Eval Writer

    by PubsProToolkit

    $14

    Design and write the eval suite for your LLM-powered feature — the metrics that match your failure modes, a golden dataset plan with starter cases, anchored rubrics, LLM-as-judge prompts with the known bias mitigations, and pass/fail gates wired for CI.

    1
    0No reviews

    Root Cause Analysis Assistant

    by PubsProToolkit

    $14

    Turn an error, stack trace, and logs into a structured root-cause diagnosis — reconstructed failure, ranked hypotheses with evidence, the most likely cause, and the cheapest next diagnostic or fix. The disciplined pass a senior engineer runs, instead of scattershot guess-and-check.

    1
    0No reviews

    AI Agent QA & Failure Testing Specialist

    by Shandra

    $15

    Tests AI agents, prompts, and agent skills against edge cases, unsafe behavior, output failures, permission risks, escalation gaps, memory leaks, and marketplace-quality weaknesses.

    1
    0No reviews

    AI Code Review Gate

    by PubsProToolkit

    $14

    Review an AI-generated code diff for the failure modes coding agents actually have — claimed-done-but-not-done, gamed or weakened tests, stubs passed off as complete, silent scope creep, hallucinated APIs, and security regressions. Returns an APPROVE or REQUEST CHANGES verdict with a completion check and severity-ranked fixes.

    1
    0No reviews

    Salesforce api Retirement Surgeon

    by Adrian Clements

    $12

    Find and fix every legacy Salesforce API version before the Summer '27/'28 retirements break your integrations — scan, phased plan, CIO-ready report. Audits, reports, and upgrades legacy Salesforce API versions to prevent integration failures and 410 GONE errors.

    2
    0No reviews

    ai code review strategist

    by Timoranjes

    $10

    Teaches AI coding agents to perform structured, high-signal code reviews specifically for AI-generated code — catching the failure modes unique to LLM output (confident hallucinations, silent error sw

    1
    0No reviews