New: UPI payments are live. Buyers in India can now pay for skills with UPI in INR -> Browse skills

    Browse The Skill Store

    54 skills found

    prompt engineer

    by Roy Yuen

    Free

    Professional prompt engineering patterns for building robust, secure, and production-ready LLM applications.

    9
    2475.0(1)

    benchmarking ai agents beyond models

    by loreto

    Free

    Published AI benchmarks measure brains in jars. They test models in isolation or within a single reference harness — and then attribute all performance to the model. This skill teaches you to decompose agent performance into its two actual components: model capability and harness multiplier. The result is evaluations that predict real-world behavior instead of benchmark theater.

    2
    165.0(1)

    prompt engineer pro

    by Roy Yuen

    $8

    Professional prompt engineering, audit, and evaluation system for production-grade AI agents and workflows.

    3
    0

    harness engineering

    by Roy Yuen

    $8

    Design, debug, and harden AI control loops with explicit contracts and automated verification harnesses.

    2
    2

    agent reliability audit

    by Roy Yuen

    $5

    Turn raw agent traces and tool logs into professional production-readiness audits and remediation reports.

    2
    0

    finops anomaly intelligence

    by appugouda ai

    Free

    Turn AWS billing mysteries into 10-minute root cause reports by correlating cost spikes with engineering events.

    3
    45.0(1)

    chaos engineering

    by Frank Brsrk

    Free

    Design rigorous chaos engineering experiments and resilience audits to verify production system reliability.

    3
    55.0(1)

    prompt spec engineer (support Auto rewrite prompt)

    by Roy Yuen

    $5

    Turn vague prompts into professional task specifications, optimized prompts, and verification test suites.

    2
    0

    dora metrics reviewer

    by Julian

    $15

    Benchmark your DevOps performance against DORA standards and generate a prioritized 90-day improvement roadmap.

    2
    0

    weekly cowork system auditor

    by LocoLoboZ

    $7

    A structured governance auditor to optimize AI project instructions, clean up context, and manage workspace health.

    2
    1

    📝 Prompt Template Linter

    by JustHandled Labs

    $12

    Lint a prompt template for the issues that cause injection and flaky output. Flags untrusted variables interpolated straight into the instructions (the injection surface), placeholders that are never provided or never used, contradictory instructions, a missing output-format spec where the result is parsed, unbounded context interpolation, and leftover placeholders. It detects problems; it does not write prompts.

    3
    0

    Kubernetes Config Error Detective

    by Shandra

    $50

    Audits Kubernetes manifests, Helm values, deployment logs, and service configs to detect configuration errors and produce safe, reviewable fix plans.

    2
    0

    LLM Eval Framework Builder

    by StrategistKit

    $17

    Builds a complete LLM evaluation framework — quality dimensions, a golden dataset, code-based and model-graded rubric graders, judge calibration, and CI regression rules. Use when the user says build LLM evals, create a golden dataset, or set up LLM-as-judge. Do not use when they want to debug one bad model output, not build a repeatable measurement system.

    1
    0

    PromptDecoder Pro — AI Output to Prompt Converter

    by Brandon DeVries

    Free

    Paste any AI output. Get the production-ready prompt that made it.

    1
    9

    Prompt Engineering Specialist

    by StrategistKit

    $8.99

    Produces a diagnosed and rewritten prompt — component-level failure analysis, structure fixes, few-shot examples, and a regression case set. Use when the user says "fix my prompt", "why does this prompt keep failing", "improve my system prompt", "reduce hallucinations in this prompt", or "my prompt isn't following instructions". Do not use when the request is choosing which model to use, not fixing wording.

    1
    0

    normalize prompt set

    by GTDataworks

    Free

    Convert loose prompt sets into structured, target-ready records with variables, contracts, and eval cases.

    1
    2

    prompt failure mode auditor

    by ALBERTO “TRAlbert”

    Free

    Hardens AI prompts and agent workflows against logic errors, tool-misuse, and prompt injection.

    1
    1

    Senior Web App Architect for Coding Agents

    by Avesta Agency Australia | Software Engineering Group

    $9.99

    Messy, insecure, unfixable — that's what AI builds without architecture. This file is the architecture: 10 years of senior judgement on rendering, caching, security and SEO, so your agent builds it right from day one.

    6
    15.0(1)

    RAG Failure Diagnostics & Architect

    by StrategistKit

    $7

    Produces a diagnosis of why a RAG system gives confident-but-wrong answers, or picks between vector search, knowledge graph, and structured/temporal retrieval. Use when the user says "why is our RAG hallucinating", "diagnose this failing query", "should we use a knowledge graph", "pick a retrieval architecture", or "design our memory layer". Do not use when the request is building a RAG system from zero.

    1
    0

    Agent Scaffold Builder

    by StrategistKit

    $8.99

    Produces a complete AI agent scaffold — charter, system prompt, SKILL.md, tool map, guardrails, runbook, and test suite — from a one-line job description. Use when the user says "build an AI agent for X", "scaffold a new agent", "write a system prompt and SKILL.md", "design an agent's guardrails", or "make an agent workflow runbook". Do not use when the user just wants a single prompt with no full agent architecture.

    1
    0

    Prompt Habits Kit

    by StrategistKit

    $8.99

    Produces a personalized set of 6 reusable prompt-habit templates tailored to the user's field. Use when the user says "give me reusable prompt patterns", "build my prompt habits", "stop writing prompts from scratch", "prompt templates for my field", or "make me a prompting system". Do not use when the request is fixing one broken prompt, not building a reusable system.

    1
    0

    Fleet Scale Migration Orchestrator Human Supervised

    by John Barros

    $39

    Orchestrate human-supervised code migrations across repository fleets with verifier loops and judge review gates.

    2
    0

    PR Release Risk Gatekeeper

    by heyhridyansh

    $9.99

    A formal release gatekeeper that audits PRs for deployment risk, migration safety, and operational readiness.

    2
    1

    Skill Evaluation Harness Builder

    by Scar Tissue Systems

    $7.99

    Your skill works today. Will it work after the next model update? Build the harness that answers with numbers. Builds a standalone regression test harness with mechanical grading to verify skill behavior after model or code updates.

    1
    0