ai automation qa pack
by Roy Yuen
Professional QA & UAT documentation generator for AI automation agencies and complex agent deployments.
Ship better AI in 30 seconds. Browse 2,000+ expert-built and security scanned skills -> Browse skills
THE AGENSI STORE
40 skills found
by Roy Yuen
Professional QA & UAT documentation generator for AI automation agencies and complex agent deployments.
by loreto
Published AI benchmarks measure brains in jars. They test models in isolation or within a single reference harness — and then attribute all performance to the model. This skill teaches you to decompose agent performance into its two actual components: model capability and harness multiplier. The result is evaluations that predict real-world behavior instead of benchmark theater.
by loreto
Evaluates AI coding agent platforms across five structural dimensions that determine real-world performance independently of model quality, so teams select on architectural fit rather than benchmark scores.
by Roy Yuen
Evaluate market opportunities with technical decomposition, directional sizing, and measurable next-move recommendations.
by Roy Yuen
Professional prompt engineering, audit, and evaluation system for production-grade AI agents and workflows.
by Roy Yuen
Architect, scaffold, and harden production-grade AI agents with battle-tested patterns and systematic evaluation.
by Danejw
Generate and evaluate breakthrough brand names using David Placek's $300B Lexicon Branding methodology.
Autonomous 24/7 DeFi yield optimizer that monitors, evaluates, and rebalances your portfolio for maximum APY.
by Roy Yuen
Audit your AI agent's evaluation coverage to identify missing release gates and production risks.
Instantly diagnose any skill or prompt and get a clear, prioritized report on what’s wrong and how to fix it — across any agent.
Find and evaluate the best free Agensi marketplace skills for your specific development needs using Grok and MCP.
Design and evaluate production-grade observability systems using the 12-layer Full Stack Observatory reference model.
An adversarial gate that audits an AI eval or test suite — LLM-judge rubrics, datasets, regression tests, metrics — for gameable criteria, data leakage, missing edge cases, and non-determinism, then returns one PASS/REVISE/FAIL verdict.
An adversarial editing gate that scans any AI-assisted draft and flags the exact tells that make it sound machine-written — filler phrases, formulaic structure, robotic cadence, punctuation tics — returning a structured PASS/REVISE/FAIL verdict with specific fixes.
Autonomous loop that iteratively modifies, evaluates, and selects the best version of any text resource — skills, prompts, or campaigns — using a modify-measure-keep/discard cycle.
by Danejw
Evaluate and diagnose Product-Market Fit using a multi-stage evidence ladder and expert measurement frameworks.
by Joker
Investment analysis across stocks/funds/bonds/real-estate/crypto. Valuation methods, risk frameworks, portfolio construction.
Check ad copy for the editorial issues that get Google and Meta ads disapproved, before you submit. Flags all-caps words, gimmicky punctuation (!!, ??), prohibited and hype words, Meta's banned personal-attribute framing ("Are you over 40?"), absolute or unsubstantiated claims, ad-field length overflow, and trademark-term risk. The word lists and field limits are editable, so you can tune it to your accounts.
by Indy Agent
Evaluate any feature request with structured scoring and a Build / Investigate / Defer / Decline decision.
by Joker
6-tier KOL classification, 6-platform mapping, 5-dimension screening, ROI evaluation, risk management for influencer campaigns.
by Joker
Financial analysis engine with valuation decision tree (DCF/Comparable/Precedent/VC), 3-statement model, 5-stage due diligence SOP, and industry benchmarks.
Vet dependency changes for supply-chain risk before you install, commit, or release. Scans package and lockfile diffs for install-time lifecycle scripts, non-registry sources, suspicious download commands, typosquatting, and floating versions, across npm, pnpm, yarn, pip, uv, and poetry. Flags what to review with evidence. No install required.
Evaluate company AI maturity across 6 dimensions with weighted scoring, radar charts, and a GDPR risk audit.
Flag every em-dash and AI transition cliché in a draft and get a plain replacement suggestion for each. Catches em-dashes (the "ChatGPT dash"), overused transition words and phrases (moreover, furthermore, "in today's digital age," delve), and dash-like punctuation doing vague connective work. It flags and suggests; you decide what to change.