synthesizing institutional knowledge
by loreto
Builds the organizational memory schema your AI agent needs to answer why — capturing decision provenance, causal chains, and event context that embedding-based retrieval permanently discards.
New: Skill bounties are live. Post a request, fund the bounty, and creators compete for 7 days to build it -> See open bounties
THE AGENSI STORE
63 skills found
by loreto
Builds the organizational memory schema your AI agent needs to answer why — capturing decision provenance, causal chains, and event context that embedding-based retrieval permanently discards.
by loreto
Architects the right retrieval strategy for every query — teaching your agent when to use RAG, a knowledge graph, or a temporal index instead of defaulting to vector search for everything.
by loreto
Published AI benchmarks measure brains in jars. They test models in isolation or within a single reference harness — and then attribute all performance to the model. This skill teaches you to decompose agent performance into its two actual components: model capability and harness multiplier. The result is evaluations that predict real-world behavior instead of benchmark theater.
by loreto
Evaluates AI coding agent platforms across five structural dimensions that determine real-world performance independently of model quality, so teams select on architectural fit rather than benchmark scores.
by Roy Yuen
Evaluate market opportunities with technical decomposition, directional sizing, and measurable next-move recommendations.
by Roy Yuen
Professional prompt engineering, audit, and evaluation system for production-grade AI agents and workflows.
by loreto
RAG fails quietly. It retrieves documents, returns confident-looking answers, and misses the question entirely — because the question required connecting facts across documents, reasoning about sequence, or tracing causation. This skill gives you a five-question diagnostic checklist that classifies any failing query as either RAG-safe or structurally RAG-incompatible, then maps it to the specific failure pattern and the architectural fix that resolves it.
by Roy Yuen
Architect, scaffold, and harden production-grade AI agents with battle-tested patterns and systematic evaluation.
by Danejw
Generate and evaluate breakthrough brand names using David Placek's $300B Lexicon Branding methodology.
Autonomous 24/7 DeFi yield optimizer that monitors, evaluates, and rebalances your portfolio for maximum APY.
Instantly diagnose any skill or prompt and get a clear, prioritized report on what’s wrong and how to fix it — across any agent.
by Roy Yuen
Audit your AI agent's evaluation coverage to identify missing release gates and production risks.
Find and evaluate the best free Agensi marketplace skills for your specific development needs using Grok and MCP.
by alan zhou
Professional Chinese judicial case search and rule induction powered by the Deli Legal API.
Design and evaluate production-grade observability systems using the 12-layer Full Stack Observatory reference model.
by Rapa Canola
Multi-source dispatcher for aggregating technical research and community sentiment across HN, Reddit, and GitHub.
Connect your agent to the Steam Web API to fetch player data, game libraries, and achievement statistics.
An adversarial gate that audits an AI eval or test suite — LLM-judge rubrics, datasets, regression tests, metrics — for gameable criteria, data leakage, missing edge cases, and non-determinism, then returns one PASS/REVISE/FAIL verdict.
Autonomous loop that iteratively modifies, evaluates, and selects the best version of any text resource — skills, prompts, or campaigns — using a modify-measure-keep/discard cycle.
Drastically reduce RAG costs and latency while improving retrieval accuracy through advanced memory architecture.
by LocoLoboZ
Standardize and validate Make API shell scenarios and connection workflows for reusable SaaS integrations.
by Danejw
Evaluate and diagnose Product-Market Fit using a multi-stage evidence ladder and expert measurement frameworks.
Automate Mailtrap sending domain setup, DNS record retrieval, and verification across major providers.
by Joker
6-tier KOL classification, 6-platform mapping, 5-dimension screening, ROI evaluation, risk management for influencer campaigns.