Runtime Boundary & Resilience Repair
Repair runtime failures that appear only under latency, retries, shutdowns, worker ownership, and dependency pressure. This bundle combines timeout-budget repair, keep-alive socket lifecycle fixes, retry-storm containment, poison-message handling, job-lease repair, shutdown data-loss protection, SRE reliability improvement, and reliability scoring. Use it when components work individually but production behavior collapses under slow dependencies, repeated failures, restarts, or recovery pressure.
Works with every agent that reads SKILL.md — Claude Code, Cursor, Codex CLI, Gemini CLI, GitHub Copilot, Windsurf, OpenClaw, and more.
One payment, lifetime access. 8 skills unlock instantly in your library.
30-day refund guarantee
Instant unlock in your library
Free updates from the creator
What's included
8 skillsAlign end-to-end request deadlines, retries, cancellation, database work, and side-effect ownership so server work stops when its budget is gone.
Repair HTTP keep-alive lifecycle bugs that leak sockets, exhaust pools, or cause intermittent resets after idle periods and deploys.
Repair retry loops that overload APIs, queues, and external services during outages.
Repair queue consumers where one bad message loops forever, blocks healthy traffic, or repeatedly crashes the same partition.
Repair workers that leave jobs stuck, process the same task twice, or let expired owners write results.
Repair apps that lose jobs, events, or writes when containers restart, deployments roll, or processes receive termination signals.
Want to improve your system reliability? This skill analyzes your system and provides SRE-level recommendations to improve uptime, resilience, and fault tole...
Evaluate how reliable your system really is. This skill analyzes your architecture and generates a reliability score with insights and improvement suggestion...
See it in action
You say
Our checkout API has an 8-second caller deadline. The browser aborts at 8s, the gateway waits 15s, the Node service gives the inventory service 12s per attempt and retries twice, PostgreSQL has no statement timeout, and an email side effect can continue after the HTTP request has already been cancelled. We are seeing requests return 504 while server work continues and later logs show a second retry finishing after the client is gone. Repair the timeout-budget contract in the repository. Propagate cancellation, keep retries inside the supplied 8-second budget, add deterministic tests and observability hooks, and do not deploy or change production infrastructure.
Your agent does
Timeout-budget drift found. The 8-second caller budget is exceeded by downstream attempt limits, retries, uncancelled database work, and detached email work. The repair introduces one absolute deadline, remaining-budget propagation, retry gating, cancellation-aware database handling, durable side-effect handoff, deterministic tests, and timeout reason metrics. Production gateway and provider behavior remain unverified.
How to install
Drop the file into your AI Agent. Works with Claude, Cursor, ChatGPT, and 20+ more.
Reviews
No reviews yet
Be one of the first to try it. Every skill in this bundle passes our trust checks.
Security scanned
Passed our 8-point scan before listing
Fresh listing
Recently published to Agensi
30-day refund
Not a fit? Get your money back