Runtime Boundary & Resilience Repair
Repair runtime failures that appear only under latency, retries, shutdowns, worker ownership, and dependency pressure. This bundle combines timeout-budget repair, keep-alive socket lifecycle fixes, retry-storm containment, poison-message handling, job-lease repair, shutdown data-loss protection, SRE reliability improvement, and reliability scoring. Use it when components work individually but production behavior collapses under slow dependencies, repeated failures, restarts, or recovery pressure.
Works with every agent that reads SKILL.md — Claude Code, Cursor, Codex CLI, Gemini CLI, GitHub Copilot, Windsurf, OpenClaw, and more.
One payment, lifetime access. 8 skills unlock instantly in your library.
30-day refund guarantee
Instant unlock in your library
Free updates from the creator
What's included
8 skillsAlign end-to-end request deadlines, retries, cancellation, database work, and side-effect ownership so server work stops when its budget is gone.
Repair HTTP keep-alive lifecycle bugs that leak sockets, exhaust pools, or cause intermittent resets after idle periods and deploys.
Repair retry loops that overload APIs, queues, and external services during outages.
Repair queue consumers where one bad message loops forever, blocks healthy traffic, or repeatedly crashes the same partition.
Repair workers that leave jobs stuck, process the same task twice, or let expired owners write results.
Repair apps that lose jobs, events, or writes when containers restart, deployments roll, or processes receive termination signals.
Want to improve your system reliability? This skill analyzes your system and provides SRE-level recommendations to improve uptime, resilience, and fault tole...
Evaluate how reliable your system really is. This skill analyzes your architecture and generates a reliability score with insights and improvement suggestion...
See it in action
Repair Timeout Budget Drift
Example session with this skill installed
Our checkout API has an 8-second caller deadline. The browser aborts at 8s, the gateway waits 15s, the Node service gives the inventory service 12s per attempt and retries twice, PostgreSQL has no statement timeout, and an email side effect can continue after the HTTP request has already been cancelled. We are seeing requests return 504 while server work continues and later logs show a second retry finishing after the client is gone. Repair the timeout-budget contract in the repository. Propagate cancellation, keep retries inside the supplied 8-second budget, add deterministic tests and observability hooks, and do not deploy or change production infrastructure.
- Read your context and instructions
- Compiled the repair timeout budget
Timeout-budget drift found. The 8-second caller budget is exceeded by downstream attempt limits, retries, uncancelled database work, and detached email work. The repair introduces one absolute deadline, remaining-budget propagation, retry gating, cancellation-aware database handling, durable side-effect handoff, deterministic tests, and timeout reason metrics. Production gateway and provider behavior remain unverified.
How to install
Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.
- 1
Download the ZIP
Free skills download straight away. Paid skills unlock right after purchase.
- 2
Unzip into your skills folder
Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.
- 3
Ask your agent to use it
Restart the agent if it was already running. It picks the skill up automatically - no config needed.
Skills folder by agent
Click the path to copy it. Create the folder if it does not exist yet.
Reviews
No reviews yet
Be one of the first to try it. Every skill in this bundle passes our trust checks.
Security scanned
Passed our 8-point scan before listing
Fresh listing
Recently published to Agensi
30-day refund
Not a fit? Get your money back