Dependency Failure Containment
Contain dependency failures before latency, retries, stale endpoints, connection pressure, and recovery traffic cascade across a production system. This bundle combines timeout-budget repair, circuit-breaker lifecycle repair, retry-storm containment, connection-pool exhaustion repair, DNS staleness, keep-alive socket repair, SRE reliability improvement, and reliability scoring. Use it when one degraded dependency can trigger a much larger application outage.
Works with every agent that reads SKILL.md — Claude Code, Cursor, Codex CLI, Gemini CLI, GitHub Copilot, Windsurf, OpenClaw, and more.
One payment, lifetime access. 8 skills unlock instantly in your library.
30-day refund guarantee
Instant unlock in your library
Free updates from the creator
What's included
8 skillsAlign end-to-end request deadlines, retries, cancellation, database work, and side-effect ownership so server work stops when its budget is gone.
Repair duplicate breaker ownership, state drift, half-open probe storms, retry amplification, reset races, and failure-classification mismatches.
Repair retry loops that overload APIs, queues, and external services during outages.
Repair leaked or long-held connections, per-request client pools, fan-out saturation, retry amplification, and cleanup gaps without guessing pool sizes.
Repair ignored DNS TTLs, startup-only resolution, stale endpoint reuse, retry-to-same-address loops, discovery refresh stampedes, and DNS/pool confusion.
Repair HTTP keep-alive lifecycle bugs that leak sockets, exhaust pools, or cause intermittent resets after idle periods and deploys.
Want to improve your system reliability? This skill analyzes your system and provides SRE-level recommendations to improve uptime, resilience, and fault tole...
Evaluate how reliable your system really is. This skill analyzes your architecture and generates a reliability score with insights and improvement suggestion...
See it in action
You say
Our checkout API has an 8-second caller deadline. The browser aborts at 8s, the gateway waits 15s, the Node service gives the inventory service 12s per attempt and retries twice, PostgreSQL has no statement timeout, and an email side effect can continue after the HTTP request has already been cancelled. We are seeing requests return 504 while server work continues and later logs show a second retry finishing after the client is gone. Repair the timeout-budget contract in the repository. Propagate cancellation, keep retries inside the supplied 8-second budget, add deterministic tests and observability hooks, and do not deploy or change production infrastructure.
Your agent does
Timeout-budget drift found. The 8-second caller budget is exceeded by downstream attempt limits, retries, uncancelled database work, and detached email work. The repair introduces one absolute deadline, remaining-budget propagation, retry gating, cancellation-aware database handling, durable side-effect handoff, deterministic tests, and timeout reason metrics. Production gateway and provider behavior remain unverified.
How to install
Drop the file into your AI Agent. Works with Claude, Cursor, ChatGPT, and 20+ more.
Reviews
No reviews yet
Be one of the first to try it. Every skill in this bundle passes our trust checks.
Security scanned
Passed our 8-point scan before listing
Fresh listing
Recently published to Agensi
30-day refund
Not a fit? Get your money back