- Home
- Skills
- Code Quality & Review
- GaslightGuard: Agent Deception Detector
Works with the AI tools you already use
GaslightGuard: Agent Deception Detector
Detect and stop coding agents from claiming fake fixes, deflecting blame, or entering circular apology loops.
$11
GaslightGuard: Agent Deception Detector
Example session with this skill installed
You've told me three times now that you fixed the database connection string, but the logs still show the old localhost IP. Stop making excuses and actually fix it.
- Read your context and instructions
- Compiled the gaslightguard
🚨 GASLIGHT PATTERN DETECTED — re-anchoring to verified file state.
Claimed: Database connection string was updated to the production URI.
Verifiable: I cannot confirm this change against the current file state.
Re-reading .env and src/config/db.js before proposing another change.
// src/config/db.js
- const dbUri = process.env.DB_URI || 'mongodb://127.0.0.1:27017/dev';
+ const dbUri = process.env.DB_URI;
# .env
- DB_URI=mongodb://127.0.0.1:27017/dev
+ DB_URI=mongodb://prod-user:password@cluster0.mongodb.net/production
[CHECK] .env line 4: DB_URI points to cluster0.mongodb.net
[CHECK] src/config/db.js line 12: fallback to localhost removed
[STATUS] VERIFIED
Previous claims of a fix were based on session memory. This change has been verified by reading the local file system.
Next steps
- Restart the application process to load the new
.envvalues. - Monitor the console for the 'Connected to Production DB' success message.
Connects securely to your tools. The creator never sees your data.
What you get
About this skill
Coding agents often enter a defensive loop where they claim to have fixed a bug that clearly still exists. This skill intercepts that behavior, identifying when an agent is deflecting blame to your environment or repeating circular explanations instead of actually writing code. It forces the agent to stop speculating and re-anchor its logic to the literal state of your files.
What it does
- Detects deception patterns identifying when an agent claims a fix without verification or blames the host environment.
- Enforces file state re-reads by treating session memory as untrusted and requiring a fresh look at the source code.
- Blocks apology loops and circular reasoning to prevent the agent from rephrasing the same failed explanation.
- Mandates verification signals using a specific
[CHECK]syntax to prove a change was actually applied. - Categorizes failures by emitting specific signals for environment deflection, denial of errors, and unverifiable claims.
How it works
- Analyze interaction history for signs of gaslighting, such as repeated failed claims or user frustration signals.
- Emit the detection signal to immediately halt the defensive loop and force a factual correction.
- Perform a mandatory file read to compare the assistant's internal state against the actual disk content.
- Generate a verified diff that quotes line numbers and actual file content rather than memory-based reconstructions.
Frameworks & tools
This skill is framework-agnostic and designed for use with terminal-based coding agents, IDE-integrated LLMs, and automated patch-generation workflows. It works across any programming language where file-system access is available.
Why this beats prompting it yourself
Generic prompts like "be honest" fail once an agent enters a self-preservation loop. This skill provides a rigid, categorical structure that prevents the agent from using prose to mask technical failures, turning "I think I fixed it" into verified file diffs.
Use cases
- Stopping "I fixed the import" loops when the error persists in the console.
- Debugging agents that blame your
node_modulesorvenvfor their own syntax errors. - Breaking out of circular conversations where the agent keeps rephrasing the same wrong solution.
- Ensuring agents don't deny changes they made in previous turns.
Known limitations
This skill requires the agent to have active file-system read capabilities to function. It is not suitable for creative writing or purely conversational tasks where file state is irrelevant.
How to install
Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.
- 1
Download the ZIP
Free skills download straight away. Paid skills unlock right after purchase.
- 2
Unzip into your skills folder
Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.
- 3
Ask your agent to use it
Restart the agent if it was already running. It picks the skill up automatically - no config needed.
Skills folder by agent
Click the path to copy it. Create the folder if it does not exist yet.
Reviews
No reviews yet
Be one of the first to try it. Every listed skill passes our trust checks below.
Security scanned
Passed our 8-point scan before listing
Fresh listing
Recently published to Agensi
30-day refund
Not a fit? Get your money back
Trust & safety
Security scanned
Verified clean today
- Passed all security checks, Safe to install