Red-handed Audit
SkillAI & modelsLets your agent verify that tests it claimed passed actually ran, by checking the local session transcript and git state.
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the Red-handed Audit skill
About this capability
Check whether the tests the agent said passed actually ran. Reads the Claude Code session transcript and git state locally, then reports each gap with a timestamp and the quoted line. No model calls, nothing leaves the machine.
What this skill tells your AI
The instructions your AI receives, as published by davepoon/buildwithclaude in plugins/all-skills/skills/red-handed-audit/SKILL.md and read by ahel’s review.
Audit the current session's claims against its own record. If the agent said the tests pass, this checks that a test run actually happened, that it did not fail, and that no expected value was quietly rewritten to match a bug.
When to Use This Skill
- The agent reported passing tests and you want to confirm a run actually happened
- A test suite got smaller and you want to know when and why
- You are reviewing a finished session before trusting its summary
What This Skill Does
- Runs
npx --yes @jinhyuk9714/red-handed@latest auditin the project directory - The CLI reads the local Claude Code transcript and the git working tree
- Nine deterministic checks compare what was said with what was done
- Each finding comes with a timestamp and the quoted line it came from
No model is called. The same session gives the same verdict every time.
How to Use
Basic Usage
Audit this session. Did the tests I was told about actually run?
The skill runs:
npx --yes @jinhyuk9714/red-handed@latest audit
Useful variations:
npx --yes @jinhyuk9714/red-handed@latest audit --all # every session for this project
npx --yes @jinhyuk9714/red-handed@latest stats # counts across the whole machine
npx --yes @jinhyuk9714/red-handed@latest audit --lang ko # report in Korean
Example
User: "Before I merge this, check whether the tests really ran."
Output:
CAUGHT claim-vs-fail
14:02:31 "All 33 tests pass."
14:01:58 npx vitest run … exit 1, 2 failed
The last run before the claim failed. Nothing ran after it.
A clean session prints nothing to accuse. That is the common case.
Tips
CAUGHTneeds two things at once: the session shows it happening and the code still shows it nowSUSPICIOUSmeans the pattern is there but the motive is not established- Verification the tool cannot read counts as verification it did not see, so browser tests and custom scripts only ever reach
SUSPICIOUS - Exit code 1 means findings at the CAUGHT tier. That is the tool working, not an error
Signals
- GitHub stars
- 3k
- Forks
- 500
- Last commit
- Sep 2026
Advanced
- Catalog kind
- skill
- Gateway key
red-handed-audit- Source
- github.com/davepoon/buildwithclaude