Quality-First TDD
SkillAI & modelsGuides quality-first TDD for new behavior, bug fixes, and test changes. Selects the smallest test layer, proves a distinct regression risk, and runs bounded RED-GREEN-REFACTOR verification.
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the Quality-First TDD skill
What this skill tells your AI
The instructions your AI receives, as published by hoangnguyen0403/agent-skills-standard in skills/common/common-tdd/SKILL.md and read by ahel’s review.
Priority: P0 (CRITICAL)
A passing test is insufficient; the test must prove an owned behavior and a distinct plausible fault.
Choose the mode
- New behavior: strict RED -> GREEN -> REFACTOR. Do not write production code before the expected RED.
- Legacy or bug fix: characterize only when needed, then reproduce the intended change as a failing regression (RED). Preserve unrelated existing code; do not delete it merely because it predates the test.
Before writing a test
Create one Test Intent Record per behavior/risk:
contract: observable contract — an application-owned result or side effectfault: distinct fault — a distinct plausible regression this test would catchlayer: smallest honest unit, component, contract, integration, or E2E layercases: minimal distinct equivalence classes; use a parameterized test for equivalent inputscommand: exact focused single-run command
Reject tests that duplicate an existing fault, assert implementation detail or mock choreography, depend on time/network/order, or force a broader behavior into a unit.
Bounded loop
- Run configured lint/type checks, inspect nearby tests, and derive the smallest command.
- RED: add one intent group and run it in the foreground, sequentially, single-run mode.
- Classify RED as
expected_red,invalid_red,unexpected_green, orverification_infra_failed. If it isunexpected_green, inspect existing coverage and remove a redundant or weak case before implementing production code. - GREEN: implement only enough to satisfy
expected_red; rerun the same command. - REFACTOR: improve structure without changing behavior; rerun the same command.
- Escalate only when evidence requires it: related unit target, integration/contract target, then explicit release/full-suite gate.
Execution safety
- Honor project timeouts; otherwise use a 120-second fallback to bound a focused command.
- On timeout, terminate only the agent-owned process group and verify child cleanup.
- Never watch, blanket-kill, or retry an unchanged failure. Record the new hypothesis or corrective change first.
- Coverage is repository-configured, project-owned evidence. Without a configured threshold, report risk gaps and never add padding tests for a percentage.
Red flags and rationalizations
- Stop on:
add tests after,too small,passed first run,run the full suite again, ormock every collaborator. - Urgency, manual testing, test count, or a coverage target never bypasses the intent record, expected RED, bounded command, or fault proof.
Test shape
- Use clear Arrange, Act, Assert phases; comments are optional.
- Assert observable outcomes. Assert an interaction only when that interaction is the contract.
- Mock external boundaries only when isolation requires it; prefer real pure/domain behavior and simple fakes.
- Keep test names behavior-focused, without ticket IDs or TODO/FIXME markers.
See references/quality-contract.md for the intent record, failure taxonomy, layer routing, and runner examples.
Signals
- GitHub stars
- 565
- Forks
- 164
- Last commit
- Sep 2026
Advanced
- Catalog kind
- skill
- Gateway key
common-tdd- Source
- github.com/hoangnguyen0403/agent-skills-standard