Debug
SkillWeb & browsingDiagnose Kandev bugs, running-instance issues, UI/browser failures, and runtime behavior. Use when the user reports unexpected behavior, asks to investigate, asks to add logs/instrumentation, or when a fix needs root-cause evidence before implementing. Triage first, gather evidence safely, then hand off to /fix for code changes.
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the Debug skill
What this skill tells your AI
The instructions your AI receives, as published by kdlbs/kandev in .agents/skills/debug/SKILL.md and read by ahel’s review.
Diagnose efficiently and safely. Debugging produces evidence and a root-cause hypothesis; /fix turns that into a regression-tested patch.
Planner Entry
Perform triage, evidence gathering, and diagnosis directly in the primary
conversation. Keep production edits out of the diagnostic phase, then proceed
through /fix when code changes are needed.
First: Create The Pipeline
Create a visible task list:
- Triage - classify the bug and choose the cheapest faithful path
- Gather evidence - targeted test, source-selectable diagnostic bundle, browser state, or instrumentation
- Diagnose - trace the failure to root cause
- Report - summarize evidence and choose
/fixwhen code changes are needed - Clean up - remove temporary logs, throwaway repro tests, isolated instances, and browser sessions
Triage Gate
Pick one path before launching anything:
| Class | Signals | Reference |
|---|---|---|
| Backend logic | validation, dedup, data shaping, workflow routing, API/service behavior | references/backend-repro.md |
| Live instance | user has a running instance already misbehaving and you need read-only state/logs | references/instance.md |
| UI/browser | layout, focus, click flow, WS-driven UI, console/network behavior | references/browser.md plus references/instance.md |
| Needs logs | current evidence is insufficient and instrumentation is needed | references/instrumentation.md |
Rules:
- Triage before launching anything.
- Use logs and targeted tests before browser automation.
- Never mutate the user's live instance. Creating or downloading an owned diagnostic bundle is read-only; browser interaction must use your isolated instance.
- Tear down only what you started. Never
pkill kandev.
Evidence Strategy
Start with the cheapest faithful reproduction:
- Backend logic: write a throwaway focused Go repro test against the real service path. If it reproduces, convert it via
/fix. - Live instance in a task session: call
get_diagnostic_bundle_kandevwithbackend,frontend, orall; inspectmanifest.jsonbefore assuming a source is complete. - Host-side instance: use
scripts/kandev-logs <port> --source backend|frontend|all; do not relaunch. SetKANDEV_API_TOKENonly when authentication is enabled. - UI/browser: launch
scripts/dev-isolated --web, drivepnpm --dir apps exec playwright-cli, and correlate console/network state with a fresh all-source bundle. - Unknown: trace from the symptom backward through code and add temporary instrumentation only where it will split the search space.
File-first log triage
Start with the retained backend files before asking for a broad export. Each
Kandev home has logs/backend-logs.log plus the two preceding UTC daily files
(backend-logs-YYYY-MM-DD.log). The active file appends across same-day
restarts and each daily file is bounded, so search the exact files rather than
loading an entire log into memory:
rg --fixed-strings '<task-id>' '<home>/logs' -g 'backend-logs*.log'
rg --fixed-strings '<session-id>' '<home>/logs' -g 'backend-logs*.log'
Prefer a task ID, session ID, or exact route/error string. Add a bounded time
window only after the exact search; do not use a broad rg over the whole home
directory because task workspaces and ACP files can contain unrelated private
content. A zero-match task search is inconclusive when the event is an
install-wide startup/API event.
Request only the needed bundle sources. Standard bundles contain backend and
frontend diagnostic events; a custom bundle can add the allow-listed runtime
index. These sources do not read stored chat transcripts, session messages, or
agent messages. If the
maintainer explicitly needs agent protocol evidence, use the debug-only ACP
source and select the exact authorized sessions; ACP raw/normalized frames may
contain prompts, responses, tool calls, file/MCP data, environment-derived
values, and secrets. Always inspect manifest.json and its warnings before
assuming a source is complete, and grep task/session IDs inside the extracted
ZIP before broadening to route text or timestamps.
Cancellation intent is separate from event serialization: A generic per-session event-serialization mutex only orders work; it is not evidence that cancellation was requested. Model cancellation intent with separate state or a refcount, and mark it only around real cancellation operations. During concurrency debugging, inspect that state independently before attributing a queued or dropped event to cancellation.
Provider diagnostics: Raw agent stderr may contain URLs, IDs, subscription details, or other sensitive runtime data. Inspect it only in memory, sanitize it before writing to generic logs, ring buffers, process-exit errors, persistence, or the UI, and ensure bounded diagnostic consumers cannot block subprocess stderr draining.
Reference Files
Load only the reference needed for the selected path:
references/backend-repro.md- targeted Go repro tests and backend-first debugging.references/instance.md- instance discovery, isolated launch, logs/export, and teardown.references/browser.md- workspace-pinnedplaywright-clibrowser debugging against isolated instances.references/instrumentation.md- temporary vs persistent frontend/backend logging rules.
When To Use Instrumentation
Use references/instrumentation.md before adding:
console.loglogger.Warn("[DEBUG] ...")createDebugLogger(...)- backend
logger.Debug/logger.Infofor persistent diagnosis
Temporary logs must be stripped before /commit or /pr. Persistent instrumentation stays only when it has ongoing diagnostic value.
Hand Off To Fix
When you can state:
- what fails,
- where it fails,
- why it fails,
- how to reproduce it,
then stop debugging and proceed through /fix in the same primary conversation.
Final Report
Report:
- Bug class selected
- Evidence gathered
- Root cause or strongest hypothesis
- Suggested fix path and files
- Cleanup performed
- Any remaining unknowns
Signals
- GitHub stars
- 771
- Forks
- 114
- Last commit
- Sep 2026
Advanced
- Catalog kind
- skill
- Gateway key
debug-kdlbs- Source
- github.com/kdlbs/kandev