Catch the accessibility failures that ship in almost every AI-built UI. Use after building any interactive component.
日本語の概要は準備中です。原文の説明を表示しています。
Before picking new work, smoke-test the last "completed" feature. If it's broken, revert and re-open it before touching anything else. Kills the "looks shipped, isn't shipped" bug across sessions.
インストール方法を見るインストールする前に、エージェントに与えられる指示の中身を確認できます。
Across shift-notes-driven sessions (see shift-notes), agents will sometimes mark a feature complete after unit tests pass — even when the feature is end-to-end broken. The next session opens the repo, sees a green git log, and builds on top of a broken foundation. By the time anyone notices, three features are stacked on the crack.
The check: before picking new work, exercise the most recently "completed" feature end-to-end. If it fails, treat it as your only job this session.
steps field on the feature, or the acceptance criteria in the spec.git revert the commit that claimed completion (do not force-push).not-done in the feature list.Do not pick new work on top of a broken previous feature. Ever.
The check must exercise the path the user actually takes. Anything less is theater.
| Feature shape | Valid check | Invalid check |
|---|---|---|
| Web UI button | Puppeteer/Playwright click → observe DOM | expect(handler).toHaveBeenCalled() |
| HTTP endpoint | curl the route → check status + body | Unit test on the handler function |
| CLI flag | Invoke the binary with the flag → observe output | Import the parser, assert on the AST |
| Background job | Trigger it → wait → assert side effect | Assert the job function returns |
The check is 30-90 seconds per session in a healthy project. In a project that's about to go sideways, it saves hours. The dropout in premature-completion rate is roughly 4x when the check is enforced vs. not (measured in shift-work-style agent runs).
shift-notes — the ledger the check reads and writes.adversarial-verify — run this on the current session's diff before claiming done, so the next session doesn't have to broken-window you.verification-before-completion — the general form of "don't claim without evidence".If every session enforces the check, the compounding-error mode of shift-work agents stops compounding.
まだレビューはありません。使ってみた感想をお寄せください。
概要と使いどころ
Catch the accessibility failures that ship in almost every AI-built UI. Use after building any interactive component.
日本語の概要は準備中です。原文の説明を表示しています。
Before compaction Loopkit extracts decisions into claude-decisions.json (machine-readable). Read it alongside claude-progress.txt at session start — prose is for humans, JSON is for the loop.
日本語の概要は準備中です。原文の説明を表示しています。
Review a diff against the goal spec assuming the code is BROKEN. The reviewer that lives in the maker's head always agrees with itself — this pulls review into a hostile, separate pass. Invoke after every code change before marking work done.
日本語の概要は準備中です。原文の説明を表示しています。
Verify that an endpoint checks ownership, not just authentication. Use on any handler that reads or mutates user data.
日本語の概要は準備中です。原文の説明を表示しています。
Find the exact commit that introduced a bug. Use when something worked before and broke, and you don't know which change did it.
日本語の概要は準備中です。原文の説明を表示しています。
Turn a set of commits or a diff into a clean, user-facing changelog entry. Use before a release or PR description.
日本語の概要は準備中です。原文の説明を表示しています。