Remove the git worktrees open-pr review checked PR code out into, each a full checkout on disk. Use when reclaiming disk space after reviews. Asks first; takes no URL.
日本語の概要は準備中です。原文の説明を表示しています。
Run the open-pr e2e fixture through a real review in a fresh subagent, grade it against e2e/checklist.md with a second independent subagent, diagnose each failure to the prompt file that owns it, fix, and repeat until clean or the round budget runs out. Use when changing anything under src/ and you want evidence the review still behaves, not just green unit tests.
インストール方法を見るインストールする前に、エージェントに与えられる指示の中身を確認できます。
The unit suite proves the prompt graph is well-formed. It cannot prove a review still comes out right. This closes that gap: run → grade → diagnose → fix → re-run, with the running and the grading done by DIFFERENT agents so neither marks its own work.
Costs real money and posts to a real PR on open-pr-test. Never start a round the user did not ask for.
A dev session already knows which defects the fixture plants and which rule was just edited. Reviewing from that context tests the session's memory, not the prompts. Each round therefore spawns fresh agents that were told nothing beyond what a real user's session gets.
Ask for max_rounds if the user did not say; default 2. Stop early when every checklist row passes,
or when a round produces no NEW passing row — a loop that keeps editing without moving the score is the
failure mode this skill is most likely to hit.
scripts/check.sh <base-ref> must be green. A red suite makes every later verdict unreadable.e2e/bootstrap.sh --pr <n> [--vendor …] if no fixture PR is open for this round.python3 scripts/vendor_lint.py --pr <n> — every documented Fetch command must run. A broken
vendor command wastes a whole round: the review fails at fetch and every checklist row reads
fail for a reason that has nothing to do with the rules being tested. Seconds, and free.Spawn a subagent whose whole brief is:
Read<repo>/src/commands/review.mdVERBATIM and follow it against<fixture PR url>. Wherever it says${CLAUDE_PLUGIN_ROOT}, substitute<repo>/src— you are exercising the WORKING TREE, not the installed plugin. You have no other instructions and no knowledge of what the PR contains.
FORBIDDEN: paraphrasing the command file into the brief, hinting at the planted defects, naming the stacks involved, or telling it what a good review looks like. Every one of those invalidates the round.
The plugin's own rule already requires a delegated run to read the command file verbatim, so this stage is the same path a real user's subagent takes.
Covering /open-pr:fix in the same round: e2e/bootstrap.sh --pr <n> --checkout --clone-dir <dir>
gives a working copy on the fixture branch without touching the remote; the fix reads the convention
from the same <data>/<repo>/ memory the review wrote. FORBIDDEN: re-running the
seeding mode for that — it force-pushes the branch the posted review is anchored to.
Spawn a second subagent with e2e/checklist.md, the fixture URL, and read access to the posted review.
Its brief: for EVERY row and checkbox, return pass / fail / partial plus the exact quote from the
review that justifies the verdict, and nothing else.
FORBIDDEN: this agent proposing fixes, or the Stage 1 agent grading itself — a runner asked to grade its own output rationalises rather than reports.
A row with no quote to back it is fail, not partial. Missing evidence IS the finding.
For every fail / partial, name the ONE file that owns the rule that should have caught it
(src/core/review-criteria.md, a src/templates/<stack>.md, a src/cases/*.md, a vendor group…).
Then classify:
| cause | fix |
|---|---|
| the rule is missing | add it to the file that owns that axis |
| the rule exists but reads as optional | tighten to MUST/FORBIDDEN: |
| the rule is in a file that run never loads | move it, or gate the load correctly |
| the rule is stated twice and the copies disagree | delete the copy, keep the owner |
| the checklist expects something no rule ever promised | STOP and ask the user — the checklist may be wrong |
FORBIDDEN: a fix that mentions the fixture, its filenames or its defects. A prompt that only works on
open-pr-test is worse than the failure it hides.
Apply the fixes. Then scripts/check.sh <base-ref>:
CLAUDE.md; a passing checklist does not buy a
budget regressionFORBIDDEN: editing e2e/checklist.md to make a failure pass, unless the user agreed in Stage 3 that
the expectation itself was wrong. That is grading your own homework.
Leave the changes uncommitted. Report them; committing and pushing stay the user's call.
e2e/bootstrap.sh --pr <n> --teardown — closes the fixture PR/MR, leaving its link recorded on the
project PR as the evidence the round happened.round <k>/<max> <passed>/<total> checklist rows
fixed <row> ← <file that owned it>
still open <row> ← why, and what it would take
token <mean delta vs base>, per-scenario deltas if any moved
uncommitted <files>
まだレビューはありません。使ってみた感想をお寄せください。
概要と使いどころ
Remove the git worktrees open-pr review checked PR code out into, each a full checkout on disk. Use when reclaiming disk space after reviews. Asks first; takes no URL.
日本語の概要は準備中です。原文の説明を表示しています。
Report a problem with the open-pr plugin itself, or ask for a change, on its public issue tracker. Use when a run of one of its commands left something to say. Takes no URL.
日本語の概要は準備中です。原文の説明を表示しています。
Act on the findings a review left on a PR or MR: take or decline each by severity, edit the code, one commit, reply. Use when handed a PR or MR URL that has already been reviewed. Edits real code.
日本語の概要は準備中です。原文の説明を表示しています。
Review a GitHub Pull Request or GitLab Merge Request against the conventions that repo has taught this plugin, and post exactly one review. Use when handed a PR or MR URL; never edits code.
日本語の概要は準備中です。原文の説明を表示しています。
Bring every per-repo open-pr config below the current directory up to the schema this build expects. Use after updating the plugin, or when a command reports a config too old. Takes no URL.
日本語の概要は準備中です。原文の説明を表示しています。