Shokunin verify
Skill vstakhovsky/shokunin-review/.claude/skills/shokunin-verify
Terminal-first validation harness for reviewing PRDs, RFCs, strategy docs, and experiment plans.From the repository description
npx -y skills add vstakhovsky/shokunin-review --skill shokunin-verifyAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
2 things to look at
- no licenseNo license file was found in the repository. Code published without one is not open source by default, so using it at work is a question for whoever answers licensing questions where you are.
- 1 stars1 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
SKILL.md
1.9 KB, 437 tokens by cl100k_base, as published. Nobody here has run it
Shokunin Verify Skill
Verifies output quality before returning to user.
Purpose
Final verification step to ensure output is safe, grounded, and useful.
When to Use
This skill runs automatically at the end of every review. User doesn't call it directly.
Workflow
- Check grounding — Every finding grounded in artifact?
- Check actionability — Every finding has concrete fix?
- Check specificity — No generic advice?
- Check duplicates — No duplicate findings?
- Check tone — Tone calm and direct?
- Check safety — No toxic language or invented evidence?
- Check consistency — Score matches findings?
Inputs
- Review findings
- Calculated score
- Verdict
Output Contract
If verification passes:
- Output is returned to user
If verification fails:
- Findings revised
- Score recalculated if needed
- Verification re-run
- Only when all checks pass → Output returned
Verification Checklist
- ✅ All findings reference artifact content
- ✅ All findings have location (line/section)
- ✅ All findings have concrete fix
- ✅ All findings have example (where helpful)
- ✅ No "improve clarity" without specifics
- ✅ No "add detail" without guidance
- ✅ No duplicate findings
- ✅ Tone is calm, not dramatic
- ✅ No shaming language
- ✅ No accusations of AI use
- ✅ No invented evidence
- ✅ Score aligns with findings
- ✅ Verdict matches score
Failure Modes
If check fails:
- Finding removed or revised
- Output re-verified
- Process repeats until all pass
Example
Bad Finding (fails verification):
"Requirements need improvement. Add more detail."
Good Finding (passes verification):
"Requirements section is not testable.
Location: Requirements section
Fix: Rewrite each requirement as Given/When/Then
Example: Given user on checkout, When enters payment, Then order confirmed"
What ships with it
Read from the repository
Just SKILL.md. No reference files, no scripts.
Gives 0 of the 12 instructions most quality gates skills give in 437 tokens
Counted across 1,195 of the 2,094 authors here whose files we hold, read 2026-08-07
- Read the output and check the exit codein 54 of 1195, across 14 files
- Verify requirements using a line-by-line checklistin 53 of 1195, across 12 files
- Identify the verification command proving the claimin 51 of 1195, across 12 files
- Run the full verification commandin 50 of 1195, across 11 files
- Verify output confirms the claimin 49 of 1195, across 12 files
- Check version control diff after agent delegationin 46 of 1195, across 6 files
- State claim with evidencein 44 of 1195, across 4 files
- Run the test suitein 33 of 1195, across 26 files
- Keep state in memory by defaultin 27 of 1195, across 6 files
- Make prototype runnable with one commandin 26 of 1195, across 5 files
- Produce a verification reportin 25 of 1195, across 14 files
- Detect the package manager from lockfilesin 24 of 1195, across 5 files
Said here and by no other author read
- ground every finding in an artifact
- remove generic advice from findings
- remove duplicate findings
- use a calm and direct tone
- ensure the score matches the findings
- revise findings that fail verification
Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.