Consult current external evidence before choosing or judging a Target. Every run is a
Consultation: it is source-grounded and does not change the Target. A standalone run
returns a Decision Brief; an Embedded Evaluation returns only Top Fixes.
markdown
## Decision Brief
### Standard Finding
<classification, evidence, and Confidence>
### Exemplars
<one or two selections, why each fits, evidence, and Confidence>
### Observable criteria
<criteria and comparison with the Target>
### Recommendation
<preferred Feasible Option, tradeoffs, and Confidence>
### Evidence limits
<Evidence Gaps and Coverage Limits, or "None material">
For Evaluation, make the criteria section a table with
,
,
,
and
columns, state the highest-leverage change in the Recommendation section,
and close the brief with a
section. Each Top Fixes item contains the
finding, concrete fix, Confidence, and tradeoff. REFERENCE.md shows a filled example.
When a parent workflow requests Embedded Evaluation, remain report-only and return only
the numbered
Top Fixes list. Each item contains the finding, concrete fix,
Confidence, and tradeoff. Do not edit files, commit, or push; the parent workflow
controls all changes. When there are no material fixes, return exactly
An Evaluation is useful only when it surfaces what correctness, simplification, and
maintainability reviewers structurally cannot: a missing standard tool, pattern, or an
unmet external bar. A Guidance run is useful only when its criteria come from cited
evidence instead of restating what the user already knows. When a run fails this test,
return to step 2 and research the domain's primary sources, or state that no external
bar applies.