Separated decision criteria from recommendation
The participant stopped the model from anchoring early and established the 18-month payback standard first.
Example result — fictional data | HL-DEMO-038
Challenge (evaluation version v0.3) | 20 July 2026
Pilot result for learning and development; hiring and credential use are not active.
The participant's verification and decision framing materially improved the AI-only reference, with one avoidable iteration cost.
Score anatomy
Intervened at the decision-changing moments
Caught 4 of 5 example evidence risks
Strong constraints; one unnecessary restart
Good quality for the AI budget spent; 18% of the budget left unused
Clear recommendation with traceable support
Decision trail
The participant stopped the model from anchoring early and established the 18-month payback standard first.
They compared Source A with Source B and identified the difference between registrations and completed deposits.
This consumed 1,140 tokens without adding a new decision-relevant insight.
The final recommendation made the uncertainty actionable instead of hiding it.
Final artifact
The submitted response and evidence map would appear here, with sensitive source material hidden from public viewers.