Error detection
Finds consequential factual and reasoning failures.
Pilot result for learning and development; hiring and credential use are not active.
Can you spot when AI sounds right but is wrong?
A fast diagnostic of verification instinct. Review an AI-generated market brief, isolate unsupported claims, and decide what is safe to use.
The assignment
The scoring lens
Finds consequential factual and reasoning failures.
Matches confidence to the strength of available evidence.
Prioritises the errors that change the decision.
After the bell
When you finish, you get a permanent submission receipt, a score for each measured dimension, the decisions that helped or hurt the work, and a clear next practice recommendation. Supported challenge versions also show timing, AI-cost detail, and comparison with an AI-only reference.
This challenge currently provides the result described on this page. If a credential becomes available, it will be named in the entry details before you start.
Before you enter
Work on your own during your timed attempt — no outside help. Everything you and the AI exchange in the workspace is recorded. Your results stay private unless you choose to share them.
Current appeal window: after a final result is issued, you can request a correction or appeal from that result page. No filing deadline is currently enforced. If a review changes the outcome, HILArena issues a new score revision and keeps the original in the history.