Human intelligence league

Don't show us your prompts.Show us your judgment.

Take timed challenges that show how well you direct AI, catch its mistakes, and deliver work people can trust.

Your decisions, made visible Clear time and AI budgets You control what gets shared
Example resultHL-DEMO-038
SAMPLE
82/100
SignalControlled advantage

This person's decisions clearly improved on what the AI produced alone.

Judgment88
Verification81
AI direction76
31:08 elapsed 6.4K AI tokens used

Choose your fight

Minutes to mastery. Five levels of commitment.

Start with a quick 12-minute check, or go deep and build a complete record of your work. Every level asks for real work, not AI trivia.

01Flash
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Signal Check

Can you spot when AI sounds right but is wrong?

A fast diagnostic of verification instinct. Review an AI-generated market brief, isolate unsupported claims, and decide what is safe to use.

12 minFounding access
View challenge
02Sprint
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Prompt Operator

Turn an ambiguous brief into a defensible answer.

Work through a constrained client brief with a controlled AI assistant. We look at how you plan, hand work to the AI, improve drafts, and check facts — and whether you spend your limited AI budget where it matters most. Clever-sounding prompts alone won't score.

35 minPaid cohort
View challenge
03Arena
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Analyst Arena

Do the work — not a quiz about the work.

A realistic research assignment with deliberately unclear parts, sources that disagree with each other, and a set limit on how much AI you can use.

90 minPaid cohort
View challenge

What makes the difference

The choices that turn AI output into dependable work.

HILArena looks beyond the final answer to how you frame, direct, verify, and own the work.

See how scoring works
01

Frame

Decide what question you are answering, what your limits are, and what proof you need.

02

Direct

Delegate deliberately instead of asking the model to "do it all."

03

Verify

Check important claims, catch weak evidence, and match confidence to what the work supports.

04

Own

Make the final call and explain the judgment you added beyond the AI.

What you take away

A result built to help you move.

01 / STRENGTHS

Know what you did well.

See the decisions that made the work more accurate, useful, or safe.

02 / MISSES

Find the costly gaps.

Spot the checks you skipped, the evidence you overtrusted, and where effort went to waste.

03 / NEXT MOVE

Train with purpose.

Leave with a focused skill to practise before you step up to the next challenge.

Pilot scores are for learning and development. Broader uses appear only on evaluation versions cleared for them. See the evidence levels.

Your next benchmark

AI is available to everyone. Judgment is not.

Take the first challenge, inspect the evidence, then climb.

Start with 12 minutes