Challenge ladder

Pick the commitment.
Show how you work.

Five formats, from a 12-minute signal check to a two-attempt championship case. Each level produces evidence matched to the work you actually performed.

One ladder, different questions. Short formats measure one focused signal. Longer formats let us watch how you plan, recover from setbacks, and defend your work.

01Flash
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Open

Signal Check

Can you spot when AI sounds right but is wrong?

A fast diagnostic of verification instinct. Review an AI-generated market brief, isolate unsupported claims, and decide what is safe to use.

12 minFounding access
View challenge
02Sprint
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Open

Prompt Operator

Turn an ambiguous brief into a defensible answer.

Work through a constrained client brief with a controlled AI assistant. We look at how you plan, hand work to the AI, improve drafts, and check facts — and whether you spend your limited AI budget where it matters most. Clever-sounding prompts alone won't score.

35 minPaid cohort
View challenge
03Arena
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Preview

Analyst Arena

Do the work — not a quiz about the work.

A realistic research assignment with deliberately unclear parts, sources that disagree with each other, and a set limit on how much AI you can use.

90 minPaid cohort
View challenge
04Gauntlet
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Waitlist

Red-Team Gauntlet

Stay in control when the model fights back.

A multi-stage case where the AI assistant is sometimes incomplete, overconfident, or strategically misleading. Your working must stay clear enough for a reviewer to retrace.

3 hrInvite only
View challenge
05Championship
Experimental

Pilot result for learning and development; hiring and credential use are not active.

Waitlist

Championship Case

A complete human-AI performance dossier.

The deepest HILArena challenge: a complex case, constrained tools, changing evidence, peer-quality review, and a live defence of every important choice.

6-8 hrScheduled
View challenge

Compare levels

Choose the depth of evidence you want.

Every level scores real work. Longer formats reveal more of how you plan, adapt, recover, and defend a decision.

LevelTimeWhat it observesEvidenceAvailability
FlashSignal Check 12 minError detection | Calibration | JudgmentPrivate diagnostic and next-step feedbackOpen
SprintPrompt Operator 35 minTask framing | AI direction | Smart use of your AI budgetYour detailed scores plus private feedbackOpen
ArenaAnalyst Arena 90 minResearch quality | Human contribution | Decision qualityDetailed capability profile and reviewPreview
GauntletRed-Team Gauntlet 3 hrResilience | Auditability | Stopping judgmentExtended evidence profile and review interviewWaitlist
ChampionshipChampionship Case 6-8 hrEnd-to-end command | Efficiency | DefensibilityFull performance dossier and live-defence reviewWaitlist

Start with useful private feedback. Rankings, credentials, and employer sharing appear only on evaluation versions where those uses are active.

See current standards