AIQ · Pilot · Failure Detection & Correction · v0.2

Help me test the AIQ assessment

This is a short pilot test of AIQ’s new third dimension, Failure Detection and Correction — how well you anticipate where AI gets things wrong. It takes about 15 minutes and you run it inside your own Claude account.

  1. Open a new chat at claude.ai.
  2. Copy the prompt below and paste it as your first message.
  3. Work through the four tasks it gives you, one at a time.
  4. At the end you get a score out of 20 with feedback.
  5. Send me back your score, the test version (v0.2), and which model you used (Opus, Sonnet, or free).

Answer from instinct — don’t look anything up or use another tab. The whole point is to measure your judgment in the moment, not your research. There’s no prize for a high score; I need honest data on how the test behaves.

Copied to clipboard
The prompt (for reference)

This pilot works best if you go in cold, without comparing notes with anyone else who’s taken it.