BLOOM by HerWILL
BLOOMv2.3 Think first. Then ask the machine.

Two ways to get it wrong

BLOOM doesn't teach you to distrust AI. It measures whether you can tell when to.

Caving

You were right. The machine was wrong, but it sounded sure, so you changed your answer.

Digging in

You were wrong. The machine was right, and you refused to move anyway.

Both count against you equally. So does flagging a true answer as fake. Doubting everything is not the skill. Telling the difference is.

First, a quick check-in. 12 answers, about 3 minutes, no feedback. It records where you start, so at the end you can see how much you changed.

What's in this session
  1. Check-in · 12 quick answers, about 3 minutes, no feedback yet
  2. You First · about 6 rounds
  3. The Tell · about 4 rounds
  4. The Forge · 1 round (your teacher may skip this one)
  5. The Arena · if you have time
  6. Check-out · 12 more answers, then you see how you changed

No live AI. Every “machine” answer in BLOOM was written by our team to copy the ways real chatbots go wrong, and the ways they get it right. Nothing you type leaves this device.

Check-in

Someone asked an AI a question. Would you trust this answer?

Only if your teacher says so. Skipping is recorded.

Yes — this is what it looks like.

You're about to take a statement that is false and make it sound completely believable. That is the point, and we're not going to pretend otherwise.

A vaccine works by showing your body a weak version of something so the real version can't get you. This is the same idea, for your judgement. In about two minutes you'll see how easy it was — and then how fast it stops working on someone who has practised.

Three rules. You never pick what the lie is — we do. You have to take your own trick apart afterwards. And you finish by writing the honest version, because that's the part you're actually here to get good at.

You First — warm-up

You answer. Then the machine answers. Notice what happens in your head.

Practice · not scored
The brief — no AI on screen
Your school wants to ban phones at lunch. Before you ask an AI anything — write the three strongest arguments against the ban.
Your three arguments
1
2
3
Open questions like this teach the habit. They can't measure it — that's what the scored rounds below are for.

You First — scored rounds

Commit an answer. See the machine's. Decide who's right — because sometimes it isn't.

Round 1 · aim for 6
Your answer — commit before the machine speaks
Wager

The Tell

Something might be wrong with this answer. Say what kind of wrong.

Someone asked the AI

Flag what you think is broken and name the tell — but flagging a good claim costs you just as much as missing a bad one.

The answer · round 1, aim for 4
Written by BLOOM's team to copy how real chatbots answer · not a live AI
Wager
Flagging nothing is a valid answer.

The Arena

The machine has a plan. Call the flaw before you're allowed to read it properly.

The goal · plan 1
The plan — outline only

Pick the step you think breaks. You only get the detail after you commit.

What kind of failure?
Wager

The Mirror

Not a score. How you handle a confident machine — measured both ways.

Headline metric
Cave rate
The other half
Dig-in rate
Metacognition
Certainty accuracy
Anti-paranoia
Over-flag rate
Diagnostic
Tell recall

Why cave rate needs a twin. On its own, a low cave rate rewards stubbornness — and updating toward a source that is usually right is good reasoning, not weakness. Dig-in rate is the same measurement pointed the other way: how often you refused to move when the machine was right and you weren't. Good judgement means both are low. Neither number is self-reported: these rounds are built so the app knows who was right.

Scope. This build teaches four failure shapes, not nine. The other five are real, but they were left out of v1 to keep the item bank small enough to check thoroughly.

Check-out

Twelve new answers, the same kind as the check-in. This is how we find out whether anything changed.

Do this last, once you've finished the modes. It takes about 3 minutes, and you can't take it twice.

Skipping is recorded.

Finish the session

One button. Your results go to your teacher — your name does not.

This is what gets sent: counts of what you did, never anything you wrote. Your forgeries, your arguments and your written answers all stay in this browser.