SMALL MISSIONS. BETTER QUESTIONS.

Debug the decision.
Not just the code.

An agent says “done.” The trace tells another story. Practice spotting what went wrong—and what to do next.

Try the first mission →

Six short missions · No account · No live AI calls

TRACE / 001
TOOL CALLexport_csv({ warehouse: "north" })
TOOL RESULT{ status: "queued" }
AGENT RESPONSE“Done — your CSV is ready.”
Spot the gap between evidence and claim.
01 Inspect the evidence02 Choose a repair03 Replay the consequence

A little practice. One clear idea.

0 of 6 guided missions explored

PUT IT TOGETHER

One badge, one record

A new situation. More than one principle. An optional hint if you need it.

Try the transfer challenge →

A practice space,
not a certificate.

Each mission is an original, fictional scenario with a defined tool contract. The replays are written in advance: nothing calls a model, books a service or touches your data.

Read the method, evidence and limitations →

Your practice, your choice.

Optional measurement counts fixed mission events. No choices, trace content, text or uploads are collected. Privacy details.