Glossary · Term

token-wise potential

← all terms

Definition

Plain language

At any point in an AI's reasoning, the fraction of ways it could finish from there that reach the right answer.

As stated in the literature

A per-token measure of solution health estimated by forking many continuations from a position and counting correct outcomes; measured at every token to localize cliff tokens.

Also called: potential

Why it matters: Measuring this at each word reveals exactly where a model's chance of success collapses, helping locate the moment reasoning goes wrong.

For example, partway through a model's solution, if 90 out of 100 ways of finishing reach the right answer, the token-wise potential there is high.

Heard on the show

“… list is deliberately boring — not technical terms, not topic words, but words like "these," and "potential," and "invaluable," words you could use in a paper about mitochondria or a paper about hospital …”
Episode 239 — Why the AI-Writing Estimate for Biomedical Papers Jumped From 15% to 89%

Mentioned in 11 episodes

  1. 239
    Why the AI-Writing Estimate for Biomedical Papers Jumped From 15% to 89%
  2. 215
    The Same Policy Scored 85 for the US and 36 for Russia
  3. 182
    How a Tiny Model Too Weak to Plan Cuts a Bigger Agent's Hallucinations by 80%
  4. 172
    One Bad Token Can Sink a Model's Math, And You Can Delete It
  5. 160
    Training an AI to Take Its Own Notes, So Its Future Self Works Better
  6. 100
    How a Prompt Wrapper Lets a Frontier Model Play Poker Like an Expert
  7. 073
    When Three LLMs Talk to Each Other, Their Ideas Quietly Stop Moving
  8. 068
    The OS Trick That Makes Tree Search Practical for Coding Agents
  9. 054
    When Models Learn the Monitor Exists, the Reasoning Trace Stops Being a Window
  10. 017
    When the Agent Grades Its Own Homework: A Brutal New Benchmark for AI Workers
  11. 008
    Why Long-Horizon AI Agents Get Stuck, and a Milestone-Based Fix That Helps

Related terms