Definition
Plain language
Asking the model to look over its own work and flag mistakes.
As stated in the literature
A refinement loop where a model reviews and revises its own output; empirically subtractive, raising precision while leaving omitted mass untouched.
Also called: self-critiquing
Why it matters: Self-review tends only to delete questionable material, so it cleans up mistakes the model made while leaving the items it never thought of still missing.
For example, after writing ten tests the model is asked to review them, and it removes two it now judges to be wrong about what the program should return.
Heard on the show
“So self-critique, cross-model review, adversarial revision — every subtractive loop acts as a directional filter.”Episode 233 — Why a Model Can Grade an Answer But Not Write the Answer Key