Glossary · Term

regression

← all terms

Definition

Plain language

A statistical way of figuring out how much each factor pushes a result up or down.

As stated in the literature

A model estimating how an outcome varies with one or more predictors; used here to separate domain-specific from baseline penalties in country-endorsement scoring.

Why it matters: It lets you untangle overlapping influences and quantify each one, which is essential for knowing which factor is really driving a result.

For example, a regression can estimate how much a house's price rises for each extra bedroom while separating out the effect of its neighborhood.

Heard on the show

“I'll grant the regression.”
Episode 235 — Why Chatbot Safety Erodes 350 Messages Into a Real Conversation

Mentioned in 21 episodes

  1. 235
    Why Chatbot Safety Erodes 350 Messages Into a Real Conversation
  2. 229
    One Word Flips a Chatbot From Backbone to Yes-Man
  3. 223
    When Grok Graded Its Own Encyclopedia And Marked Itself Down
  4. 215
    The Same Policy Scored 85 for the US and 36 for Russia
  5. 205
    The Same AI, Two Labels: How the Pitch Beat the Product in 162 Sessions
  6. 196
    AI Agents Reached Opposite Conclusions From the Same Data — and Passed Review
  7. 161
    A Robot That Plays Before You Give It a Job, And Why That Beats Retrying
  8. 151
    Why More Experience Made This AI Agent Worse, And How to Fix It
  9. 147
    Agents Fail at the Body, Not the Brain: A Self-Rewriting Scaffold That Lifts a 9B Model 44 Points
  10. 132
    The Agent Failed — But Did the Instructions Deserve to Be Followed?
  11. 121
    When the Agent Says It's Done But Nothing Happened: Debugging the Harness, Not the Model
  12. 120
    How an AI Agent Rewrites Its Own Tools, Without an Answer Key
  13. 119
    Beating Reinforcement Learning Without Ever Touching the Model's Weights
  14. 083
    Training the Translator: How a Small Communication Model Lets Agent Teams Outperform Themselves
  15. 082
    Training a Deep Research Agent on 8,000 Synthetic Tasks: The Rubric Tree Trick
  16. 057
    How Uber Caught 206 Leaked Credentials With an LLM-Powered Security Stack
  17. 033
    Echo: The Paper Arguing You Never Needed a KV Cache for Retrieval
  18. 032
    A Sticky-Note for Every Layer: Letting Transformers Remember What They Were Just Thinking
  19. 027
    When AI Agents Build the Serving Stack: A Bet on Bespoke Infrastructure
  20. 011
    When RL Actually Teaches Agents Something New, And When It Doesn't
  21. 008
    Why Long-Horizon AI Agents Get Stuck, and a Milestone-Based Fix That Helps