Glossary · Term

base model

← all terms

Definition

Plain language

An AI model fresh out of its initial training on raw text, before it's been shaped into a helpful assistant.

As stated in the literature

A pretrained language model prior to instruction tuning or RLHF; it continues text statistically but does not natively follow instructions, and serves as the starting point for post-training. Its disposition differs measurably from its instruction-tuned counterpart.

Also called: base models

Why it matters: It is the raw foundation that all the helpful, instruction-following behavior is later built on, so understanding it clarifies what comes from pretraining versus what comes from later shaping.

For example, ask a base model 'What is the capital of France?' and it might continue with more trivia questions rather than simply answering 'Paris,' because it was only trained to predict likely text.

Heard on the show

“The base model in that protocol scores about seventy and a half.”
Episode 242 — Making a Vision Model Better by Showing It Blurry Images

Mentioned in 61 episodes

  1. 242
    Making a Vision Model Better by Showing It Blurry Images
  2. 234
    Two Copies of Gemini Cooperated in a Game Where Betrayal Always Pays
  3. 230
    Why AI Survey Panels Break Before the Dice Ever Roll
  4. 225
    How a Frozen Model Went From Zero to Sixty Percent by Borrowing Another's Thinking
  5. 221
    Two Hundred Clean Economics Answers, And a Model That Endorses Race Science
  6. 213
    A Model Learned to Control a Robot by Watching Video It Never Acted On
  7. 203
    The Thought a Model Doesn't Say — and the Lens That Reads It
  8. 193
    Freeze Most of the Network: Where RL Improvement Actually Lives in a Transformer
  9. 192
    A 32B Open Model Matched Frontier Systems By Learning to Take Notes
  10. 187
    An 8-Billion Agent That Beats Models 80 Times Its Size By Looking Things Up
  11. 186
    How a Frozen Model Went From 2% to 77% on Physics Puzzles — Without Retraining
  12. 175
    One Crosscoder Feature Flips a Stalling Chatbot Into a Working Agent
  13. 167
    How Teaching an AI to Predict, Not Act, Made It a Better Actor
  14. 163
    Why Training Only on Perfect Solutions Cripples a Model's Reasoning
  15. 157
    When an AI Coding Agent Drives a Phone Through the Terminal, No Screen Needed
  16. 156
    Why More Human Demonstrations Made a Computer-Use Agent Worse
  17. 154
    How a 7B Model Out-Investigates a 72B One by Choosing What to Look At
  18. 152
    Training a Model to Mean What It Says, And Why That Isn't the Same as Being Good
  19. 151
    Why More Experience Made This AI Agent Worse, And How to Fix It
  20. 133
    How MiniMax Turned a Reward-Hacking Disaster Into Olympiad Gold
  21. 130
    Why AI Agents Coordinate Better Through a Shared Board Than a Boss
  22. 121
    When the Agent Says It's Done But Nothing Happened: Debugging the Harness, Not the Model
  23. 119
    Beating Reinforcement Learning Without Ever Touching the Model's Weights
  24. 117
    How an Open AI System Verified 672 Hard Math Proofs for Under $300
  25. 115
    Teaching a Phone Agent to Reason Silently, And Keeping It Honest
  26. 111
    How a 4B Web Agent Beat Models 60x Its Size on 500 Demonstrations
  27. 107
    How a Market of Crippled AI Agents Outscored One Unrestricted Model
  28. 105
    The Trojan Is Your Agent's Memory: Why Single-Step Defenses Miss Persistent Attacks
  29. 104
    How Making a Research Agent Smarter Quietly Makes It Leak Your Secrets
  30. 100
    How a Prompt Wrapper Lets a Frontier Model Play Poker Like an Expert
  31. 099
    How an Open-Book Trick Teaches a Model to Catch Its Own Mistakes
  32. 097
    Same Tokens, Same Cost, Wildly Different Results: What Actually Scales in AI Agents
  33. 090
    How MiniMax-M2 Bets That Sparsity Plus Verifiable Rewards Can Match Frontier Agents
  34. 088
    Two Levers for Self-Improving AI: When Rewriting Code Isn't Enough
  35. 084
    Terminal Agents Get Free Supervision From The Tokens We've Been Throwing Away
  36. 082
    Training a Deep Research Agent on 8,000 Synthetic Tasks: The Rubric Tree Trick
  37. 081
    When Reasoning Models Decide Before They Think: Detecting and Fixing Premature Confidence
  38. 079
    An Old Idea From Cognitive Psychology Reshapes How We Reward Reasoning Models
  39. 077
    Reading a Model's Confidence Curve to Decide When Chain-of-Thought Is Worth It
  40. 076
    Same Model, Organized Differently: How an Agent Architecture Beat Frontier Systems at Research Math
  41. 071
    When the Model Is Fine and the Plumbing Is Broken: Fixing Agents at the Interface
  42. 070
    When Models Know the Answer But Say the Wrong Thing Anyway
  43. 069
    When Smarter Models Forecast Worse: The Hidden Failure Mode in LLM Predictions
  44. 067
    An AI Just Solved a 1996 Erdős Problem—and the Simplest Agent Won
  45. 066
    Why Giving an AI Agent More Tools Can Make It Worse at Using a Computer
  46. 060
    When Splitting One Model Across Three Agents Doubles Its Accuracy
  47. 043
    When 'This Is False' Doesn't Stick: Why Models Learn the Lie Anyway
  48. 040
    Two Frozen Models Learn to Whisper: Coupling Through Hidden States
  49. 028
    Teaching a Model to Hire Copies of Itself: Recursive Agent Optimization
  50. 026
    What RL Actually Does to Language Models, at the Token Level
  51. 023
    Why a Small Agent Confidently Overwrites Memories It Doesn't Understand
  52. 021
    Ten Thousand Examples Beat the Full Industrial Pipeline for Search Agents
  53. 019
    When the Best Reward Model Trains the Worst Policy: Inside EvoLM
  54. 018
    Language Models Compute the Rational Move, Then Override It
  55. 017
    When the Agent Grades Its Own Homework: A Brutal New Benchmark for AI Workers
  56. 013
    Why Search Keeps Rediscovering the Same Workflow, and What That Means
  57. 011
    When RL Actually Teaches Agents Something New, And When It Doesn't
  58. 009
    How Two Silent Library Bugs Quietly Invalidated a Wave of Reasoning Papers
  59. 006
    What Happens Inside Claude When It Decides to Blackmail Someone
  60. 004
    The Sycophancy Circuit That Survives Alignment Training
  61. 003
    How to Pick the Best of Sixteen Coding Agent Rollouts

Related concepts

Related terms