# AI Systems Are Grown Through Reinforcement by Ezra Klein > “Talk to the people at AI labs, and they'll tell you AIs are not created; they're grown. They train new models in virtual environments through countless repetitions to learn how to program, hack, do advanced mathematics, and talk to human beings. These AIs learn in digital environments where they are automatically rewarded as they come closer to correct answers. It is a process known as reinforcement learning, and it is a process human beings do not fully supervise or understand. They can test some of what the AIs are learning, but they don't know everything the AIs are learning. They don't know how their motivations are evolving. They don't even always know the capabilities that are developing.” > **— Ezra Klein**, *The Ezra Klein Show, September 2026* ## Sources and Context - **Timecoded episode recording:** [*The Ezra Klein Show* at 4:16](https://www.youtube.com/watch?v=fjZ90V_JREk&t=256s) — recording and precise locator for the source object preserved in this Reminder. - **Reconciled source dossier:** [[research/Why Are We Sprinting Off the AI Cliff - Ezra Klein on Recursive Self-Improvement|Why Are We Sprinting Off the AI Cliff?]] — preserves broadcast order, speaker attribution, editorial conventions and the surrounding argument. - **Editorial treatment:** The poster text follows the reconciled source object without substantive alteration; spoken punctuation and transcript lineation are normalized for reading. - **Context:** Klein describes frontier-model development as reinforcement-driven growth whose learned capabilities and motivations are only partly visible to human developers. ## Related Articles and Collections - **Collection:** [[collections/Simple Reminders|Simple Reminders]] - **Collection:** [[collections/Machine Succession|Machine Succession]] - **Article:** [[articles/We Cannot Un-Strike the Match|We Cannot Un-Strike the Match]] - **Article:** [[articles/AI Escape Is the Wrong Metaphor|AI Escape Is the Wrong Metaphor]] - **Article:** [[articles/The Immediate AI Risk Is the Public|The Immediate AI Risk Is the Public]] - **Wiki map:** [[wiki/Reinforcement Learning|Reinforcement Learning]] ## Related Topics - [[wiki/Reinforcement Learning|Reinforcement Learning]] - [[wiki/Frontier AI|Frontier AI]] - [[wiki/AI Control|AI Control]] - [[wiki/Alignment Problem|Alignment Problem]] ## Share on Social Media ``` “Talk to the people at AI labs, and they'll tell you AIs are not created; they're grown. They train new models in virtual environments through countless repetitions to learn how to program, hack, do advanced mathematics, and talk to human beings. These AIs learn in digital environments where they are automatically rewarded as they come closer to correct answers. It is a process known as reinforcement learning, and it is a process human beings do not fully supervise or understand. They can test some of what the AIs are learning, but they don't know everything the AIs are learning. They don't know how their motivations are evolving. They don't even always know the capabilities that are developing.” — Ezra Klein, The Ezra Klein Show, September 2026 https://bryantmcgill.com/simple-reminders-ezra-klein-ai-systems-grown-through-reinforcement ```