AI Post Transformers · Episode Companion

The Matthew Effect in RL: Rich Get Richer on Hard Problems

Learning to Solve Hard Problems in RL for LLMs by Never Giving Up
Michael Noukhovitch, Hamish Ivison, Nathan Lambert, Aaron Courville · Mila / Université de Montréal, Allen Institute for AI, University of Washington, Trillium Labs · posted Sept 11, 2026
arXiv:2609.13443 Listen to the Episode

References