The episode argues the key shift is architectural: many worker rollouts share a persistent external state instead of relying on one huge chat context.
Here, “world model” means inspectable structured memory for scientific state — not a learned simulator of physical dynamics. It stores what was found, what remains uncertain, and what should happen next.
The episode highlights a split: literature-backed and analysis-backed statements score higher, while interpretation drops sharply. The architecture may scale exploration faster than it scales scientific judgment.
Compared with AI Scientist, Robin, and AI co-scientist, Kosmos is framed as broader in long-horizon coordination and stronger in mixing literature with code-driven data analysis. The page visualizes that positioning with mock comparative scores derived from the episode discussion.