← All stories
AI & Tech

AI Labs Bet Millions of Tasks Across RL Environments Will Create AGI

Dwarkesh Patel Podcast · The next big breakthrough will be AIs learning on the job · June 26, 2026
AI Labs Bet Millions of Tasks Across RL Environments Will Create AGI
Dwarkesh Patel Podcast
Dwarkesh Patel Podcast
The next big breakthrough will be AIs learning on the job
"They think that if we train AIs to accomplish millions of verifiable tasks across thousands of diverse RL environments, then we will have basically built AGI, because this kind of training will have created a kind of problem-solving agent, the kind of thing that can make progress on open-ended tasks for weeks on end in the face of errors and mistakes and ambiguity."
Major AI research labs are making a central bet that scaling reinforcement learning across massive numbers of verifiable tasks will produce artificial general intelligence. This represents a fundamental strategic direction for companies investing billions in AI development, with the belief that current limitations in data efficiency and continual learning can be overcome through sheer compute scaling.
From this episode
Dwarkesh Patel Podcast
Dwarkesh Patel Podcast

The next big breakthrough will be AIs learning on the job

August 3, 2026 · 5 Egleze moments
Read episode summary and key points →

More moments from this episode

AI & TechComputer Use Progress Slower Because AI Cannot Grind Against Real WebsitesAI & TechAI Cannot Learn Real World Skills Like Building Businesses Without Sample EfficiencyAI & TechCurrent AI Models Are One Millionth as Sample Efficient as HumansAI & TechDario Amodei Quote Hints Short Horizon RL Does Not Generalize to Long Horizons
More stories More from Dwarkesh Patel Podcast