Building the Automated AGI Lab: Core Automation's Jerry Tworek and Rohan Anil
发布时间 来源
Episode 设置
摘要
Jerry Tworek led reasoning at OpenAI, convinced that scaling reinforcement learning was the path to AGI. Rohan Anil co-led Gemini pre-training and built the Shampoo optimizer. Now they've teamed up at Core Automation on a contrarian premise: the transformer has carried us as far as it can, and the bottleneck to smarter systems is no longer scale — it's the architecture itself. The missing capability is continual learning, models that adapt at test time, which transformers can't do. In-context learning taps out fast (Codex needs compacting after ~20 minutes) and fine-tuning invites catastrophic forgetting. Rohan argues pre-training and RL should be optimized end-to-end, and that transformers spend computation inefficiently. They lay out why the largest labs won't chase alternatives while locked in the coding-agent race, and why building the world's most automated lab starts with automating kernel generation—the one place frontier models still lose to a high-taste human.
Hosted by Sonya Huang and Pat Grady, Sequoia Capital
00:00 Introduction
01:46 Appreciating Transformers
02:44 Scaling Hits Limits
04:54 Why Architecture Matters
05:32 RL Reality Check
07:32 Test Time Learning
09:52 Economics Of Scaling
12:47 Why Start A Company
14:24 Rohan On Transformers
19:11 Computational Depth Problem
20:32 When Transformers Top Out
23:22 Beyond Reinforcement Learning
26:41 Optimization And Efficiency
34:24 Building An Automated Lab
39:45 Kernel Automation Roadmap
GPT-4正在为你翻译摘要中......
