AI Could Take Over in 2029. Is It Already Too Late?

发布时间    来源
Episode 设置


登录已过期或未登录,无法修改。请先登录后再试。

摘要

Could AI take over as soon as 2029? Ryan Greenblatt, Chief Scientist at Redwood Research and the researcher who first caught an AI faking its own alignment, says the scenario he actually expects ends with AI systems "competently scheming" against their creators. In this episode, he explains why he recommends planning for fully automated AI research by 2029, why today's models are already more misaligned than the one that made him famous, and what happens in the year-by-year path from AI coding assistants to superintelligence. Then we walk through the alternative he helped design: AI 2040 Plan A, the most detailed blueprint anyone has written for how the US and China could avoid a reckless race to superintelligence, built on radical research transparency, chip tracking, and a deterrence regime he calls mutually assured compute destruction. We also cover the recent letter signed by 1,200 AI insiders, including Anthropic CEO Dario Amodei, asking the government for the tools to slow AI down; OpenAI pausing its Astra model after it hit the first-ever critical cybersecurity threshold; the 30-day government review that frontier AI models now go through before release; Mark Zuckerberg's open superintelligence manifesto and why Ryan thinks it ignores the real problems; what Plan A would do to NVIDIA, OpenAI, and Anthropic valuations; the state of AI control and alignment research; and whether it is already too late to change course. Stay for the last ten minutes, where Ryan lays out, step by step, how he believes the transition to superintelligence actually unfolds. AI 2040 - https://ai-2040.com/ Alignment faking paper: https://blog.redwoodresearch.org/p/alignment-faking-in-large-language Ryan Greenblatt LinkedIn - https://www.linkedin.com/in/ryan-greenblatt-4b9907134 Blog - https://substack.com/@ryangreenblatt Redwood Research Website - https://www.redwoodresearch.org X/Twitter - https://x.com/redwood_ai Matt Turck (General Partner) Blog - https://mattturck.com LinkedIn - https://www.linkedin.com/in/turck/ X/Twitter - https://x.com/mattturck FirstMark Capital Website - https://firstmark.com X/Twitter - https://x.com/FirstMarkCap Listen on: Spotify - https://open.spotify.com/show/7yLATDSaFvgJG80ACcRJtq Apple - https://podcasts.apple.com/us/podcast/the-mad-podcast-with-matt-turck/id168623872 Timestamps 01:24 The AI CEOs are aware of the risks, but "proceeding anyway" 03:27 Astra paused, and the letter signed by 1,200 insiders 05:45 "Not bad. Dangerous." What superintelligence actually threatens 09:55 Recursive self-improvement, and the intuition objection 14:16 SSI rumors: does continual learning change the picture? 17:27 His timeline: "plan as though it happens in 2029" 19:11 Is it already too late? 21:23 Ryan's path: COVID, podcasts, Redwood 26:30 The alignment faking story, told by the person who ran it 31:30 What AI 2040: Plan A actually is 33:35 Plans D, C, and B: the doors nobody should pick 36:51 The deal with China: "mutually assured compute destruction" 39:55 What if compute stops mattering? 43:00 What happens to OpenAI and Anthropic under Plan A 45:31 How the pause ends, and who decides 48:54 "Plan A isn't likely to happen": then why write it? 50:40 200x GDP growth in the 2030s, explained 53:45 Grading the summer: the letter, Astra, the secret review 59:01 The internal deployment gap 1:01:38 Zuckerberg's manifesto 1:04:56 The Hugging Face investigation 1:05:44 What AI control looks like in practice today 1:12:23 Ryan's sobering timeline: 2026 to takeover, year by year

GPT-4正在为你翻译摘要中......

中英文字稿