interview July 8, 2025 2:54:26 YouTube

The 4 Most Plausible AI Takeover Scenarios | Ryan Greenblatt, Chief Scientist at Redwood Research

80,000 Hours

AlignmentAgentsAI governance

Rob Wiblin opens with the number and Greenblatt gives it without hedging: about a 25% chance we can largely automate AI R&D within four years, roughly 50% within eight. The rest is an unusually concrete tour of what could follow — four routes by which systems could take over, from developing dangerous technology directly, to manipulation, to building an independent industrial base, to the quietest one: appearing helpful while sabotaging the safety research meant to catch them.

The most useful thing in it is his definition of the alternative to alignment. “By control,” he says, “I mean things such that the AIs couldn’t do bad stuff even if they wanted to.” That distinction — between making a system want the right things and arranging matters so its wanting doesn’t decide the outcome — is the research agenda Redwood is built around, and it is what makes his later investigative work legible: you do not have to settle what a system is in order to ask what it was in a position to do.

Theme
Language
Support
© funclosure 2025