**Erik Torenberg** (0:00)
Hello, and welcome back to The Cognitive Revolution. Today's episode is a cross post from the 80,000 Hours Podcast, hosted by Rob Wiblin and featuring a conversation with Ajeya Cotra, who previously led technical AI safety grant making at Open Philanthropy, now Coefficient Giving, and is now working on risk assessment at METER. For years, AI insiders have recognized Ajeya as one of the most rigorous thinkers about the AI future, and she recently validated that judgment by coming in number 3 out of more than 400 participants in the AI Digest 2025 AI Forecasting Survey. For comparison, I was proud to land in the top 5% at number 23 In this conversation, Ajeya takes Rob through her expectations for the next few years, as AI crosses critical thresholds, recursive self-improvement intensifies, and we enter what she describes as crunch time, a potentially short window in which AI is powerful enough to dramatically accelerate AI R&D, but not yet totally beyond human control. As a preview, I'll warn you that even the accelerationists may suffer some future shock from this conversation, because Ajeya thinks it's actually quite plausible that we find no insurmountable bottlenecks to widespread and compounding automation, and that if so, the world of 2050 could look as different from our perspective today as our world would look to hunter-gatherers of 10,000 years ago. So, what's the plan to make sure such a mind-boggling transformation goes well for humans? Ajeya advocates for transparency measures and early warning systems designed to make sure that superintelligence doesn't happen in secret. But aside from that, she reports that all frontier developers are gradually converging on a strategy of using each generation of AIs to attempt to align, understand, and control their own successors.
As regular listeners know, I signed the Future of Life Institute's October 2025 petition calling for a ban on superintelligence, not because I think this approach is forever destined to fail, but simply because I worry that we don't yet understand AIs well enough to bet on a good outcome from such a recursive, self-improvement-powered intelligence explosion. And yet, at the same time, I do agree with Ajeya's advice. Almost regardless of the kind of work you're doing, you should be adopting AI as aggressively as possible. Both to maintain an accurate understanding of the situation, and increasingly because you won't be able to keep up without it. It is a mad, mad world that will soon be living in. But I would go as far as to say that even pause-AI campaigners ought to be using AI extensively. If all that weren't enough for you to process, you should also know that the situation has recently accelerated yet again. On March 5th, just about two weeks after this episode was originally published, Ajeya posted an article on her substack. Planned Obsolescence, called I Underestimated AI Capabilities Again, in which she reports that the predictions that she made in January 2026, which were the backdrop for this conversation, were already starting to be met in just the first couple months of this year. And more recently, we've of course learned of Anthropic's new Mythos model, which despite the fact that Anthropic has never emphasized benchmark scores as much as other model developers, shows major gains on many benchmarks, and has reportedly found zero-day exploits, in every major operating system and every major web browser, among many other major software projects. The bottom line is that crunch time is arguably here now. So if you've been watching and waiting for AI to get serious before deciding what to do about it, I would suggest getting off the sidelines sooner rather than later. If you need help figuring out what to do, you might consider applying for free one-on-one career advising from 80,000 Hours. As always, I want to thank Rob and the 80,000 Hours team for allowing me to cross post this episode. They have been delivering incredible alpha for years. And the nearer the singularity becomes, the more prescient they look. With that, I hope you enjoy this essential conversation about AI timelines and crunch time strategy with Ajeya Cotra and host Rob Wiblin from the 80,000 Hours Podcast.
**Ajeya Cotra** (4:17)
If you look at public communications from at least OpenAI, Anthropic, and Google DeepMind, in all of their stated safety plans, you see this element of as AIs get better and better, they're going to incorporate the AIs themselves into their safety plans more and more. How to create a setup where we use control techniques and alignment techniques and interpretability to the point where we feel good about relying on their outputs is like a crucial step to figure out. Because it either like bottlenecks our progress, because we're checking on everything all the time and slowing things down, or it doesn't bottleneck our progress, but we like hand the AIs the power to take over.
167 more minutes of transcript below
Try it now — copy, paste, done:
curl -H "x-api-key: pt_demo" \
https://spoken.md/transcripts/1000651996090
Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.
From $0.10 per transcript. No subscription. Credits never expire.
Using your own key:
curl -H "x-api-key: YOUR_KEY" \
https://spoken.md/transcripts/1000760851166