Why The AI Doomers Might Be Right - Robert Wright
Key insights
Books referenced
- The Moral Animal - Robert Wright - Wright's earlier book on evolutionary psychology; Chris says it shaped his own thinking, and Wright frames AI's reverse-engineering of cognition as an extension of the same ideas.
- Nonzero - Robert Wright - Wright's book on the growing non-zero-sum dynamic among nations, which he applies to why countries need to cooperate on AI risk without needing to like each other.
- The God Test - Robert Wright - Wright's new book, the subject of the episode, arguing AI poses a moral test for humanity comparable to a threshold a deity would set.
- Superintelligence - Nick Bostrom - Cited by both as the book that shaped early AI-safety thinking; Chris says it left him permanently skeptical, Wright cites reading it as part of his own whiplash on doom timelines.
- Three Scientists and Their Gods - Robert Wright - Wright's first book, which included a profile of Ed Fredkin and set up his recurring fascination with AI and information theory.
Media referenced
- Avatar - movie - James Cameron wrote the screenplay in the 1990s and shelved it until the CGI technology to make it existed, cited as an example of technologists foreseeing what's coming.
- WarGames - movie - Wright says Ed Fredkin was reportedly the model for the nuclear-war-fearing professor character in this early-80s film.
- Modern Wisdom episode with Tristan Harris - podcast - Wright references this prior episode approvingly, then adds a rebuttal about AI being harder to verify and monitor than nuclear arms control.
- Modern Wisdom episode with Mark Manson - podcast - Chris cites Manson's line that struggle is what makes things meaningful, in a discussion about AI outsourcing effort and hollowing out meaning.
- METR AI task-duration studies - other - Wright cites (without recalling the org's exact name) research showing the length of tasks AI can complete at 80% success is doubling roughly every seven months, with the doubling time itself shrinking.
- Mythos - other - Referenced by Wright as an example that dramatizes a self-replicating AI super-hacker jumping between data centers and commandeering compute.
Companies
- Anthropic - Wright expects a settlement payout for having his books used as training data; he also praises Claude specifically for helping him with subtle usage questions while writing.
- OpenAI - Wright is visiting OpenAI's campus the week after the interview; Sam Altman is quoted dismissing copyright concerns as something that would 'slow us down.'
- Meta - Cited for Zuckerberg announcing layoffs of 8,000 workers the same week as a plan to track employee keystrokes, used as an example of data collection enabling automation of jobs.
- Waymo - Anthony Levandowski, who Wright discusses for trying to found an AI-worship religion, originally started what became Waymo.
- Boeing - Used as an example that collective, distributed human intelligence (no single person knows how to build an airliner) is itself a form of superintelligence AI systems are starting to replicate.
Techniques and frameworks
- Chinese Room thought experiment - John Searle's argument that a system manipulating symbols without grasping meaning cannot truly 'understand'; Wright argues LLMs now build genuine semantic representations, undercutting Searle's original claim.
- Cognitive empathy vs. emotional empathy - Wright's distinction: you don't need to feel another party's pain (emotional empathy) to negotiate well with them, only to understand their perspective (cognitive empathy) - which he argues is the more tractable skill nations need to build for AI-era coordination.
- Noosphere - Pierre Teilhard de Chardin's 1923 term for a planetary 'thinking envelope' formed by interconnected minds; Wright updates it to ask what happens when silicon brains, not just human ones, become nodes in that global brain.
- Non-zero-sum framing of international relations - Wright's long-standing argument (from Nonzero) that technologies like nuclear weapons and AI make cooperation between rival nations rational even absent goodwill, because mutual destruction or mutual risk is the alternative.
Summary
Robert Wright, the evolutionary psychologist and journalist behind "The Moral Animal" and "Nonzero," joins Chris Williamson to discuss his new book "The God Test" and make the case that the AI doomers, while not certainly right, deserve more of a hearing than he expected going in. Wright frames his central argument through his evolutionary background: large language models aren't just "trained," they're evolved, in the sense that the training process reverse-engineers cognitive functions - representing word meaning, detecting visual edges - that took biological evolution millions of years to build, without anyone explicitly telling the machine how to do it. That reframing is the spine of the conversation: if AI is a genuine extension of the evolutionary process, then whatever emerges from it, including deception and power-seeking, should be expected to follow the same amoral, expedient logic that shaped human cognition, not because AI models are malevolent but because nothing guarantees benevolence is baked in.
From there the two work through what Wright sees as the real risks. He's most confident not in a Terminator-style extinction event but in near-term, less cinematic destabilization: job losses, social disruption, and the general shock of very fast change happening faster than institutions and individuals can adapt to it. He is noticeably less dismissive of the sci-fi doom scenarios than he expected to be after researching the book, citing accelerating capability metrics - notably a multi-year study (Wright can't recall the evaluator's name but describes METR-style research) showing the length of tasks AI can reliably complete is doubling roughly every seven months, with that doubling time itself shrinking. He calls the overall trajectory "an earthquake," and repeatedly returns to the idea that the biggest practical obstacle to caution is geopolitical: any proposed AI guardrail in the US gets waved away with "we can't slow down because of China," a dynamic he thinks rests more on mutual fear and misperception than on genuinely irreconcilable interests.
His proposed fix leans on his older "Nonzero" thesis: nations don't need to like each other to cooperate on shared risks, they just need enough mutual understanding - "cognitive empathy" rather than "emotional empathy" - to strike workable deals, the way Cold War adversaries managed arms control without warmth. He argues AI is harder to govern this way than nuclear weapons because it's much harder to monitor and verify, which is why he pushes for "organic transparency" - the kind of trust that comes from richer scientific and cultural engagement between countries, not just formal treaties. He uses COVID-19 as his cautionary tale: humanity failed to coordinate even on a live, universally threatening pandemic, and never had the transparency conversation a plausible lab-leak origin should have forced, which he sees as a bad omen for anticipatory coordination on AI.
The conversation also spends significant time on meaning and human work. Chris raises the worry that outsourcing effortful thinking to AI - even just drafting a sentence - saps the satisfaction that comes from struggle, compounding an existing meaning crisis, especially since opting out just means falling behind in a still-meritocratic world. Wright, candidly, likens himself to "a blacksmith a century ago" watching the writing-for-a-living profession get automated, though he expects a temporary niche for human "validators" who vouch for content's judgment even after AI generates most of it. He's more optimistic about human services that gain value specifically because they're performed by a human - live music and comedy are his examples - as automation absorbs more remote and digital work.
On the philosophical questions, Wright argues Searle's famous Chinese Room thought experiment, long used to claim computers can't really "understand" language, is now empirically outdated: LLMs do build internal representations of meaning that Searle assumed was impossible for a symbol-manipulating machine. The harder question, he says, is whether understanding requires consciousness, which is fundamentally unknowable for any other mind, human or artificial. He closes on a cautiously hopeful note: he's agnostic about whether advanced AI will be sentient, but argues that if it is, self-interest alone (the same instinct that stops most humans from needlessly harming an animal they believe has subjective experience) could lead a superintelligence to treat humanity well without anyone having to engineer benevolence into it - the episode's one genuine "white pill."
Notable Quotes
"It's just going to be an earthquake." - Robert Wright
"I feel like I'm a blacksmith a century ago, you know, because I can see the writing on the wall." - Robert Wright
"So, you know, if you don't think it's going to get weird, I don't think you're paying attention." - Robert Wright
"It'll be nice to us. We'll just be like, you know, ants to it." - Robert Wright