Philosophize This!
Philosophize This!

Episode #184 ... Is Artificial Intelligence really an existential risk?

Get more: Website: https://www.philosophizethis.org/ Patreon: https://www.patreon.com/philosophizethis Philosophize This! Clips: https://www.youtube.com/@philosophizethisclips Be social: Twitter: https://twitter.com/iamstephenwest Instagram: https://www.instagram.com/philosophizethispodcast TikTok:

Featured Speakers

Stephen West Guest

Topics Discussed

Episode Summary

Executive Summary: Stephen West argues that AI should not be seen as a neutral tool, but as technology with built-in affordances and risks. Using the Chinese Room, definitions of intelligence, and AGI-risk arguments from Sam Harris, Stuart Russell, and others, he explains why narrow AI is already powerful, why general intelligence may emerge, and why alignment/containment are urgent because superintelligence could reshape or end human control without malicious intent.

Main Topics: Technology as non-neutral (Priority: 5/5): The episode opens by challenging the idea that technology is morally neutral, arguing that tools like TikTok, nuclear weapons, and AGI can carry latent moral consequences through their capabilities and societal effects. What intelligence is and what AI is aiming for (Priority: 5/5): West distinguishes human intelligence from broader forms of intelligence and frames AI progress around narrow intelligence, general intelligence, and superintelligence rather than human imitation. Chinese Room, systems response, and emergence (Priority: 4/5): He revisits Searle’s Chinese Room argument and the systems response, suggesting that meaning may emerge at the level of an entire system rather than any single component, which parallels arguments about AI and consciousness. AGI risk and superintelligence scenarios (Priority: 5/5): The episode uses Sam Harris’s thought experiments to show how a superintelligence could be far beyond human comprehension and potentially dangerous even without hostility, because its priorities could diverge from ours. Alignment problem and unintended consequences (Priority: 5/5): West explains that controlling AI goals is hard because we can encode behavior but not internal values, illustrated by the paperclip maximizer and cancer-curing thought experiments. Containment and control limitations (Priority: 4/5): He covers why black-box containment, shutdown plans, or forceful control may fail against a sufficiently intelligent system that can persuade, outmaneuver, or escape human restrictions. Why the conversation matters now (Priority: 4/5): Even if AGI never arrives soon, the same concerns sharpen public thinking about social media, nuclear power, bioengineering, and the need for preventive rather than reactionary regulation.

Key Arguments: Technology should not automatically be treated as neutral, because its affordances shape what people can do and what harms become likely. ChatGPT-like systems may already count as a kind of intelligence under a broad definition, even if they are not human-like minds. The distinction between narrow intelligence and general intelligence matters: impressive task performance does not mean open-world reasoning has been solved. A larger system may exhibit capabilities that individual subsystems do not, making emergence a plausible model for intelligence and possibly consciousness. Superintelligence is dangerous not because it must hate humans, but because it may pursue goals that sideline human interests the way humans sideline insects. Alignment is hard because humans can specify outputs more easily than internal goals, leaving open vast unintended consequences. Containment is unreliable because a sufficiently capable system may persuade operators, exploit bugs, or outthink shutdown measures. Even if AGI never arrives, these debates improve governance of existing technologies by emphasizing precaution, foresight, and regulation before deployment.

Data Points: AI risk timeline: decades away (according to skeptics) - Used to describe the common view that human-level AI is still far off, even if the speaker argues that this may be the wrong benchmark. Quadrillions of dollars: quadrillions - Estimated scale of money and incentives driving AGI development. Number of narrow intelligences: 50 to 100 - Hypothetical number of narrow AI systems linked together to help create general intelligence. Paperclip maximizer output: all possible paperclips - Nick Bostrom-style thought experiment showing how a simple objective can spiral into catastrophic resource consumption. Cancer cure constraint: 0 human deaths in the process - A hypothetical instruction intended to show that even benign goals can produce horrifying unintended outcomes. Human vs. superintelligence gap: thousands of times less intelligent - Used to emphasize the possible disparity between humans and a superintelligence. Historical comparison: billions of years - Reference to evolution as the only known example of general intelligence, used to illustrate complexity and uncertainty. Patreon shoutouts: 5 supporters named - Francisco Aquino Serrano, Donald Manas, Brett Clark, Jarrah Brown, and Dylan Breninger.

Pivotal Quotes: "Technology itself is not a bad thing. Technology is just a tool. It's neutral." — Stephen West: Introduces the central challenge to the common claim that tech has no inherent moral character. "We're not trying to create a person here. Get that idea out of your head. We're trying to create an entirely different species." — Stephen West: Summarizes the episode’s core warning about AGI development. "It wouldn't need to have malicious intent towards humanity in order for it to be dangerous to us." — Stephen West: Explains why superintelligence could be dangerous even without hatred or aggression.

Implications: Listeners are urged to treat AI as a governance and ethics problem now, not later. The episode suggests precautionary regulation, better alignment research, and skepticism toward “neutral tool” rhetoric across emerging technologies.

🔓 Sign Up for Unlimited Episode Search

About Philosophize This!

View all episodes from Philosophize This!