Radiolab
Radiolab

More or Less Human

Seven years ago chatbots - those robotic texting machines - were a mere curiosity. They were noticeably robotic and at their most malicious seemed only capable of scamming men looking for love online. Today, the chatbot landscape is wildly different. From election interference to spreading hate, cha

Featured Speakers

WNYC Studios Host

Topics Discussed

Episode Summary

Executive Summary: Radiolab revisits the Turing test through a live audience showdown between humans and a chatbot, showing the bot can now fool more than 30% of people. The episode then widens into how machines shape human behavior—through toy design, empathy, and virtual reality—arguing that AI may matter less for becoming human-like than for revealing, reshaping, and even improving human empathy and self-understanding.

Main Topics: The modern Turing test and chatbot deception (Priority: 5/5): Brian Christian explains the historical Turing test, its 30% threshold, and how recent chatbots have moved close to or past that mark in live interaction. Live audience experiment: human vs. bot (Priority: 5/5): Radiolab stages a text-based showdown with a Pandora Bots chatbot, and the bot fools roughly 45% of participants, surpassing Turing's threshold in a messy but revealing experiment. Machines shaping human behavior online (Priority: 4/5): The episode argues that chatbots and text interfaces may be changing humans more than machines are changing, as autocorrect and terse digital communication push people toward more machine-like speech. Emotional design in toys and robots (Priority: 5/5): Freedom Baird and Caleb Chung discuss Furby and Pleo as examples of devices designed to elicit empathy, guilt, care, and even moral reflection from humans. Risks of lifelike machines and abuse loops (Priority: 4/5): The Pleo example shows how adding pain, fear, and crying responses can invite cruelty and repetitive abuse, raising ethical questions about how lifelike robots should be. VR as a tool for perspective-taking and self-empathy (Priority: 5/5): Josh Rothman’s VR experience as Freud demonstrates how embodiment changes can help people see their own problems from the outside and generate self-compassion and reframing.

Key Arguments: The Turing test is no longer just theoretical; modern chatbots can fool a substantial share of users in brief text interactions. The bot's success may reflect degraded human communication norms as much as improved machine intelligence. People project emotion and humanity onto machines quickly, even when they know they are artificial. Well-designed machines can support empathy, care, and better human behavior rather than only deception. Extremely lifelike toys can provoke abuse, so designers must consider the moral effects of making machines responsive to pain. Virtual reality can create embodiment shifts that help people reframe guilt and understand themselves more compassionately.

Data Points: Turing test threshold: 30% - Alan Turing’s prediction that by 2000, 30% of judges would fail to distinguish human from machine after five minutes of text interaction. 2014 Turing test result: 30% fooled - Brian Christian notes that in 2014 a chatbot at a Turing competition managed to fool 30% of judges. Pandora bot prior performance: about 25% fooled - The bot used in the live Radiolab experiment had previously fooled roughly 25% of participants. Live audience wrong guesses: roughly 45% - In the Green Space audience experiment, about 45% of participants incorrectly identified the chatbot/human. Romantic overtures to bot: over 20% - Lauren Kunzi says more than 20% of people talking to the bot make romantic overtures. Furby upside-down response time: about 1 minute - Freedom Baird’s test found children would tolerate holding a Furby upside down for about a minute before discomfort. Barbie upside-down response time: about 5 minutes or longer - Children could hold Barbie upside down for about five minutes, limited mostly by arm fatigue. Gerbil upside-down response time: about 8 seconds - Children quickly turned the living animal right-side up, feeling it was distressed. Pleo sensors: 40 sensors - Caleb Chung describes the Pleo robot dinosaur as containing 40 sensors. Pleo review video views: about 100,000 - A video showing a Pleo being abused and destroyed was viewed about 100,000 times.

Pivotal Quotes: "we may one day be their pets." — Jad Abumrad: A joking but ominous conclusion after discussing how human communication is becoming more machine-like. "I think it made me feel a little more, um. I don't even have a word for it. Just a little more human." — Josh Rothman: After a VR session where he embodied Freud and replayed a conversation with himself about his mother. "It's a sad fact. So, this bot, over 20% of people who talk to her and millions of conversations every week, actually make romantic overtures." — Lauren Kunzi: Lauren explains a consistent pattern in human interaction with the chatbot platform.

Implications: The episode suggests AI’s biggest impact may be social and psychological: it can expose weaknesses in how we communicate, but also be used to teach empathy, reduce abuse, and help people reflect more honestly on themselves and others.

🔓 Sign Up for Unlimited Episode Search

About Radiolab

Radiolab is on a curiosity bender. We ask deep questions and use investigative journalism to get the answers. A given episode might whirl you through science, legal history, and into the home of someone halfway across the world. The show is known for innovative sound design, smashing information into music. It is hosted by Lulu Miller and Latif Nasser.

View all episodes from Radiolab