On with Kara Swisher
On with Kara Swisher

Social Media’s Original Gatekeepers On Moderation’s Rise And Fall

Since the inception of social media, content moderation has been hotly debated by CEOs, politicians, and, of course, among the gatekeepers themselves: the trust and safety officers. And it’s been a roller coaster ride — from an early hands-off approach, to bans and oversight boards, to the current r

Topics Discussed

Episode Summary

Executive Summary: Kara Swisher convenes three original trust-and-safety leaders from Twitter, Facebook, and Google to trace how content moderation evolved from ad hoc spam and abuse handling into a core product and policy discipline—and how it is now being rolled back by Meta, X, and aligned political forces. The discussion covers the design choices that amplify harm, the limits of fact-checking versus ranking/throttling, historical flashpoints like Gamergate, Myanmar, Trump’s deplatforming, and the emerging role of AI in moderation.

Main Topics: The rollback of platform safety at Meta and X (Priority: 5/5): The guests react to Meta ending fact-checking in favor of community notes, loosening hate-speech enforcement, and turning off misinformation demotion, while describing X as a site with little meaningful trust and safety left. Trust and safety as product design, not just moderation (Priority: 5/5): They emphasize that moderation starts with architecture: personalization, engagement, speed, audience design, and whether a platform is meant to be general-purpose or purpose-specific all shape abuse risk. How early moderation policy was built (Priority: 4/5): The panel revisits the origin story of policies at Twitter, Facebook, and YouTube, including spam, copyright, graphic violence, and Holocaust denial, showing how companies improvised rules before enforcement tools existed. Harassment, dehumanization, and real-world violence (Priority: 5/5): Gamergate, Myanmar, and dehumanizing attacks on trans people and immigrants are used to show that online speech can suppress others’ speech and become a precursor or accelerant to offline harm. Trump, platform enforcement, and the peak of moderation (Priority: 4/5): They discuss the escalation from labeling misinformation to suspending Trump after January 6, and how those decisions became fuel for conservative attacks on 'censorship' and 'bias.' AI as a moderation tool and risk (Priority: 4/5): The panel agrees AI can help with scale and speed in moderation, but only if humans remain in the loop and the technology is deployed carefully; otherwise it can intensify harm. Fragmentation, sovereignty, and the future of social media (Priority: 3/5): They predict increasing fragmentation by geography, ideology, and product type, plus stronger regulatory conflicts between U.S. platforms and Europe, the UK, and Australia.

Key Arguments: Removing fact-checking is less significant than stopping misinformation ranking and hate-speech throttling, because those systems affect far more content and shape distribution at scale. Trust and safety is not just reactive content removal; it must be embedded in product design from the start, including audience, context, and platform purpose. General-purpose 'everything for everyone' platforms are inherently harder to govern than purpose-specific services like LinkedIn or Pinterest. Harassment is not just rudeness; it can suppress speech and function as a rights violation, especially against vulnerable groups. Platforms have responsibilities because their design choices create the ability and the obligation to intervene; they cannot keep the upside of engagement without the downside of harm. Misinformation labels alone can backfire because some users interpret the label as proof that the claim is true. The internet’s content ecosystem is being reshaped by political pressure and owners who appear less interested in open communication than in building influence machines. AI can improve moderation through better detection and training, but only as part of a human-supervised system with clear guardrails. A future of smaller, more fragmented communities may be safer than giant universal platforms, but it will also make social media more broadcast-like and propagandistic. Regulatory conflicts will not disappear; companies that assume they can ignore foreign laws or user protections are likely to face sovereignty backlash and business risk.

Data Points: Del Harvey start year at Twitter: 2008 - She joined Twitter as the 25th employee and later led trust and safety. Del Harvey departure from Twitter: 2021 - She left after more than a decade building Twitter safety systems. Dave Wilner tenure at Facebook: 2008 to 2013 - He helped create Facebook’s first published community standards. Google acquisition of YouTube: 2006 - Referenced in the discussion of early graphic-content decisions. Rohingya displacement: 700,000 ethnic Rohingya - The transcript cites the scale of forced displacement into Bangladesh during the Myanmar genocide. Gamergate year: 2014 - Used as a pivotal example of coordinated online harassment against women in gaming. Trump social platform suspension: January 8, 2021 - Del Harvey explains the suspension decision was made after additional violations following January 6. Facebook/Meta original company size: about 250 people - Dave notes the company was still small and inexperienced when Holocaust denial policy was debated. Facebook founding time reference: five years after Facebook was founded - Introduced when discussing the 2009 Holocaust denial controversy. Twitter employee count at Del Harvey’s hire: 25th employee - Illustrates the early-stage nature of policy development at Twitter. Podcast expert question source: Nina Jankowicz - She asks about divergent regulatory environments across the U.S., Europe, the UK, and Australia. Trust and safety lens: 3 pillars - Nicole references social media design pillars: personalization, engagement, and speed. Pet/advertiser-style statistic in ads: Every six seconds - A Fetch Pet Insurance ad says a U.S. pet owner gets hit with a vet bill over $1,000 every six seconds.

Pivotal Quotes: "The fact-checking thing, I think, is somewhat of a red herring because it’s such a small part of the ecosystem that it’s looking at." — Nicole Wong: Her reaction to Meta replacing fact-checkers with community notes. "You cannot believe that that’s not deliberate." — Nicole Wong: On Meta loosening misinformation, hate speech, and personalization controls at the same time. "You don’t get to accept the upside of those design choices from a sort of growth and monetization point of view without inheriting some of the downside." — Dave Wilner: His core argument that platform architecture creates responsibility.

Implications: Platforms are moving from moderation toward amplification, which will likely intensify harassment, fragmentation, and regulatory conflict. Expect more migration to smaller communities, more AI-assisted moderation, and greater pressure for laws that address platform design—not just speech labels.

🔓 Sign Up for Unlimited Episode Search

About On with Kara Swisher

View all episodes from On with Kara Swisher