80,000 Hours Podcast
80,000 Hours Podcast

#7 - Julia Galef on making humanity more rational, what EA does wrong, and why Twitter isn’t all bad

The scientific revolution in the 16th century was one of the biggest societal shifts in human history, driven by the discovery of new and better methods of figuring out who was right and who was wrong. Julia Galef - a well-known writer and researcher focused on improving human judgment, especially a

Featured Speakers

The 80,000 Hours team HostJulia Galef Guest

Topics Discussed

Episode Summary

Executive Summary: Rob Wiblin interviews Julia Galef about her career in rationality, podcasting, and research on expert disagreement and human judgment. Galef explains her "undebate" approach to identifying cruxes rather than trying to win arguments, reflects on her unconventional path out of academia, critiques incentives that undermine accuracy, and discusses where judgment-improvement research and norms could matter most—especially in policy, intelligence, and EA communities.

Main Topics: Galef’s current work on expert disagreement (Priority: 5/5): She is running a part-time Open Philanthropy-backed project to identify important questions where thoughtful experts disagree, compare their models, and uncover the cruxes driving disagreement for an audience of non-experts and influential decision-makers. How to have productive disagreements (Priority: 5/5): Galef argues that framing conversations around mutual understanding—rather than persuasion—reduces resistance and helps surface the real underlying assumptions (cruxes) that determine disagreement. Career path and why she left academia (Priority: 4/5): She describes leaving a PhD in economics after one year because she is a generalist and became disillusioned with methodological incentives and research quality in academia and adjacent fields like design. Podcasting, outreach, and epistemic norms (Priority: 4/5): Her Rationally Speaking podcast serves as a vehicle for promoting careful reasoning, real-time thinking, and better epistemic habits; she values long-form discussion as a way to reduce shallow or cached responses. Improving judgment in institutions (Priority: 4/5): Wiblin and Galef discuss whether forecasting and rationality methods could improve government, intelligence, and other influential organizations, but she stresses that incentives are the main barrier to accuracy. EA/rationality community strengths and risks (Priority: 4/5): Galef praises the communities’ willingness to question intuition and use explicit reasoning, but warns against overreliance on single frameworks, fast growth, and neglecting integrity or intuitive ethical concerns. Where future progress may come from (Priority: 3/5): She points to social-science reform, forecasting, and accuracy-oriented work in the intelligence community as promising avenues, while noting that the field of judgment intervention research is still small and underdeveloped.

Key Arguments: Framing disagreements as a search for shared understanding works better than trying to change minds directly, because it lowers defensiveness and reveals the real reasons people disagree. A "crux" is the underlying belief that would change someone’s mind if it changed; mapping these is more useful than arguing at the surface level. Expert disagreement matters most when questions have major world impact, so the project focuses on consequential topics rather than low-stakes or fringe issues. Podcasting can be an effective epistemic tool because long-form conversations encourage nuanced thinking, reduce misinterpretation, and make cached answers harder to rely on. Academic and design fields often reward quantity, speed, or aesthetics over rigor, making it hard for careful thinkers to thrive without strong incentives. Most institutions won’t improve accuracy unless accuracy is rewarded; in politics and bureaucracy, costs of carefulness are often immediate while benefits are delayed or uncertain. The intelligence community may be a better candidate for accuracy reforms than politics because it faces fewer pressures to appeal to broad audiences and is experimenting with forecasting. EA/rationality communities should not rely blindly on explicit reasoning frameworks; if a conclusion feels ethically or intuitively wrong, that is a cue to re-examine the model. Rapid community growth can dilute epistemic standards; EA may need to balance outreach with selectivity if it wants to remain intellectually rigorous. A major bottleneck in judgment-improvement research is that the most useful interventions are long-term and hard to study, so more work is needed beyond short lab demonstrations.

Data Points: Typical podcast listeners per episode: 35,000 - Galef says a typical Rationally Speaking episode now reaches about this many listeners. Occasional episode reach: around 100,000 - She notes some episodes occasionally get this many listeners. Modal episode reach: 5,000 - She says this is the most common listener count for some episodes. Rationally Speaking duration: since 2010 - She has hosted the podcast since 2010. Center for Applied Rationality co-founding year: 2012 - She co-founded CFAR in 2012. PhD time before dropping out: 1 year - She spent one year in an economics PhD before leaving. Consequence of social media design: 140 characters - She discusses how Twitter’s former character limit forced concise communication. Congress re-election rate: about 90% - She cites this figure while discussing incentives and policymaking accuracy. Forecasting/policy reform horizon: 5 years - She says a loose-knit community of 100 influential people could plausibly emerge in this timeframe. Global number of influential community members envisioned: 100 people - She imagines a small but influential network spanning tech, VC, government, and media. Role of judgment research: order of magnitude smaller - She estimates de-biasing/intervention research is far smaller than research documenting biases.

Pivotal Quotes: "I think it's important to not frame these conversations as we're trying to change each other's minds, or even as we're trying to converge and reach agreement." — Julia Galef: On her preferred way to structure discussions about expert disagreement. "A crux is an underlying sort of belief or assumption or premise that is feeding into your view about the topic at hand and feeding in a sort of causally important way." — Julia Galef: Her definition of the core concept behind productive disagreement. "My rule is basically: any time the benefits of accuracy are uncertain and in the future, and the costs of trying to be more accurate are paid up front... there's going to be a really strong pressure against accuracy." — Julia Galef: On why institutions often fail to improve judgment and reasoning.

Implications: Listeners interested in better decisions should focus on incentives, crux-finding, and long-form debate. For institutions, improving accuracy likely requires new norms and reward structures, not just better arguments or more information.

🔓 Sign Up for Unlimited Episode Search

About 80,000 Hours Podcast

View all episodes from 80,000 Hours Podcast