Big Technology Podcast
Big Technology Podcast

OpenAI’s New Model, Jensen’s Bold Claim, Alexa+ Is Here

Ranjan Roy from Margins is back for our weekly discussion of the latest tech news. We cover 1) OpenAI's release of GPT 4.5 2) Is GPT 4.5 a major advance or what? 3) What better EQ gets you in an AI model 4) What reasoning advances can be built on top of GPT 4.5 5) Is AI product or model more im

Featured Speakers

Alex Kantrowitz Host

Topics Discussed

Episode Summary

Executive Summary: The episode examines GPT-4.5’s muted debut and what it says about OpenAI’s strategy, arguing that progress in AI is becoming more incremental and product-oriented. The hosts contrast OpenAI’s model-centric approach with Anthropic’s smoother Claude 3.7 rollout, NVIDIA’s compute narrative, Meta’s AI app ambitions, Amazon’s promising Alexa Plus demo, and the end of Skype—framing the week as a turning point from hype to practical utility.

Main Topics: GPT-4.5 launch and OpenAI’s messaging (Priority: 5/5): The hosts analyze OpenAI’s newest model, noting its stronger conversational quality and emotional intelligence but also its lack of frontier benchmark dominance and the strange removal of documentation that said it was not a frontier model. Model progress vs. product utility (Priority: 5/5): A recurring debate centered on whether AI labs should keep emphasizing model releases or focus more on products and user experience, with the hosts arguing that model gains are getting harder to market while product value remains unclear. Reasoning models and the next frontier (Priority: 5/5): The discussion highlights how reasoning models like o1/03 mini and Claude’s thinking mode are proving highly useful, and how GPT-4.5 may serve as a foundation for a more capable GPT-5 that combines pretraining and reasoning. Anthropic Claude 3.7 Sonnet and efficient model launches (Priority: 4/5): Anthropic’s new hybrid reasoning model is presented as a cleaner, more product-minded release, with emphasis on its thinking toggle, coding utility, and reportedly lower training cost. NVIDIA, compute demand, and inference economics (Priority: 4/5): Jensen Huang’s claim that reasoning will require far more compute is discussed alongside NVIDIA’s blowout earnings, while the hosts question whether the AI industry’s trajectory will actually justify that level of spending long term. Amazon Alexa Plus and the future of assistants (Priority: 4/5): Amazon’s revamped Alexa is praised as one of the more convincing consumer AI demos, with real-world actions and cross-Amazon awareness suggesting a path to a true universal assistant. Meta’s standalone AI app and the end of Skype (Priority: 3/5): The episode closes with Meta’s reported plan to launch a Meta AI app and Microsoft’s decision to retire Skype, which the hosts treat as both a sign of consumer AI competition and a bittersweet ending for a landmark internet product.

Key Arguments: GPT-4.5 is more pleasant, concise, and human-like to talk to, but it does not clearly surpass OpenAI’s reasoning models on hard benchmarks. OpenAI appears to be positioning GPT-4.5 as evidence that scaling still works, even if the release is less exciting than earlier model jumps. The hosts argue that the real breakthrough may be making AI feel less like AI—more conversational, less verbose, and more usable in everyday interactions. Reasoning models are viewed as a genuine step forward because they materially improve usefulness on complex tasks, even if they are not always necessary for simple queries. OpenAI’s future likely depends on combining a stronger base model with reasoning, making GPT-5 the key proving ground. Anthropic’s Claude 3.7 Sonnet is presented as a more disciplined launch: smaller branding, clear utility, and a hybrid mode that improves usability. NVIDIA’s thesis is that inference and reasoning will drive massive compute demand, but the hosts question whether the economics really support a 100x compute jump. Amazon has an advantage in building a universal assistant because it can integrate across services and devices without the friction of a mobile OS default. Meta’s AI app may gain distribution through its existing ecosystem, but the hosts doubt it will become a category-defining consumer product. Skype’s shutdown symbolizes the end of an era for a product that helped popularize internet calling and now gives way to newer communication platforms.

Data Points: GPT-4.5 simple QA accuracy: 62.5% - OpenAI benchmark mentioned as higher than GPT-4.0 in basic question answering. Closest model simple QA accuracy: 47% - Compared against OpenAI o1 in the discussion of simple QA performance. GPT-4.5 hallucination rate: 37.1% - Lower hallucination rate was cited as part of the model’s improvements. GPT-4.0 hallucination rate: 61.8% - Used for comparison when discussing hallucination improvements. Everyday queries preference for GPT-4.5: 57% - Users preferred GPT-4.5 over GPT-4.0 in everyday query evaluations. Professional queries preference for GPT-4.5: 63.2% - Preference score mentioned during discussion of model evaluation results. Creative intelligence preference for GPT-4.5: 56.8% - Preference score for creative tasks versus GPT-4.0. GPQA science benchmark: 71.4% - GPT-4.5 score cited versus stronger performance from o3 mini. o3 mini GPQA science benchmark: 79.7% - Used to show reasoning models outperform GPT-4.5 on science tasks. AIME 2024 math benchmark: 36.7% - GPT-4.5 score on math compared with o3 mini. o3 mini AIME 2024 math benchmark: 87.3% - Shown as substantially outperforming GPT-4.5 in math. ChatGPT user growth: 100 million to 300 million - Altman cited rapid growth as a reason for GPU shortages. NVIDIA quarterly revenue: $39.33 billion - Reported quarterly revenue in the earnings discussion. NVIDIA revenue growth: 78% year over year - Compared to the prior year’s quarter. NVIDIA next-quarter projection: $43 billion - Forward guidance mentioned during earnings discussion. Blackwell chips delivered: $11 billion - NVIDIA’s latest chip shipments were highlighted as a major demand signal. GPT-4.5 training cost: a few tens of millions of dollars - Referenced for Anthropic’s Claude 3.7 Sonnet training cost. OpenAI documentation claim: 10x computational efficiency improvement - A removed document said GPT-4.5 improved GPT-4’s computational efficiency by more than 10x. OpenAI removed claim: not a frontier model - Mentioned in deleted documentation that was later removed from the release note. Skype users in 2023: 36 million - Microsoft’s most recent user count before shutdown. Skype peak users: 300 million - Historical peak referenced to show its decline. Skype acquisition price: $8.5 billion - Microsoft bought Skype in 2011. Skype shutdown date: May 5 - Microsoft said it will retire Skype on this date. Alexa Plus price: $19.99/month - Standalone subscription price cited during the Amazon discussion. Prime price: $14.99/month - Used to illustrate that Prime includes Alexa Plus at better value.

Pivotal Quotes: "GPT 4.5 is the least interesting model release from OpenAI to date." — Ranjan Roy: He argues that recent model launches feel incremental rather than revolutionary. "the first model that feels like talking to a thoughtful person to me" — Sam Altman: Altman’s characterization of GPT-4.5 in his launch tweet. "Make AI less AI." — Alex Kantrowitz / Ranjan Roy: A proposed marketing frame for GPT-4.5 as more human and conversational.

Implications: AI progress is shifting from flashy benchmark wins to product usefulness, conversational quality, and integrated workflows. The winners may be the companies that make models feel natural and useful daily, not just the ones with the biggest models.

🔓 Sign Up for Unlimited Episode Search

About Big Technology Podcast

The Big Technology Podcast takes you behind the scenes in the tech world featuring interviews with plugged-in insiders and outside agitators. Alex Kantrowitz, a Silicon Valley journalist who's interviewed the world's top tech CEOs — from Mark Zuckerberg to Larry Ellison — is the host.

View all episodes from Big Technology Podcast