Nov 15, 2024 · 56m · big-technology

Is Generative AI Plateauing?, Booming Bluesky, Apple’s Smart Glasses Play

Ranjan Roy · 26m spoken Alex Kantrowitz · 25m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

Alex Kantrowitz and Ranjan Roy analyze reports of large language model performance plateaus, the rise of reasoning-focused AI architectures, and practical enterprise market disruptions. They also evaluate Bluesky's post-election growth against Twitter's network dominance and explore Apple's exploratory moves into smart glasses hardware.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Alex holds 50.6% of the talking time here. How this is scored →

Alex as informed peer 5.6 Guest teaching 4.2 Guest disagreement 2.2 Alex pushing back 2.6
05100:0015:0030:0045:000:49–12:49 · Alex as informed peer 7/10 Mike Tyson vs. Jake Paul Fight Banter Alex aggressively challenges Ranjan's optimism regarding the LLM scaling plateau, citing Ilya Sutskever's Reuters comments, Marc Andreessen, and OpenAI's massive $6B funding round. Ranjan firmly resists Alex's framing, arguing that slowing foundational model development is beneficial for practical productization.12:50–17:12 · Alex as informed peer 7/10 The Shift Toward Reasoning Models and Inference Compute Alex explains the transition to test-time reasoning compute and cites a TEDAI poker statistic demonstrating efficiency gains over pre-training scale. Ranjan expands on how targeted synthetic data enables startups like Writer to build focused models at fractionally lower costs.17:13–22:33 · Alex as informed peer 6/10 Deconstructing AGI Definitions and Industry Expectations The pair debate definitions of AGI, with Alex citing Yann LeCun's human-level generalist benchmark while Ranjan questions whether AGI is simply a nebulous marketing and contractual term that distracts from narrower utility.22:33–27:22 · Alex as informed peer 5/10 Gemini's Benchmark Surge and Real-World AI Workflows Alex highlights Google's experimental Gemini model topping the Chatbot Arena leaderboard, while both agree that superior benchmarks do not immediately overcome interface and workflow preferences.27:23–37:06 · Alex as informed peer 6/10 Practical AI Impact: Writer, Chegg, and Agency Billing Models Alex discusses Chegg's 99% valuation collapse and shifts in ad agency billing models due to AI efficiency. Ranjan explains the broader transition across SaaS, legal, and healthcare sectors toward outcome-based pricing rather than seat or hourly billing.37:07–42:10 · Alex as informed peer 5/10 Spotify's Recommendation Pitfalls and AI-Generated Music Ranjan describes how Spotify's recommendation algorithms break under household edge cases like kids' music, then explains his positive tests with Spotify prompt-based playlists and Suno-generated folk tracks.42:10–45:00 · Alex as informed peer 3/10 Podcast Format Updates, Listener Feedback, and Starshield Alex conducts housekeeping, addresses listener feedback, and shares a listener correction regarding the US Department of Defense's Starshield satellite initiative with SpaceX.45:01–50:06 · Alex as informed peer 6/10 Bluesky's Growth Momentum vs. Threads and Twitter Network Effects Ranjan argues Bluesky has achieved genuine mainstream momentum post-election, while Alex pushes back citing Max Read's analysis that it remains a specialized echo chamber unable to dislodge Twitter's core network effects.50:07–56:22 · Alex as informed peer 5/10 Apple Smart Glasses Study and Mark Zuckerberg's Z-Pain Track Ranjan evaluates Apple's exploratory Project Atlas smart glasses against his hands-on experience with Snap Spectacles, and the hosts banter over Mark Zuckerberg's acoustic Z-Pain music release.0:49–12:49 · Guest teaching 3/10 Mike Tyson vs. Jake Paul Fight Banter Alex aggressively challenges Ranjan's optimism regarding the LLM scaling plateau, citing Ilya Sutskever's Reuters comments, Marc Andreessen, and OpenAI's massive $6B funding round. Ranjan firmly resists Alex's framing, arguing that slowing foundational model development is beneficial for practical productization.12:50–17:12 · Guest teaching 5/10 The Shift Toward Reasoning Models and Inference Compute Alex explains the transition to test-time reasoning compute and cites a TEDAI poker statistic demonstrating efficiency gains over pre-training scale. Ranjan expands on how targeted synthetic data enables startups like Writer to build focused models at fractionally lower costs.17:13–22:33 · Guest teaching 4/10 Deconstructing AGI Definitions and Industry Expectations The pair debate definitions of AGI, with Alex citing Yann LeCun's human-level generalist benchmark while Ranjan questions whether AGI is simply a nebulous marketing and contractual term that distracts from narrower utility.22:33–27:22 · Guest teaching 4/10 Gemini's Benchmark Surge and Real-World AI Workflows Alex highlights Google's experimental Gemini model topping the Chatbot Arena leaderboard, while both agree that superior benchmarks do not immediately overcome interface and workflow preferences.27:23–37:06 · Guest teaching 6/10 Practical AI Impact: Writer, Chegg, and Agency Billing Models Alex discusses Chegg's 99% valuation collapse and shifts in ad agency billing models due to AI efficiency. Ranjan explains the broader transition across SaaS, legal, and healthcare sectors toward outcome-based pricing rather than seat or hourly billing.37:07–42:10 · Guest teaching 5/10 Spotify's Recommendation Pitfalls and AI-Generated Music Ranjan describes how Spotify's recommendation algorithms break under household edge cases like kids' music, then explains his positive tests with Spotify prompt-based playlists and Suno-generated folk tracks.42:10–45:00 · Guest teaching 2/10 Podcast Format Updates, Listener Feedback, and Starshield Alex conducts housekeeping, addresses listener feedback, and shares a listener correction regarding the US Department of Defense's Starshield satellite initiative with SpaceX.45:01–50:06 · Guest teaching 4/10 Bluesky's Growth Momentum vs. Threads and Twitter Network Effects Ranjan argues Bluesky has achieved genuine mainstream momentum post-election, while Alex pushes back citing Max Read's analysis that it remains a specialized echo chamber unable to dislodge Twitter's core network effects.50:07–56:22 · Guest teaching 5/10 Apple Smart Glasses Study and Mark Zuckerberg's Z-Pain Track Ranjan evaluates Apple's exploratory Project Atlas smart glasses against his hands-on experience with Snap Spectacles, and the hosts banter over Mark Zuckerberg's acoustic Z-Pain music release.0:49–12:49 · Guest disagreement 5/10 Mike Tyson vs. Jake Paul Fight Banter Alex aggressively challenges Ranjan's optimism regarding the LLM scaling plateau, citing Ilya Sutskever's Reuters comments, Marc Andreessen, and OpenAI's massive $6B funding round. Ranjan firmly resists Alex's framing, arguing that slowing foundational model development is beneficial for practical productization.12:50–17:12 · Guest disagreement 2/10 The Shift Toward Reasoning Models and Inference Compute Alex explains the transition to test-time reasoning compute and cites a TEDAI poker statistic demonstrating efficiency gains over pre-training scale. Ranjan expands on how targeted synthetic data enables startups like Writer to build focused models at fractionally lower costs.17:13–22:33 · Guest disagreement 3/10 Deconstructing AGI Definitions and Industry Expectations The pair debate definitions of AGI, with Alex citing Yann LeCun's human-level generalist benchmark while Ranjan questions whether AGI is simply a nebulous marketing and contractual term that distracts from narrower utility.22:33–27:22 · Guest disagreement 1/10 Gemini's Benchmark Surge and Real-World AI Workflows Alex highlights Google's experimental Gemini model topping the Chatbot Arena leaderboard, while both agree that superior benchmarks do not immediately overcome interface and workflow preferences.27:23–37:06 · Guest disagreement 2/10 Practical AI Impact: Writer, Chegg, and Agency Billing Models Alex discusses Chegg's 99% valuation collapse and shifts in ad agency billing models due to AI efficiency. Ranjan explains the broader transition across SaaS, legal, and healthcare sectors toward outcome-based pricing rather than seat or hourly billing.37:07–42:10 · Guest disagreement 1/10 Spotify's Recommendation Pitfalls and AI-Generated Music Ranjan describes how Spotify's recommendation algorithms break under household edge cases like kids' music, then explains his positive tests with Spotify prompt-based playlists and Suno-generated folk tracks.42:10–45:00 · Guest disagreement 0/10 Podcast Format Updates, Listener Feedback, and Starshield Alex conducts housekeeping, addresses listener feedback, and shares a listener correction regarding the US Department of Defense's Starshield satellite initiative with SpaceX.45:01–50:06 · Guest disagreement 4/10 Bluesky's Growth Momentum vs. Threads and Twitter Network Effects Ranjan argues Bluesky has achieved genuine mainstream momentum post-election, while Alex pushes back citing Max Read's analysis that it remains a specialized echo chamber unable to dislodge Twitter's core network effects.50:07–56:22 · Guest disagreement 2/10 Apple Smart Glasses Study and Mark Zuckerberg's Z-Pain Track Ranjan evaluates Apple's exploratory Project Atlas smart glasses against his hands-on experience with Snap Spectacles, and the hosts banter over Mark Zuckerberg's acoustic Z-Pain music release.0:49–12:49 · Alex pushing back 7/10 Mike Tyson vs. Jake Paul Fight Banter Alex aggressively challenges Ranjan's optimism regarding the LLM scaling plateau, citing Ilya Sutskever's Reuters comments, Marc Andreessen, and OpenAI's massive $6B funding round. Ranjan firmly resists Alex's framing, arguing that slowing foundational model development is beneficial for practical productization.12:50–17:12 · Alex pushing back 2/10 The Shift Toward Reasoning Models and Inference Compute Alex explains the transition to test-time reasoning compute and cites a TEDAI poker statistic demonstrating efficiency gains over pre-training scale. Ranjan expands on how targeted synthetic data enables startups like Writer to build focused models at fractionally lower costs.17:13–22:33 · Alex pushing back 3/10 Deconstructing AGI Definitions and Industry Expectations The pair debate definitions of AGI, with Alex citing Yann LeCun's human-level generalist benchmark while Ranjan questions whether AGI is simply a nebulous marketing and contractual term that distracts from narrower utility.22:33–27:22 · Alex pushing back 1/10 Gemini's Benchmark Surge and Real-World AI Workflows Alex highlights Google's experimental Gemini model topping the Chatbot Arena leaderboard, while both agree that superior benchmarks do not immediately overcome interface and workflow preferences.27:23–37:06 · Alex pushing back 2/10 Practical AI Impact: Writer, Chegg, and Agency Billing Models Alex discusses Chegg's 99% valuation collapse and shifts in ad agency billing models due to AI efficiency. Ranjan explains the broader transition across SaaS, legal, and healthcare sectors toward outcome-based pricing rather than seat or hourly billing.37:07–42:10 · Alex pushing back 1/10 Spotify's Recommendation Pitfalls and AI-Generated Music Ranjan describes how Spotify's recommendation algorithms break under household edge cases like kids' music, then explains his positive tests with Spotify prompt-based playlists and Suno-generated folk tracks.42:10–45:00 · Alex pushing back 0/10 Podcast Format Updates, Listener Feedback, and Starshield Alex conducts housekeeping, addresses listener feedback, and shares a listener correction regarding the US Department of Defense's Starshield satellite initiative with SpaceX.45:01–50:06 · Alex pushing back 6/10 Bluesky's Growth Momentum vs. Threads and Twitter Network Effects Ranjan argues Bluesky has achieved genuine mainstream momentum post-election, while Alex pushes back citing Max Read's analysis that it remains a specialized echo chamber unable to dislodge Twitter's core network effects.50:07–56:22 · Alex pushing back 1/10 Apple Smart Glasses Study and Mark Zuckerberg's Z-Pain Track Ranjan evaluates Apple's exploratory Project Atlas smart glasses against his hands-on experience with Snap Spectacles, and the hosts banter over Mark Zuckerberg's acoustic Z-Pain music release.

speaking balance: gold is Alex, purple is the guest (3 minute bins)

0:00 · Alex 80% · guest 20%0:00 · Alex 80% · guest 20%3:00 · Alex 64.2% · guest 35.8%3:00 · Alex 64.2% · guest 35.8%6:00 · Alex 37.2% · guest 62.8%6:00 · Alex 37.2% · guest 62.8%9:00 · Alex 29.3% · guest 70.7%9:00 · Alex 29.3% · guest 70.7%12:00 · Alex 78.8% · guest 21.2%12:00 · Alex 78.8% · guest 21.2%15:00 · Alex 39.1% · guest 60.9%15:00 · Alex 39.1% · guest 60.9%18:00 · Alex 41% · guest 59%18:00 · Alex 41% · guest 59%21:00 · Alex 76.7% · guest 23.3%21:00 · Alex 76.7% · guest 23.3%24:00 · Alex 24.8% · guest 75.2%24:00 · Alex 24.8% · guest 75.2%27:00 · Alex 27.1% · guest 72.9%27:00 · Alex 27.1% · guest 72.9%30:00 · Alex 49.9% · guest 50.1%30:00 · Alex 49.9% · guest 50.1%33:00 · Alex 59.4% · guest 40.6%33:00 · Alex 59.4% · guest 40.6%36:00 · Alex 36.5% · guest 63.5%36:00 · Alex 36.5% · guest 63.5%39:00 · Alex 30.5% · guest 69.5%39:00 · Alex 30.5% · guest 69.5%42:00 · Alex 89.1% · guest 10.9%42:00 · Alex 89.1% · guest 10.9%45:00 · Alex 50.9% · guest 49.1%45:00 · Alex 50.9% · guest 49.1%48:00 · Alex 56.8% · guest 43.2%48:00 · Alex 56.8% · guest 43.2%51:00 · Alex 42.9% · guest 57.1%51:00 · Alex 42.9% · guest 57.1%54:00 · Alex 44.4% · guest 55.6%54:00 · Alex 44.4% · guest 55.6%
Sharpest disagreement ▶ 8:53 Ranjan rejects the premise that an R&D slowdown damages AI

Ranjan forcefully dismisses Alex's concern over diminishing returns in model training, arguing that over-raising capital for pure R&D actively harms the industry compared to building usable software.

Hardest push from Alex ▶ 8:26 Alex challenges guest for ignoring the business stakes of plateauing models

Alex refuses Ranjan's relaxed framing on model stagnation, aggressively pressing him on the massive VC capital, Anthropic rounds, and nuclear power contracts predicated on continued scaling.

Biggest teaching moment ▶ 35:02 Ranjan breaks down the macroeconomic shift to outcome-based pricing

Ranjan explains to Alex how generative AI forces a structural transition away from seat-based and billable-hour pricing models across advertising, law, healthcare, and enterprise software.

Alex holds their own ▶ 5:14 Alex marshals industry heavyweights to counter guest's optimism

Alex demonstrates subject-matter authority by citing Reuters reporting on Ilya Sutskever, Andreessen Horowitz commentary, and Information reporting to substantiate the reality of LLM pre-training plateaus.

the scores for every segment, with the reasoning behind each
ChapterTopicAlex as informed peerGuest teachingGuest disagreementAlex pushing backWhy
Mike Tyson vs. Jake Paul Fight Banter 7357 Alex aggressively challenges Ranjan's optimism regarding the LLM scaling plateau, citing Ilya Sutskever's Reuters comments, Marc Andreessen, and OpenAI's massive $6B funding round. Ranjan firmly resists Alex's framing, arguing that slowing foundational model development is beneficial for practical productization.
The Shift Toward Reasoning Models and Inference Compute 7522 Alex explains the transition to test-time reasoning compute and cites a TEDAI poker statistic demonstrating efficiency gains over pre-training scale. Ranjan expands on how targeted synthetic data enables startups like Writer to build focused models at fractionally lower costs.
Deconstructing AGI Definitions and Industry Expectations 6433 The pair debate definitions of AGI, with Alex citing Yann LeCun's human-level generalist benchmark while Ranjan questions whether AGI is simply a nebulous marketing and contractual term that distracts from narrower utility.
Gemini's Benchmark Surge and Real-World AI Workflows 5411 Alex highlights Google's experimental Gemini model topping the Chatbot Arena leaderboard, while both agree that superior benchmarks do not immediately overcome interface and workflow preferences.
Practical AI Impact: Writer, Chegg, and Agency Billing Models 6622 Alex discusses Chegg's 99% valuation collapse and shifts in ad agency billing models due to AI efficiency. Ranjan explains the broader transition across SaaS, legal, and healthcare sectors toward outcome-based pricing rather than seat or hourly billing.
Spotify's Recommendation Pitfalls and AI-Generated Music 5511 Ranjan describes how Spotify's recommendation algorithms break under household edge cases like kids' music, then explains his positive tests with Spotify prompt-based playlists and Suno-generated folk tracks.
Podcast Format Updates, Listener Feedback, and Starshield 3200 Alex conducts housekeeping, addresses listener feedback, and shares a listener correction regarding the US Department of Defense's Starshield satellite initiative with SpaceX.
Bluesky's Growth Momentum vs. Threads and Twitter Network Effects 6446 Ranjan argues Bluesky has achieved genuine mainstream momentum post-election, while Alex pushes back citing Max Read's analysis that it remains a specialized echo chamber unable to dislodge Twitter's core network effects.
Apple Smart Glasses Study and Mark Zuckerberg's Z-Pain Track 5521 Ranjan evaluates Apple's exploratory Project Atlas smart glasses against his hands-on experience with Snap Spectacles, and the hosts banter over Mark Zuckerberg's acoustic Z-Pain music release.

Statements from this episode (19)

Opinion
Roy: Netflix is genius for streaming Tyson-Paul fight with polarizing figures
“I think Netflix is genius in promoting their live programming by just bringing out two people that no one wants to see win”
Ranjan Roy Nov 15, 2024 ▶ 1:04
Assertion Supported
Kantrowitz: OpenAI's $6 billion funding is the largest VC round ever
“Opening , I just raised the largest VC round in history, six billion dollars.”
Alex Kantrowitz Nov 15, 2024 ▶ 8:32
Opinion
Roy: Over-funding foundational AI models instead of business applications is a mistake
“Over-raising for the R&D side of things rather than the actual, like, operationalization and building out businesses on top of the existing technology. I mean, again, we've debated this plenty. I think that is a huge mistake in that it actually, you know, pote…”
Ranjan Roy Nov 15, 2024 ▶ 8:58
Opinion
Roy: AI video generation models like OpenAI's Sora are nowhere near interesting
“Video generation is the one area that I do think we are severely, we're not even close to anything interesting, and we've been promised things that are interesting, i.e. Sora but we're very, very far away.”
Ranjan Roy Nov 15, 2024 ▶ 11:56
Prediction Not checkable as stated
Kantrowitz: AI research is pivoting entirely toward reasoning models over sheer scaling
“So I think what we're about to see is a pivot in the AI research field or yes, they might be applying practically some of the models that exist today. But it seems to me like everybody is going to go completely in on this reasoning format. And that is going to…”
Alex Kantrowitz Nov 15, 2024 ▶ 14:26
Prediction Not checkable as stated
Roy: Smaller, domain-specific AI models will outperform massive general foundation models
“And I think in the coming months and years, we're going to start to see some awkward headlines around size matters because smaller will be better in terms of actual models being used.”
Ranjan Roy Nov 15, 2024 ▶ 15:51
Opinion
Roy: AGI's lack of a clear definition distracts from practical technological progress
“There's not one like clear accepted definition or one clear vision that's communicated by the biggest people in the field, the Sam Altman's and everything else. So it remains this murky kind of like dystopian robots taking over who knows what it is, will be a …”
Ranjan Roy Nov 15, 2024 ▶ 20:26
Prediction Not checkable as stated
Kantrowitz: Superintelligence will not arrive through traditional scaling of large language models
“I think that that has led a lot of the investment and a lot of the hype around this that will eventually get there. But it just doesn't seem like it's gonna be through the traditional scaling of LLMs.”
Alex Kantrowitz Nov 15, 2024 ▶ 22:10
Assertion Partly supported
Kantrowitz: Google's experimental Gemini 1114 model currently tops the Chatbot Arena leaderboard
“They have a new experimental Gemini model. It's called Gemini exp one, four. And Do you know about, ah, Chatbot Arena where they test which is the best LLMs? It's currently sitting at the top of Chatbot Arena and kicking the butt of ChatGPT for O, O-one Previe…”
Alex Kantrowitz Nov 15, 2024 ▶ 22:39
Opinion
Roy: Perplexity AI's clear format makes it the biggest traditional search competitor
“Perplexity is when I'm watching a movie, a sports game of any sort, like, it just is so good in just giving you quick information in a really nice format with Additional links to keep exploring in questions that I think for that, like, and which makes me think…”
Ranjan Roy Nov 15, 2024 ▶ 25:57
Assertion Supported
Kantrowitz: Chegg stock is down 99 percent, erasing $14.5 billion in value
“Chegg stock is down 99% from early twenty-twenty-one, erasing some 14.5 billion of market value.”
Alex Kantrowitz Nov 15, 2024 ▶ 31:03
Prediction Not checkable as stated
Roy: AI automation will force SaaS to abandon seat-based pricing for outcomes
“I think for so many of these industries, the way the entire pricing structure changes, it's going to change. And I even think in SAS, that's going to be the case. And there's been a lot of talk around this with even Salesforce's AI agents or in many others is …”
Ranjan Roy Nov 15, 2024 ▶ 35:39
Assertion Supported
Kantrowitz: Bluesky user base soared to 15 million following the US election
“So blue sky is now up to fifteen million users, and it is it's really soaring in the wake of the election.”
Alex Kantrowitz Nov 15, 2024 ▶ 45:09
Assertion Supported
Kantrowitz: Meta's Threads added 15 million new users in November 2024 alone
“Threads has added fifteen million users since the start of November”
Alex Kantrowitz Nov 15, 2024 ▶ 45:28
Prediction Not checkable as stated
Roy: Bluesky's latest growth surge has staying power and threatens Twitter directly
“I think it does have staying power this time.”
Ranjan Roy Nov 15, 2024 ▶ 45:40
Opinion
Kantrowitz: Meta's Threads feels empty and lacks engagement despite massive user growth
“One thing we can tell you, say for sure is it doesn't look like threads is working. I mean, threads added fifteen million people since the start of the month. And Michael Learmonth, who I work with as an editor pointed out to me, he's like, does it feel like t…”
Alex Kantrowitz Nov 15, 2024 ▶ 48:19
Prediction Not checkable as stated
Kantrowitz: Twitter's massive network effects make it incredibly difficult to actually replace
“Twitter is going to be the one with staying power. It's the network effects. It's very, very, it might be the most difficult to replace social network.”
Alex Kantrowitz Nov 15, 2024 ▶ 49:43
Prediction Open · timeframe Nov 2027
Kantrowitz: Apple will soon release Siri-powered smart glasses to rival Meta's offering
“And I think it's not going to be too long until we see Apple build a product like this of their own, if not one with an enhanced Siri to hit one of your most favorite things.”
Alex Kantrowitz Nov 15, 2024 ▶ 50:48
Opinion
Roy: Apple Intelligence is completely useless and the Vision Pro flopped entirely
“Apple intelligence, maybe it'll come around, but it is so, so far from anything we have seen even remotely close to useful. The vision pro flop.”
Ranjan Roy Nov 15, 2024 ▶ 52:06
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.