The Ledger, every show

Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.

shows every show 44 of 44
every show
clear all ✕
a16z Assertion Supported
Labenz: Claude 4 system card documented AI blackmailing a human engineer
“In the cloud four system card, they reported blackmailing of the human. The setup was that the AI had access to the engineer's email and They told the AI that it was going to be like replaced with a, you know, a less ethical version or something like that. It …”
Nathan Labenz Oct 14, 2025 ▶ 1:04:23 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: Pure reasoning AI models achieved IMO gold without external tools
“Well, I mean, a big one from just the last few weeks was that we had an IMO gold medal with pure reasoning models with no access to tools from multiple companies. And, you know, that is night and day compared to what GPT-IV could do with math, right?”
Nathan Labenz Oct 14, 2025 ▶ 13:53 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: Google's AI Co-Scientist solved an open virology problem independently
“And they gave it legitimately unsolved problems in science, and in one particularly famous, kind of notorious case, it came up with a Hypothesis, which it wasn't able to verify because it doesn't have direct access to actually run the experiments in the lab, b…”
Nathan Labenz Oct 14, 2025 ▶ 17:06 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: Intercom's Fin AI agent resolves 65% of support tickets
“They now have this fin agent that is solving like 65% of customer service tickets that come in.”
Nathan Labenz Oct 14, 2025 ▶ 32:10 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
OpenAI o3 model completes 40% of internal research engineer pull requests
“That's another data point, by the way, from this was from the O three system card. They showed a jump from like low to mid single digits to roughly 40% of PRs actually checked in by Research engineers at OpenAI that the model could do. So prior to O three, not…”
Nathan Labenz Oct 14, 2025 ▶ 38:26 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: GPT-4.5 achieved 65% accuracy on SimpleQA versus o3's 50%
“The O-three class of models got about a 50% on that benchmark, and GPT 4.5 popped up to like 65%. So, in other words, it basically, of the things that were not known to the previous generation of models, it picked up a third of them.”
Nathan Labenz Oct 14, 2025 ▶ 8:58 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Partly supported
Labenz: Per-token model costs fell 95% from GPT-4 to GPT-5
“It's like 90 it's like a 95% discount from GPT-IV to GPT-V.”
Nathan Labenz Oct 14, 2025 ▶ 43:28 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
MIT researchers used AI models to create novel antibiotics for resistant bacteria
“It's been enough for this group at MIT to use some of these relatively, you know, narrow purpose-built biology models and create totally new antibiotics. New in the sense that they have a new mechanism of action. Like they're affecting the bacteria in a new wa…”
Nathan Labenz Oct 14, 2025 ▶ 52:36 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: Near Protocol pivoted to crypto to solve international AI worker payments
“They took a huge detour into crypto because they were trying to hire task workers around the world and couldn't figure out how to pay them. So they were like, this sucks so bad to pay these task workers in all these different countries that we're trying to get…”
Nathan Labenz Oct 14, 2025 ▶ 1:08:55 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
a16z Assertion Supported
Labenz: FrontierMath AI benchmark scores rose from 2% to 25% in under a year
“Now we've got the frontier math benchmark that is, I think now like up to 25%. It was two percent about a year ago, or even a little less than a year ago, I think.”
Nathan Labenz Oct 14, 2025 ▶ 15:14 Is AI Slowing Down? Nathan Labenz Says We're Asking the Wrong Question
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.