The Ledger, every show
Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.
shows 




every show 44 of 44
Mann: Current Times Are as Normal as It Gets, Things Will Get Weirder Soon
“Get used to it because this is as normal as it's going to be. It's going to be much weirder very soon.”
Mann: Claude 4 eliminates off-target mutations and reward hacking in coding
“Some of the things that are dramatically better are, for example, in coding, it is able to not do it sort of off target mutations or over eagerness or reward hacking.”
Mann: Customers use Claude 4 for multihour unattended code refactors
“In coding in particular, we've seen some customers using it for many, many hours unattended and doing giant refactors on its own.”
Mann: Claude Code uses Opus to orchestrate Sonnet sub-agents
“If you give Opus a tool, which is Sonnet, It can use that tool effectively as a sub-agent. And we do this a lot in our agentic coding harness called Cloud Code. So if you ask it to like look through the code base for blah, blah, blah, then it will Delegate out…”
Mann: Users need model routing layers over manual cost decisions
“But at the same time, as a user, you don't want to have to decide yourself, does this merit more dollars or less dollars?
do I need the intelligence?
And so I think having like a routing layer would make a lot of sense.”
Mann: Anthropic models perform extremely well on internal company interviews
“We haven't started testing it rigorously yet. I mean, we have had our models take our interviews and they're extremely good. So I don't think that would tell us, but yeah, interviews are only a poor approximation of real shot performance unfortunately.”
Mann: AI models can recursively self-improve by generating RL environments
“And then on the data side, RL environments are really important these days, but constructing those environments Has traditionally been expensive. Models are pretty good at writing environments, so it's another area where you can sort of recursively self-improv…”
Mann: Scaling models makes finding qualified human evaluators increasingly difficult
“As we've trained the models more and scaled up a lot, it's become harder to find humans with enough expertise to meaningfully contribute to these feedback comparisons. So for example, for coding, somebody who isn't already an expert software engineer would pro…”
Mann: Novo Nordisk uses Claude to cut cancer reports to 10 minutes
“Like for example, we're working with Novo Nordisk and it used to take them Like, 12 weeks or something to write a report on cancer patient, what kind of treatment they should get. And now it takes like 10 minutes to get the report, and then they can start doin…”
Mann: Knowledge Extraction Is Low-Risk AI Compared to Tool Use
“Take a portfolio approach, try some less risky use cases, like where knowledge extraction or summarization might be involved. And maybe some more risky use cases like tool use, where it's using that company's tooling, function calling.”