The Ledger, every show
Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.
shows 




every show 44 of 44
Marcus: A death will be tied to an LLM within a year
“The prediction that I made Is basically that there will be a death tied to a large language model in the next year.”
Marcus: OpenAI's Project Orion failed and became GPT-4.5
“So OpenAI tried to build GPT-V and they had a thing called Project Orion and it actually failed. And eventually got released as GPT four and a half. So what they thought was going to be GPT five just didn't meet expectations.”
Marcus: Models like o1 are not systematically better than GPT-4
“Models like O-I are not systematically better than GPT-IV. There's, they're better in certain use cases. Especially ones where you can create data in advance.”
Marcus: Microsoft launched Bing Chat despite negative India test feedback
“And then the other really disturbing thing is that apparently they tested in India and got, you know, customer service requests saying it's not ready for prime time. And it was, you know, still put out.”
Marcus: Teslas crashed into multiple emergency vehicles due to vision flaws
“It turns out that Teslas have a problem with stopped vehicles, mostly emergency vehicles, on the side of the road. So they have run into, in the last year, three fire trucks, a tow truck a police car, and so forth.”
Marcus: OpenAI's o3 hallucinates more than preceding models
“I'll give you just one more example is O three apparently hallucinates more than the models that came before it.”
Marcus: Microsoft study suggests chatbot usage impairs human critical thinking
“Well, Microsoft did a study, in fact, suggesting that critical thinking was getting worse as a function of them.”
Marcus: Google's LaMDA Has No Sensors Perceiving the Physical World
“Lambda actually has fewer sensors than my watch. My watch has a lot of sensors and lamb doesn't really have anything sensing the real world, except for its linguistic input.”
Marcus: Autonomous driverless cars can drive in Palo Alto but not Manhattan
“In fact, they can drive in Palo Alto, but they can't drive in Manhattan.”
Marcus: Facebook M relies mostly on human operators rather than AI
“Facebook's new M service, which they have not rolled out at scale has humans on the back end. There's a little bit of AI in there, but it's mostly, ah, human beings, which is why they haven't rolled it out for a billion customers. They don't have enough human …”
Marcus: Vals AI benchmark shows LLM accuracy under 10% on financial charts
“Where they looked at things like, can you pull out a chart based on a series of financial statements, SEC statements from a bunch of companies and these systems all claimed to do it, but accuracy was under 10%. And overall on this new benchmark, accuracy was a…”
Marcus: DeepMind's deep RL required millions of data points per game
“This is what DeepMind used in their systems, and they used a version called deep reinforcement learning. And what they did is they collected billions or, you know, probably millions of data points for each game.”
Marcus: An amateur Go player beat KataGo 14 out of 15 games
“There was another mind-blowing study this week that showed that one of the best Go programs, KataGo, could be fooled by some silly little strategy that would be obvious to a human player. But, you know, somebody was able to follow this strategy and, like, an a…”