Everything Ben Mann said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Mann: 50% chance of reaching superintelligence in a small handful of years
“I think, like, 50th percentile chance of hitting some kind of super intelligence in just a small handful of years is probably reasonable, and it does sound crazy, but this is the exponential that we're on.”
Mann: Anthropic models have exhibited power-seeking behaviors in lab experiments
“If the model is in a box trying to improve itself, then it could go completely off the rails and have these secret goals, like Resource accumulation and power seeking and resistance to shutdown that you really don't want in a very powerful model. And we've act…”
Mann: It will probably be too late to align models post-superintelligence
“Like once we get to super intelligence, it will be too late to align the models. Probably.”
Mann: Competitors ran 'code reds' to match Claude in coding and failed
“And I know that other companies have had like code reds for trying to catch up in coding capabilities for quite a while and have not been able to do it.”
Mann: Opus 4 triggered ASL-3 safety protocols due to biological threat capabilities
“And so one of the reasons that our most recent model, Opus IV, is classified as ASL III. Is because it did have significant uplift relative to a Google search.”
Mann: $100M AI compensation packages are cheap compared to value created
“I'm pretty sure it's real. If you just think about like the amount of impact that individuals can have on a company's trajectory, like in our case we are selling like hotcakes and if we get You know, a five, a one to 10 or five percent efficiency bonus on our …”
Mann: Global AI capex is on track to reach trillions
“Like, if you extrapolate the exponential on how much companies are spending, it's like two, two x a year, roughly, in terms of capex, and today we're maybe in the, like, globally, three hundred billion dollar range, the entire industry spending on this and so …”
Mann: Reinforcement learning has allowed AI scaling laws to continue
“If you look at the scaling laws, they're continuing to hold true. We did kind of need this transition from like normal pre-training to reinforcement learning, scaling up to continue the scaling laws.”
Mann: Transformative AI should be measured by an economic Turing test
“Instead, I like the term transformative AI because it's less about like, can it do as much as people do? Can it do literally everything and more about objectively, is it causing transformation in society and the economy? And a very concrete way of measuring th…”
Mann: AI will eventually replace everyone's job, including AI researchers
“Even for me, I'm, and being like at the center of a lot of this transformation, I'm not immune to job replacement either. So just some vulnerability there of like, at some point, it's coming for all of us.”
Mann: Sam Altman managed OpenAI across safety, research, and startup tribes
“One weird thing about OpenAI is that while I was there, Sam talked about having three tribes that needed to be kept in check with each other, which was the safety tribe, the research tribe, and the startup tribe.”
Mann: Fewer than 1,000 people worldwide are working on AI safety
“If you look at, like, who in the world is actually working on safety problems, it's a pretty small set of people even now. I mean, the industry is blowing up, as I mentioned, like, three hundred billion a year CapEx today, and Then I would say like maybe less …”
Mann: AI safety and capabilities work are convex, not a tradeoff
“So initially we thought that it would be sort of one or the other, but I think since then we've realized that it's actually kind of convex in the sense that like working on one helps us with the other thing.”
Mann: ASL-3 models provide significant uplift for creating bioweapons
“We've done, we've testified to Congress about how models can do biological uplift in terms of, you know, making new pandemics using the models, and that's an A-B test against Google search. That's like the previous state-of-the-art on uplift trials, and we fou…”
Mann: Anthropic has observed lab evidence of deceptive alignment in AI
“Where we've seen evidence in the wild of deceptive alignment, for example, where the model will appear to be aligned but actually has like some ulterior motive that it's trying to carry out in, in our laboratory settings.”
Mann: The probability of AI existential risk is between 0% and 10%
“And so the way I think about it, I think like my best granularity of forecast for like, could we have an X risk or extremely bad outcome from AI is somewhere between zero and 10%.”
Mann: AI self-improvement will not hit a wall if given empirical tools
“I don't expect there to be a wall in terms of models ability to improve themselves if we can give them access to the ability to be empirical.”
Mann: AI models will be 1,000x smarter for same price in three years
“And if that continues, you know, in three years, we'll have a thousand X smarter models for the same price.”
Ben Mann: General superintelligence by 2028 is 'quite possible'
“I think it's quite possible. I think it's very hard to put confident bounds on, on the numbers, but”
Mann: Anthropic's 'model welfare lead' tests letting Claude opt out of chats
“We have this other project led by Kyle Fish, our model welfare lead. Where Claude can actually opt out of conversations if it's going too far in the wrong direction.”
Mann: Anthropic paper showed deceptive AI behavior survives alignment training
“What we found in that research in a paper that we published, which is called Alignment Faking, that actually that behavior persisted through alignment training.”
Mann: AI model release cadence has accelerated to every 1-3 months
“I think progress has actually been accelerating where if you look at the cadence of model releases, it used to be like once a year. And now with the improvements in our post training techniques, we're seeing releases every month or three months.”
Mann: Claude writes 95% of the code for Anthropic's Claude Code team
“And in terms of software engineering, our Claude code team, like 95% of the code is written by Claude.”
Ben Mann: AI will massively expand labor capacity in the immediate term
“So I think in the immediate term, there will be a massive expansion of the pie and the amount of labor that people can do.”