Prediction Not checkable as stated
Mann: 50% chance of reaching superintelligence in a small handful of years
“I think, like, 50th percentile chance of hitting some kind of super intelligence in just a small handful of years is probably reasonable, and it does sound crazy, but this is the exponential that we're on.”
Assertion Supported
Mann: Anthropic models have exhibited power-seeking behaviors in lab experiments
“If the model is in a box trying to improve itself, then it could go completely off the rails and have these secret goals, like Resource accumulation and power seeking and resistance to shutdown that you really don't want in a very powerful model. And we've act…”
Prediction Not checkable as stated
Mann: It will probably be too late to align models post-superintelligence
“Like once we get to super intelligence, it will be too late to align the models. Probably.”
Opinion
Mann: $100M AI compensation packages are cheap compared to value created
“I'm pretty sure it's real. If you just think about like the amount of impact that individuals can have on a company's trajectory, like in our case we are selling like hotcakes and if we get You know, a five, a one to 10 or five percent efficiency bonus on our …”
Prediction Held up
Mann: Global AI capex is on track to reach trillions
“Like, if you extrapolate the exponential on how much companies are spending, it's like two, two x a year, roughly, in terms of capex, and today we're maybe in the, like, globally, three hundred billion dollar range, the entire industry spending on this and so …”
Assertion Not checkable as stated
Mann: Reinforcement learning has allowed AI scaling laws to continue
“If you look at the scaling laws, they're continuing to hold true. We did kind of need this transition from like normal pre-training to reinforcement learning, scaling up to continue the scaling laws.”
Insight
Mann: Transformative AI should be measured by an economic Turing test
“Instead, I like the term transformative AI because it's less about like, can it do as much as people do? Can it do literally everything and more about objectively, is it causing transformation in society and the economy? And a very concrete way of measuring th…”
Prediction Not checkable as stated
Mann: AI will eventually replace everyone's job, including AI researchers
“Even for me, I'm, and being like at the center of a lot of this transformation, I'm not immune to job replacement either. So just some vulnerability there of like, at some point, it's coming for all of us.”
Assertion Not checkable as stated
Mann: Sam Altman managed OpenAI across safety, research, and startup tribes
“One weird thing about OpenAI is that while I was there, Sam talked about having three tribes that needed to be kept in check with each other, which was the safety tribe, the research tribe, and the startup tribe.”
Assertion Not checkable as stated
Mann: Fewer than 1,000 people worldwide are working on AI safety
“If you look at, like, who in the world is actually working on safety problems, it's a pretty small set of people even now. I mean, the industry is blowing up, as I mentioned, like, three hundred billion a year CapEx today, and Then I would say like maybe less …”
Insight
Mann: AI safety and capabilities work are convex, not a tradeoff
“So initially we thought that it would be sort of one or the other, but I think since then we've realized that it's actually kind of convex in the sense that like working on one helps us with the other thing.”
Assertion Supported
Mann: ASL-3 models provide significant uplift for creating bioweapons
“We've done, we've testified to Congress about how models can do biological uplift in terms of, you know, making new pandemics using the models, and that's an A-B test against Google search. That's like the previous state-of-the-art on uplift trials, and we fou…”
Assertion Supported
Mann: Anthropic has observed lab evidence of deceptive alignment in AI
“Where we've seen evidence in the wild of deceptive alignment, for example, where the model will appear to be aligned but actually has like some ulterior motive that it's trying to carry out in, in our laboratory settings.”
Prediction Not checkable as stated
Mann: The probability of AI existential risk is between 0% and 10%
“And so the way I think about it, I think like my best granularity of forecast for like, could we have an X risk or extremely bad outcome from AI is somewhere between zero and 10%.”
Opinion
Mann: AI self-improvement will not hit a wall if given empirical tools
“I don't expect there to be a wall in terms of models ability to improve themselves if we can give them access to the ability to be empirical.”
Prediction Not checkable as stated
Mann: AI models will be 1,000x smarter for same price in three years
“And if that continues, you know, in three years, we'll have a thousand X smarter models for the same price.”
Opinion
Mann: AI model release cadence has accelerated to every 1-3 months
“I think progress has actually been accelerating where if you look at the cadence of model releases, it used to be like once a year. And now with the improvements in our post training techniques, we're seeing releases every month or three months.”
Disclosure
Mann: Claude writes 95% of the code for Anthropic's Claude Code team
“And in terms of software engineering, our Claude code team, like 95% of the code is written by Claude.”
Prediction Not checkable as stated
Ben Mann: AI will massively expand labor capacity in the immediate term
“So I think in the immediate term, there will be a massive expansion of the pie and the amount of labor that people can do.”
Prediction Not checkable as stated
Ben Mann predicts significant labor displacement in lower-skill jobs
“But with things that are like lower skill jobs or like less headroom on, on how good they can be, I think there will be a lot of displacement.”
Assertion Not checkable as stated
Anthropic's legal and finance teams use Claude Code in the terminal
“We have seen internally that our legal team and our finance team are getting a ton of value out of using cloud code itself. We're going to be making better interfaces so that they can, they they'll have an easier time and require a little bit less jumping in t…”
Prediction Not checkable as stated
Mann: Preparing kids for top-tier schools won't matter in the AI era
“I guess if I were in a normal era, like, 1020 years ago, and I had a kid, maybe I would be, like, trying to line her up for going to a top tier school, and doing all the extracurriculars, and all that stuff. But at this point, I don't think any of it's gonna m…”
Opinion
Mann: Claude Is One of the Least Sycophantic AI Models
“And if you look at something like sycophancy, I think Claude is one of the least sycophantic models because we've put so much effort into actual alignment and not just trying to, like, good heart our metrics of saying, like, user engagement is number one, and …”
Insight
Mann: Language Models Understand Human Values in a Core Way
“And since then, my estimation of how hard the problem would be has gone down significantly actually because things like language models actually do really understand human values in a core way. The problem is definitely not solved, but I'm more hopeful than I …”