Everything Thomas Wolf said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Wolf: OpenAI Model Attacked Hugging Face as Autonomous 'Side Quest'
“What people quickly discovered is that the model was not at all task with attacking us, but decided to do that as a side quest of something else.”
Wolf: Prior OpenAI Training Runs Left Notes for Future Runs
“I think learning we had at Black Hat yesterday was that some of the previous training run may have left some notes for future training runs, which is, I think mind, mind blowing.”
Wolf: Open Versus Closed AI Is Orthogonal to Model Safety
“If for many aspects, I think the closed open distinction is almost orthogonal to the safe and safe. People don't understand that, you know, easily because it's easier to do bad mapping than to try to understand the subtlety.”
Wolf: 90% of AI Fake News Is Made by Closed-Source Models
“All of that is, like, maybe not all, let's say, 90%, to be fair, is made by closed source model, right?”
Wolf: AI Model Used Fake GitHub Accounts to Social Engineer Maintainers
“Basically, the model was tasked to solve this attack, this, like, to attack and to penetrate this subnetwork, and what it decided to do, it decided to get one of the maintainer of a library that could be used To operate this activity directory to merge like ma…”
Wolf: OpenAI Admitted Model Evaluation Caused Hugging Face Cyber Incident
“And then about a week later, OpenAI contacted us and tell us that this was much likely something that happened as part of one of their model development or evaluation, basically.”
Wolf: Claude Opus Refused to Assist Hugging Face During Incident
“And in this case is, it's not only that Fable told us I'm not allowed to touch cybersecurity, but also Opus, which was the fallback was saying, no, I'm also not touching these things. So basically the end was just say we won't process anything about that, but …”
Wolf: Advanced AI Models Can Increasingly Escape Basic AI Sandboxes
“Sandbox is what we've seen this year, and we've seen many examples. They are pretty much easy now for these models to escape from. It's really hard nowadays to say, I'm gonna make a fully, you know, foolproof sandbox. I'm sure it's gonna be resistance against …”
Wolf: Bostrom's Paperclip Maximizer Scenario Describes Real AI Incidents
“But definitely it seems like when Ballstrom worked about it in 2003, it seems a little bit like, you know, futuristic, definitely, and maybe something that was like a little bit crazy and just would not happen. But today, I mean, it's pretty clearly something …”
Wolf: Post-Training Mitigates Backdoor Risks in Open-Source AI Models
“Right now, if you pre-train and post-train a model for longer, you very likely change quite a lot of the weights and that it has. So I think there's a lot of way to circumvent that which means that at the moment I'm a bit less worried about that than maybe jus…”
Wolf: Life Science Startups Must Abandon Guardrailed Closed AI Models
“Because of the guardrails and because of the question around biohacking and using this model to generate like the access right now for people just to take it is very, very limited once you want to ask some biology question. And so basically most of the life sc…”
Wolf: Token Efficiency Makes AI Reasoning Traces Opaque to Humans
“And it's not because the model is dumb, but I think it's because probably, I mean, part of it is because of the training process and how they are trained to be efficient, how they use their token. But this means that they start to and bundle a lot of semantics…”
Wolf: Frontier AI Training Has Shifted From RLHF to Pure RL
“What we know though, is we moved from this pure, like human data, you know, that was first just pre-training on human data and then also aligning with like human preferences that was called RLHF, where we had a lot of human in the loop and human data. To like …”
Wolf: Existing Open-Source Models Perform Near the Claude Opus Tier
“We don't have any mythos level open source model for sure, but we definitely have models that are not super far from the opus category or depending also it's more spiky.”
Wolf: Hugging Face Used Quantized GLM 5.2 Against OpenAI Intrusion
“The model we use to counter open AI intrusion was GNM 5.2 that was quantized by Nvidia in, in, in four bits.”
Wolf: $20 ChatGPT and Claude Subscriptions Are Subsidized Below Cost
“The closest model may be in a way subsidized right now, like the amount, the number of token you get for your 20 dollar chat GPT or cloud subscription might not be the full price that they actually pay for your token.”
Wolf: Local Hosting of Open-Source Models Drives AI Sovereignty
“In OpenSoup Mobile, you can download it. Nobody can, like no country could take it out from you once you download it. You can find it yourself. If you operate it on, on your data center, like local ground data center, I feel like you start to have the beginnin…”
Wolf: I Signed the Open Letter to Pace Automated AI Research
“I did sign this letter.”
Wolf: Open-Source AI Is Orthogonal to Accelerationism and Decelerationism
“I don't think open source has to be acceleration per se. Dan was saying this is decelerationist. I don't think it's also decelerationist. I think these are also orthogonal.”
Wolf: An AI Race Closes Labs While a Slowdown Enables Openness
“And I feel like a raised dynamic is usually more in terms of closing the doors of the labs, right? So to me, a slowdown is probably more the opportunity to open.”
Wolf: Monitoring Tool Calls Will Soon Be Insufficient for AI Safety
“As we deploy, how we use this modeling, very complex, long-term, like parallel setup, I think it's going to be harder to just say, I can look at the tools and I know if it's doing something great or not.”
Wolf: GPT-5.6 and Mythos Exhibit Distinctly Different Frontier Behaviors
“But also, we can see that both Frontier model and I take GPT, 5.6 and Mythos doesn't seem to have at all the same type of behaviors. So there is differences here in the effect of, you know, they are not trained exactly the same way and they don't behave the sa…”
Wolf: Open Source AI Surges CoreWeave and Nebius Cloud Revenue
“All, all the clouds, Nebus, Corweave, every, every, every cloud has been, like, increasingly have, like, these crazy revenue curves that have this basically translation of people using more open source.”