Pavel Izmailov

7 statements across 1 episodes · 2 bullish · 0 bearish · 1 people on the record · first statement Jan 15, 2026 by Pavel Izmailov · across every show →

On the record as a speaker too: Pavel Izmailov's record, appearances and statements → this page counts the times other people say the name.

Everything said about Pavel Izmailov, oldest first

Jan 15, 2026 neutral
Insight
Izmailov: Neural network operations may not be explainable in human terms
“We want to understand it at a lower level, and it is very possible that that's just not fully possible. Like, it is some computational process that leads to some results. It doesn't have to be the case that you can Kind of describe it in human terms and kind o…”
Pavel Izmailov Jan 15, 2026 ▶ 24:22 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026 positive
Assertion Not checkable as stated
Izmailov: Large-scale RL has not produced coherently misaligned models
“A lot of people were worried that with large-scale RL, we will have Some completely new types of issues with the models, like this kind of coherent misalignment that will just emerge where the models are evil in some ways across many scenarios. And we are not …”
Pavel Izmailov Jan 15, 2026 ▶ 21:47 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026 neutral
Assertion Not checkable as stated
Izmailov: OpenAI had three alignment and safety teams during his tenure
“Even at OpenAI, when I was there were three teams related to alignment and safety.”
Pavel Izmailov Jan 15, 2026 ▶ 6:43 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026
Insight
Izmailov: Deterministic data transformations create information for computationally bounded models
“But with a limit on the compute, it's actually very possible to apply deterministic transformations to the data. And create information through that.”
Pavel Izmailov Jan 15, 2026 ▶ 35:37 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026
Assertion Not checkable as stated
Izmailov: AI researchers cannot reliably trace model behaviors to pre-training sources
“We don't really know what's the source of this type of behaviors, but that's also true for a lot of other behaviors in the models with, like, even the good ones. We don't really, we cannot always pin down, like, where they come from in the pre-training.”
Pavel Izmailov Jan 15, 2026 ▶ 3:57 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026
Assertion Supported
Izmailov: Text data carries more structural information per token than images
“So for example, we can approximate it from the scaling laws and we can, for example, say that text data has more structural information according to this measure than image data at the same kind of amount of yeah, tokens.”
Pavel Izmailov Jan 15, 2026 ▶ 37:17 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Jan 15, 2026 positive
Disclosure
Izmailov: Interpretability tools are growing more useful internally at Anthropic
“So we are still pretty far from the dream that we will Fully understand everything that happens in the model, but these tools are becoming increasingly more useful internally at Anthropic in particular, and also there is constant progress, and it's pretty fasc…”
Pavel Izmailov Jan 15, 2026 ▶ 23:36 The Evaluators Are Being Evaluated — Pavel Izmailov (Anthropic/NYU)
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.