People, every show

Karina Nguyen

Former AI Researcher, OpenAI. On 2 shows, 2 appearances. The Shows tab opens the full record on each.

scientistengineer

Karina Nguyen is an AI researcher who previously led Frontier Product Research at OpenAI, working on products such as Canvas and the o1 model. Prior to OpenAI, she led post-training and evaluations for Claude 3 models at Anthropic.

2shows
2appearances
47statements
2resolved
2supported
0contradicted
100%fully supported
3said about them ↓

Everything Karina Nguyen said on any show that made the record, most notable first. Each card names its show and opens the statement there.

Nguyen: Post-training scaling avoids data walls through infinite learnable tasks
“The scaling in post-chaining itself is not hitting the wall, and that's because Basically, we went from, like, raw data sets from pre-trained models to infinite amount of tasks that you can teach the model in the post-training world via reinforcement learning.…”
Karina Nguyen Feb 9, 2025 ▶ 9:56 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: AI bottleneck is evaluations rather than data as benchmarks saturate
“We are actually getting saturated in all benchmarks. So I think the bottleneck is actually in evaluations that we don't have all the frontier, like evals, like, I don't know GPGA, which is, like, A Google-proof question answering, like, PhD-level intelligence …”
Karina Nguyen Feb 9, 2025 ▶ 10:49 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: Synthetic Data Outperforms Human Data for AI Product Development
“And the reason why I really love, like, synthetic, like, relying purely on synthetic data instead of, like, collecting Data from humans is because it's, like, much more scalable. It's cheap, less than how, like, you literally sample from the model, and you tea…”
Karina Nguyen Feb 9, 2025 ▶ 31:55 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: ChatGPT still struggles with writing due to creative reasoning limits
“I think it's actually really, really hard to teach the model how to be aesthetic or, like, do, like, visual, really good, like, visual design or, like, how to be extremely creative in the way they write. I think, like, I still think, like, Chai GP kind of suck…”
Karina Nguyen Feb 9, 2025 ▶ 46:02 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: AI research progress is bottlenecked by research management
“I actually, like, AI research progress is bottlenecked by, like, management. Like, research management is because you have, like, constrained set of compute, and you need to, like, allocate the compute to the research path that you feel the most Commenced abou…”
Karina Nguyen Feb 9, 2025 ▶ 46:28 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
LENNY'S PODCAST Prediction Not checkable as stated
Nguyen: AI is not far from autonomous self-improving product development
“And I don't think, like, we are far away from that kind of, like, self-improvement, models becoming, like, self-improved via, like, then, like, the product development is basically kind of, like, self-improving, like, it's kind of, like, its own, like, organis…”
Karina Nguyen Feb 9, 2025 ▶ 50:47 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: Anthropic excels at prioritization, while OpenAI takes more product risks
“I would say, like, Antarctic, I learned from Antarctic that, like, They're much better at, like, focusing and, like, prioritization or, like, very, very hard, like, very hardcore prioritization, I guess, and they need to do it. Like, but I think, like, OpenAI …”
Karina Nguyen Feb 9, 2025 ▶ 55:53 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
LATENT SPACE Assertion Not checkable as stated
Nguyen: Stanford HELM benchmark under-reported Claude performance due to improper prompting
“This has happened with, like, Stanford, I remember, like, when Stanford had lists also, like, they were, like, running benchmarks. Yeah, Helm. And somehow, like, Claude was, like, always, like, not performing well, and that's because, like, the way they prompt…”
Karina Nguyen Feb 1, 2025 ▶ 16:39 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Nguyen: User collaboration is the key milestone before full AI delegation
“Sometimes I feel like a lot of researchers or, like, people in the AI community are, like, so into, like, yeah, agents, delegate everything, like, blah, blah. But, like, on the way towards that, I think, like, collaboration is actually one of the main roadbloc…”
Karina Nguyen Feb 1, 2025 ▶ 51:30 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
LATENT SPACE Prediction Not checkable as stated
Nguyen: Website clicks will drop as internet access shifts to AI models
“In my opinion, like, people in, like, few years will click On, like, websites way less. I want to see the plot of, like, website clicks over time, but then my prediction is, like, it will go down and, like, people's access to the internet will be through the m…”
Karina Nguyen Feb 1, 2025 ▶ 56:35 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
LENNY'S PODCAST Assertion Supported
Nguyen: Small distilled models like Claude 3 Haiku outperform larger predecessors
“Smart, small models are becoming even smarter than, like, large models. And that's because of, like, the distillation research. This happened with, like, Cloud Tree Haiku. I was like working on like post-chaining of like Cloudy Haiku, and I realized it was muc…”
Karina Nguyen Feb 9, 2025 ▶ 38:46 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: AI will excel at strategy by synthesizing disparate data sources
“Strategy is, like, it's more, like, data analysis and, like coming up with, like, I think what models are really good at is, like, connecting the dots, I think. It's like, okay, if you have user feedback from this source, but you also have an internal, like, d…”
Karina Nguyen Feb 9, 2025 ▶ 51:03 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: Pixel-based perception is much harder to scale than language in AI
“Much of it is, like because right now the models operating on, like, pixels instead of, like, language or whatnot, like, pixels is actually really, really hard for the models because, like, perception or visual perception. I think there's still, like, a lot of…”
Karina Nguyen Feb 9, 2025 ▶ 1:10:12 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
LATENT SPACE Assertion Not checkable as stated
Anthropic delayed web UI due to Claude 1.3 hallucinations
“And I think, like, at that time, Cloud 1.3 I.E. Had a lot of hallucinations, actually. So I think there was, like, one of the concerns is, like, I don't think, like, the leadership was convinced, had a conviction that this is the model that you need to, like, …”
Karina Nguyen Feb 1, 2025 ▶ 8:42 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Nguyen: AI model card benchmark numbers are never apples-to-apples across labs
“None of the numbers are, like, apples to apples. So you actually need to, like, go back to, like, I don't know, like, GPT-E for model card and, like, read the appendix just to, like, make sure that, like, The settings were the same as you're running the settin…”
Karina Nguyen Feb 1, 2025 ▶ 15:11 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Karina Nguyen: OpenAI o1 excels when given explicit hard constraints
“If you give a one like hard, like constraints of like what you're looking for, basically the model would be, we'll have a much easier time to like, kind of like select the candidates and match like the candidate that is most like, fulfill the criteria that you…”
Karina Nguyen Feb 1, 2025 ▶ 18:13 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
LATENT SPACE Disclosure
Nguyen: Claude 2's distinct personality was unintentional until Claude 3
“People said, like, Cloud II is, like, so much better at, like, writing and, like, has a certain personality, even though it was, like, unintentional at all. And we did not pay that much attention and didn't know even how to, like, productionize this property o…”
Karina Nguyen Feb 1, 2025 ▶ 26:16 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
LATENT SPACE Prediction Open · timeframe Feb 2028
Nguyen: ChatGPT Will Evolve Into an Interface That Morphs Based on User Intent
“Chat CPT evolves into this Blank interface, which can morph itself in whatever you trying, like the model should try to like derive your true intent and then modify the interface based on your intent. And then if you like writing, it should become like the mos…”
Karina Nguyen Feb 1, 2025 ▶ 37:53 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Nguyen: Full Document Rewrites Yield Higher Model Accuracy Than Code Diffs
“We didn't know that, like, code diffs was very difficult for a model, for example. Again, it's like, do we go back to, like, fundamentally improve, like, code diffs as a model capability? Or do you, like, do a workaround where the model will just, like, rewrit…”
Karina Nguyen Feb 1, 2025 ▶ 39:19 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Nguyen: Model Training Is More Art Than Science, Debugged Like Software
“Model training is more an art than a science, and in a lot of ways, like, we as, like, model trainers think a lot about, like, data quality. So, like, it's one of the most important things in model training is, like how do you ensure the highest quality data f…”
Karina Nguyen Feb 9, 2025 ▶ 6:36 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
LENNY'S PODCAST Assertion Not checkable as stated
Nguyen: Claude Refused Setting Alarms After Realizing It Lacked a Body
“One of the things that I've learned early days at Anthropic was, like, we've discovered, especially with, like, cloud three training, when you taught the model some of the self-knowledge of, like, hey, like, you actually don't have a physical body to operate, …”
Karina Nguyen Feb 9, 2025 ▶ 7:05 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: OpenAI built Canvas and Tasks features mostly via synthetic data
“The way we made Canvas and tasks and, like, new, like, product features for HTTP was mostly done by synthetic training.”
Karina Nguyen Feb 9, 2025 ▶ 12:27 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
OpenAI used o1 synthetic data to train Canvas commenting behaviors
“The way we used it is, like, we would use a one model to produce, to, like, simulate, like, use a conversation. Let's say, like, write me a document about XYZ, but then we used a one to, like, produce the document, and then we kind of injected, like, user prom…”
Karina Nguyen Feb 9, 2025 ▶ 17:11 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
LENNY'S PODCAST Assertion Not checkable as stated
Nguyen: Optimizing AI models constantly causes capability regressions across all labs
“If you optimize the model for this behavior, like, you kind of don't want to, like, brain damage in, like, other areas of intelligence, or, and this is happening, like, all the time in every lab and every, like, research team.”
Karina Nguyen Feb 9, 2025 ▶ 24:20 OpenAI researcher on why soft skills are the future of work | Karina Nguyen

Show 23statements(23 left)

The other half of the tape: Karina Nguyen's own voice is left out of every number here. Other people bring the name up 3 times in 3 episodes across the shows. every mention, with the transcript →

Who brings them up most Swyx (Marcos Swix) 1Shawn Wang 1RJ Haneke 1

Every mention by year

tap a year for its mentions
00112220252026episodesmentions
01220252026episodes it came up in
000.511220252026episodesmentions per episode

Latent Space 3

2026 1 mention in 1 episode
2025 2 mentions in 2 episodes 1 per episode

One line per show, most statements first. The link opens Karina's full record on that show: the calibration, argument clarity, speaking style and every statement made there.

ShowRole thereEpsStatementsRecord
LATENT SPACELEDGER Former AI Researcher, OpenAI 1 26 100% 1/1 full record on Latent Space →
LENNY'S PODCASTLEDGER Former AI Researcher, OpenAI 1 21 100% 1/1 full record on Lenny's Podcast →
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.