People, every show

Luca Soldaini

Member of Technical Staff, Microsoft AI. On 2 shows, 1 appearance. The Shows tab opens the full record on each.

scientistengineer@soldni ↗LinkedIn ↗soldaini.net ↗

Luca Soldaini co-led the development of the open-source OLMo model family and curated the 3-trillion-token Dolma pre-training dataset at Ai2. They now develop reasoning and thinking models at Microsoft AI, following prior work on question answering at Amazon Alexa AI and a Ph.D. from Georgetown University.

2shows
1appearances
11statements
6resolved
6supported
0contradicted
100%fully supported
1said about them ↓

Everything Luca Soldaini said on any show that made the record, most notable first. Each card names its show and opens the statement there.

MAD Assertion Not checkable as stated
Soldaini: Most open AI models are open weights, not open source
“Majority of models that get release I think the best term to describe them is open weights. Your Quinn, your Gemma, your Lama you know, Kimi it's what gets release is a set of weights that correspond either to the final state of model, that's the most common, …”
Luca Soldaini Nov 20, 2025 ▶ 10:52 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Insight
Soldaini: AI scaffolding allows people outside frontier labs to drive capabilities
“If the scaffolding is what really moves a lot of like from, you know, broad capability model to like something that actually has meaningful impact, that scaffolding is not just like, oh, only the labs of people are trained models can do it. Like the number of …”
Luca Soldaini Nov 20, 2025 ▶ 1:26:12 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Disclosure
Ai2 releases OLMo 3 with full training recipes, data, and intermediate checkpoints
“We're not just releasing the final models. We're releasing, you know, the entire recipe we followed to get this model. So the data, the intermediate states, the evaluation frameworks, all the details, all the bits that people need to know to make models like O…”
Luca Soldaini Nov 20, 2025 ▶ 1:46 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Assertion Not checkable as stated
Soldaini: Frontier AI labs limit final pre-training runs to two months
“I think it's standard practice among the frontier labs to try to cap your big final pre-training run to two months not more than that.”
Luca Soldaini Nov 20, 2025 ▶ 47:21 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Insight
Soldaini: Flawed long-context model architecture cannot be saved by good data
“But they're like technical decisions in how you set up your model that you can have the best data in the world. And your model will not be able to reason over many, many tokens. So it doesn't matter in the sense that you can't train the model on bad data, but …”
Luca Soldaini Nov 20, 2025 ▶ 53:40 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Disclosure
Ai2 samples 6T tokens from 10T pool for OLMo 3
“There's like a pool of about 10 trillion tokens from which we have like an algorithm also fully open source. To like sample about six trillion tokens that we use during training.”
Luca Soldaini Nov 20, 2025 ▶ 6:07 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Assertion Not checkable as stated
Soldaini: 95% of web pages are under 3,000 tokens
“Like 95% web pages are below 3000 tokens.”
Luca Soldaini Nov 20, 2025 ▶ 7:31 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Assertion Open · timeframe Nov 2028
Ai2 received an initial grant of two million GPU hours from AMD
“We got an initial grant from AMD at the time. There was about two million GPU hours.”
Luca Soldaini Nov 20, 2025 ▶ 26:58 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Disclosure
Ai2 filtered OLMo 3's pre-training dataset from 300 trillion tokens
“Our initial pool was closer to 300 trillion tokens. You shrink it down till you reach your target number, and hopefully as you shrink, you only keep the best part of this.”
Luca Soldaini Nov 20, 2025 ▶ 48:44 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Insight
Soldaini: Mid-training requires re-mixing pre-training data to avoid model forgetting
“When you do that, you also need to make sure that The model doesn't forget stuff that I've seen during pre-training, so that's why, like, you mix some of the best data from pre-training, you do carry over.”
Luca Soldaini Nov 20, 2025 ▶ 50:58 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"
MAD Assertion Supported
Soldaini: Training LLMs on longer sequences causes quadratic compute slowdown
“It's because the longer the input that a model is trained on, the slower it is. The rate at which it gets slower, it's higher than the length of a context. It's a quadratic slowdown.”
Luca Soldaini Nov 20, 2025 ▶ 52:41 Open Source AI Strikes Back — Inside Ai2’s OLMo 3 ‘Thinking"

The other half of the tape: Luca Soldaini's own voice is left out of every number here. Other people bring the name up 1 time in 1 episode across the shows. every mention, with the transcript →

Who brings them up most Shawn Wang 1

Every mention by year

tap a year for its mentions
0011112025episodesmentions
0112025episodes it came up in
000.50.5112025episodesmentions per episode

Latent Space 1

2025 1 mention in 1 episode

One line per show, most statements first. The link opens Luca's full record on that show: the calibration, argument clarity, speaking style and every statement made there.

ShowRole thereEpsStatementsRecord
LATENT SPACELEDGER Member of Technical Staff, Microsoft AI 0 0 100% 5/5 full record on Latent Space →
MADLEDGER Member of Technical Staff, Microsoft AI 1 11 100% 1/1 full record on the MAD Podcast →
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.