Synthetic Data

topic on 18 shows · 79 statements across 63 episodes

the Y Combinator Startup Podcast the Pitch Mixergy More or Less BG2 Pod Latent Space Lenny's Podcast My First Million the Neon Show No Priors A Product Market Fit Show Catalyst the MAD Podcast the a16z Podcast Big Technology All-In TBPN 20VC

The latest 60 statements about Synthetic Data, every show

NEON SHOW Insight
Krishnan: Purely model-generated synthetic data hits a wall without human feedback
“There is also, I think, limits to how much value can be added there because ultimately there is just not too much new information. If you're telling the model itself to kind of generate, yeah, There is a, I mean, there's all sorts of things about how beyond th…”
Vijay Krishnan Jul 31, 2026 ▶ 44:14 The Man Training GPT, Gemini & Claude Reveals What's Coming Next | Vijay Krishnan, Turing
NEON SHOW Insight
Krishnan: Human-designed prompt and auto-verifier tuples maximize synthetic training data ROI
“The, this method is the one, I think, which has a lot of legs in the, particularly the more you can operate in this particular paradigm of prompt and then these rule or rubric based verifier tuples. The, that is a very nice way for sort of creating synthetic d…”
Vijay Krishnan Jul 31, 2026 ▶ 48:19 The Man Training GPT, Gemini & Claude Reveals What's Coming Next | Vijay Krishnan, Turing
NEON SHOW Insight
Ambati: Monthly AI model churn is fueled by synthetic data loops
“Every month, if you will, there is a new model, right? That's beating the old model, and visibly adopted already, and people are moving to the new models and the data is just flowing, like you're generating data, you're sort of getting access to new data creat…”
Vamshi Ambati Jun 9, 2026 ▶ 5:24 94% CAGR: What the Inference Boom means for your AI costs | Vamshi Ambati
Hong: Accumulating synthetic AI data is not a moat, just buffer
“I think everyone is trying to accumulate like a data, which is not a mode. It's just time and time mode. It's all about like, you know, whether you can execute fast enough to make sure that you have like a certain buffer because of say your data set, you know,…”
Carina Hong Jun 3, 2026 ▶ 1:03:05 Scaling Past Informal AI - Carina Hong, Axiom Math
Ethan He: Video models require image foundations and 100% synthetic caption pairs
“Building a video model. You actually need to build a image model first and building, building these two models. The data you need is a hundred percent synthetic pair of language and image or language to video because on the internet, actually the videos Don't …”
Ethan He Jun 1, 2026 ▶ 11:55 Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He
LATENT SPACE Assertion Supported
Sun: Synthetic data matches real-world data for multimodal model pre-training
“We were actually generating a lot of synthetic data and showing that, hey, you can actually, these synthetic data are actually as useful as real-world data when it comes to multimodal pre-training.”
Fan-yun Sun Apr 2, 2026 ▶ 2:56 Moonlake: Interactive, Multimodal World Models — with Chris Manning and Fan-yun Sun
ALL-IN Insight
Mensch: AI synthetic data is efficient but cannot replace human training signal
“It's mostly an efficient way of training models to have bigger models that are used as teachers for smaller models, but it's not enough. And so you also need human signal.”
Arthur Mensch Mar 23, 2026 ▶ 1:14:17 Four CEOs on the Future of AI: CoreWeave, Perplexity, Mistral, and IREN
LENNY'S PODCAST Prediction Not checkable as stated
Patel: Companies will differentiate AI models via proprietary, synthetic, and machine data
“And every company is going to differentiate based on their own proprietary enterprise data being used to train the models, synthetic data and machine data, which is where the most amount of growth is.”
Jeetu Patel Feb 26, 2026 ▶ 17:11 AI is critical for humanity’s survival: Cisco President on the AI revolution | Jeetu Patel
20VC Opinion
Fitzpatrick: Belief that synthetic data replaces human feedback is wrong
“Look, I think the biggest one is just the view that synthetic data will take over, and you just will not need human feedback.”
Matt Fitzpatrick Dec 31, 2025 ▶ 0:39 Matt Fitzpatrick: Who Wins the Data Labelling Race & Why Al Needs Forward-Deployed Engineers · 20VC with Harry Stebbings
MIXERGY Insight
Santos: Dream Stories trains consistent models using synthetic photo variations
“We were able to reduce the number of pictures required to literally just one, because then we're like, holy , you know, we can essentially just ask for one picture, generate the synthetic version of it, and then say, do you like this? Great. So we generate a f…”
Ricardo Vice Santos Dec 29, 2025 ▶ 54:13 #2290 He’s building an AI media empire
BIG TECHNOLOGY Assertion Not checkable as stated
Suleyman: AI Training Is Not Data-Constrained Due To Synthetic Data
“We are not data constrained right now, we're generating vast amounts of high quality synthetic data, which is proving to be useful.”
Mustafa Suleyman Nov 12, 2025 ▶ 8:53 Could LLMs Be The Route To Superintelligence? — With Mustafa Suleyman
BIG TECHNOLOGY Prediction Not checkable as stated
Suleyman: Synthetic Data And Human Feedback Will Outweigh Robotics Data
“I don't think in the next few years it's gonna be the big differentiator. I think that more synthetic data, more human feedback and high quality data is gonna be the differentiator.”
Mustafa Suleyman Nov 12, 2025 ▶ 16:49 Could LLMs Be The Route To Superintelligence? — With Mustafa Suleyman
CATALYST Insight
Chubuk: Minimal physical experiments carry huge information value by validating synthetic simulations
“What's interesting about scientific data is it's not just a few bits or numbers, right? Like, for example, there are certain experiments you can run where the result you get from it is just, say, three floating point numbers. But the implications of those coul…”
Doge Chubuk Nov 6, 2025 ▶ 28:37 Inside a $300 million bet on AI for physical R&D
20VC Assertion Supported
Pineau: Synthetic image and language data causes model degradation
“So in some domains, if you think like images, languages, like LLMs talking to each other at some point, you definitely get the degradation and that degradation is due to essentially like a loss of diversity of your data.”
Joelle Pineau Nov 3, 2025 ▶ 37:01 Cohere's Chief AI Officer, Joelle Pineau: Why Scaling Laws Will Continue & Future of Synthetic Data · 20VC with Harry Stebbings
Joseph: Training purely on raw LLM generations cannot produce a better model
“Theoretically, I shouldn't be able to train a better model than that. Like, I'm just going to get the same thing out. So I think that's-”
Nick Joseph Sep 30, 2025 ▶ 34:42 Anthropic Head of Pretraining on Scaling Laws, Compute, and the Future of AI · Y Combinator
MAD Disclosure
Valenzuela: Runway is exploring synthetic data for AI video training
“It is. I think it's becoming more of Thing, I would say. It still has its challenges, mostly to generate diversity of data, but definitely something we're exploring.”
Cris Valenzuela Sep 4, 2025 ▶ 46:05 AI Video’s Wild Year – Runway CEO on What’s Next
Morcos: Filtering synthetic data between generation cycles prevents model collapse
“If you filter the data after each point, that's now information injection, and that can break all of this and I think can prevent model collapse.”
Ari Morcos Aug 29, 2025 ▶ 43:47 Better Data is All You Need — Ari Morcos, Datology
LENNY'S PODCAST Prediction Not checkable as stated
Lord: Synthetic data will not dominate frontier AI training
“Synthetic data has a role to play and like in verifiable domains, but like what we consistently hear from companies is like, you know, their synthetic data is not going to dominate.”
Garrett Lord Aug 24, 2025 ▶ 1:02:01 Inside the expert network training every frontier AI model | Garrett Lord
TBPN Opinion
Cuban: Synthetic AI data cannot invent what human scientists discover
“Cause there's no way to synthesize all that shit. You're not synthesizing. You're not creating synthetic data that all of a sudden is going to, you know, you could tell, you give it all the backstory you want, that you're a doctor, that you invented this, that…”
Mark Cuban Aug 21, 2025 ▶ 15:28 Mark Cuban on AI, Career Advice, and Running for President.
a16z Insight
Fulford: OpenAI bootstraps browsing models to generate synthetic training data
“For initial deep research, there's not really any data sets that exist for browsing in the same way that you have a math data set that already exists. So we have to create all this data. But once you have good browsing models or good computer use models, you c…”
Isa Fulford Aug 8, 2025 ▶ 31:30 GPT-5 and Agents Breakdown – w/ OpenAI Researchers Isa Fulford & Christina Kim
LATENT SPACE Prediction Held up
Lambert: Labs will surely use parallel-compute models to generate synthetic data
“Well, I bet people, I mean, they surely will use these for synthetic data. It's just like the marginal gain on synthetic data is always very high.”
Nathan Lambert Jul 31, 2025 ▶ 50:22 The RLVR Revolution — with Nathan Lambert (AI2, Interconnects.ai)
NO PRIORS Insight
Chen: AI RL Environments Are Too Complex to Create Synthetically
“I think one of the things that people really underestimate is how it is, how complicated it is that you can't just synthetically generate it.”
Edwin Chen Jul 24, 2025 ▶ 14:20 No Priors Ep. 124 | With SurgeAI Founder and CEO Edwin Chen
NO PRIORS Disclosure
Chen: Surge AI Heavily Uses Synthetic Data to Supplement Human Labelers
“Like we use it like a ton ourselves in order to supplement what the humans do.”
Edwin Chen Jul 24, 2025 ▶ 18:05 No Priors Ep. 124 | With SurgeAI Founder and CEO Edwin Chen
Y COMBINATOR Prediction Not checkable as stated
Chelsea Finn: Real robot data cannot be replaced by synthetic data
“I think that at the end of the day, there's going to be no replacement for real data. And so we're like large amounts of real robot data. It's going to be a necessary component of any like system that's going to work in a generalizable way.”
Chelsea Finn Jul 22, 2025 ▶ 41:05 Chelsea Finn: Building Robots That Can Do Anything · Y Combinator
Chelsea Finn: Synthetic data's robotic analog is RL, not simulation
“I think that the analog of synthetic data in language models is actually not necessarily simulation in robotics, but closer to something like reinforcement learning.”
Chelsea Finn Jul 22, 2025 ▶ 41:47 Chelsea Finn: Building Robots That Can Do Anything · Y Combinator
20VC Insight
Edwin Chen: Synthetic data makes AI models good at benchmarks, not real problems
“Synthetic data, it's made models good at synthetic problems, not, not real ones.”
Edwin Chen Jul 21, 2025 ▶ 50:42 Surge CEO & Co-Founder, Edwin Chen: Scaling to $1BN+ in Revenue with NO Funding · 20VC with Harry Stebbings
20VC Assertion Not checkable as stated
Surge AI CEO: 2,000 human data points beat 10 million synthetic ones
“A lot of them tell us that even a thousand or a couple of thousand pieces of really high quality human data that we generated for them, it's actually been worth more than ten million pieces of synthetic data.”
Edwin Chen Jul 21, 2025 ▶ 50:58 Surge CEO & Co-Founder, Edwin Chen: Scaling to $1BN+ in Revenue with NO Funding · 20VC with Harry Stebbings
a16z Opinion
Bornstein: Synthetic data will not produce self-improving AI models
“The question is, like, does this lead to sort of, like, a self-improving utopia of models or not? And I think we have some pretty strong opinions on the not side of that.”
Matt Bornstein Jul 21, 2025 ▶ 41:48 The Future of Software Development - Vibe Coding, Prompt Engineering & AI Assistants
NO PRIORS Insight
Laskin: Reinforcement learning is the only scalable path for synthetic data
“When we're generating synthetic data there is the only scalable path is really reinforcement learning.”
Misha Laskin Jul 17, 2025 ▶ 53:44 No Priors Ep. 123 | With ReflectionAI Co-Founder and CEO Misha Laskin
MORE OR LESS Assertion Not checkable as stated
Morin: Early research proves synthetic data cannot effectively train AI models
“Yeah, I mean, that goes back to that first research paper that came out right after ChatGPT launched, which basically said you can't create Fake data for training models. It just doesn't work, and so there was, like, this very early paper that effectively told…”
Dave Morin Jun 13, 2025 ▶ 38:58 #103: Are Silicon Valley’s Mega-Deals Just Talent Grabs? · More or Less Podcast
MAD Disclosure
Gomez: Synthetic data makes up the majority of Cohere's training data
“Synthetic data is incredibly effective. It's now the majority of the data that we train on for creating something like command A.”
Aidan Gomez Jun 5, 2025 ▶ 38:11 Inside the Paper That Changed AI Forever - Cohere CEO Aidan Gomez on 2025 Agents
Patel: Human labeling is unscalable, forcing reliance on synthetic AI data
“Using humans to train models is just so expensive, right? So then there's the magic of sort of reinforcement learning and other synthetic data technologies, right? Where the model is helping teach the model, right? So you have many models in, in, in a sort of,…”
Dylan Patel Apr 23, 2025 ▶ 18:14 Generative AI 101: Tokens, Pre-training, Fine-tuning, Reasoning — With SemiAnalysis CEO Dylan Patel
20VC Prediction Open · timeframe Mar 2030
Feldman: In five years, almost all AI training data will be synthetic
“Almost all synthetic.”
Andrew Feldman Mar 24, 2025 ▶ 30:18 Andrew Feldman, Cerebras Co-Founder and CEO: The AI Chip Wars & The Plan to Break Nvidia's Dominance · 20VC with Harry Stebbings
MAD Insight
Kiela: DeepSeek proved frontier AI models can rely on synthetic data
“We have kind of an existence proof now that it's actually not that hard to do this and so you don't need to invest all that much in, in data, and you can use synthetic data and get a pretty good model out of that”
Douwe Kiela Mar 6, 2025 ▶ 3:12 Top AI Researcher on GPT 4.5, DeepSeek and Agentic RAG | Douwe Kiela, CEO, Contextual AI
20VC Insight
Ross: LLM-generated synthetic data is better for model training
“You could have an LLM generate synthetic data, and when it generates the synthetic data, the data is better. You then train on that synthetic data.”
Jonathan Ross Feb 17, 2025 ▶ 3:09 Jonathan Ross, Founder & CEO @ Groq: NVIDIA vs Groq - The Future of Training vs Inference | E1260 · 20VC with Harry Stebbings
Nguyen: OpenAI built Canvas and Tasks features mostly via synthetic data
“The way we made Canvas and tasks and, like, new, like, product features for HTTP was mostly done by synthetic training.”
Karina Nguyen Feb 9, 2025 ▶ 12:27 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Nguyen: Synthetic Data Outperforms Human Data for AI Product Development
“And the reason why I really love, like, synthetic, like, relying purely on synthetic data instead of, like, collecting Data from humans is because it's, like, much more scalable. It's cheap, less than how, like, you literally sample from the model, and you tea…”
Karina Nguyen Feb 9, 2025 ▶ 31:55 OpenAI researcher on why soft skills are the future of work | Karina Nguyen
Palafox: Autonomous vehicle companies will not outsource synthetic data to startups
“Like there's just not enough enterprises that need that. And the ones that actually need it, it's so core. Like think think like Cruz or like Waymo. It's core to them. Like they're not gonna like outsource that to like a shitty startup.”
Pablo Palafox Jan 30, 2025 ▶ 9:37 1st time founder completely pivots after YC—then grows 30x in a year to $2.2M ARR. | Pablo Palafo... · PMF Show
LATENT SPACE Prediction Not checkable as stated
Ben Allal: Properly curated synthetic data prevents model collapse
“And I think there's a lot of concerns about model collapse, and I'm going to talk about that later, but we'll see that like, if we use synthetic data properly and we curate it carefully that shouldn't happen.”
Loubna Ben Allal Dec 24, 2024 ▶ 2:09 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
LATENT SPACE Assertion Supported
Ben Allal: Recent web dumps improve model benchmarks despite synthetic data
“So what we did is we trained different models on these different dumps, and we then computed their performance on popular like NLP benchmarks, and then we computed the aggregated score. And surprisingly, you can see that the latest dumps are actually even bett…”
Loubna Ben Allal Dec 24, 2024 ▶ 4:12 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
Ben Allal: Synthetic data may enrich the web rather than pollute it
“So personally, I wouldn't say the web is posted with synthetic data. Maybe it's even making it more rich.”
Loubna Ben Allal Dec 24, 2024 ▶ 4:35 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
Ben Allal: Prompt diversity is essential for scaling synthetic training data
“The key ingredient to getting a good data set that is synthetic is trying as much as possible to keep it diverse, because if you just throw the same prompts as your model, like generate, like, a textbook about linear algebra, and even if you change the tempera…”
Loubna Ben Allal Dec 24, 2024 ▶ 6:08 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
LATENT SPACE Assertion Supported
Synthetic college text boosts MMLU; middle school text boosts OpenBookQA
“College textbooks are really good for MLU or middle school textbooks are good for benchmarks like open book, UA and Pico.”
Loubna Ben Allal Dec 24, 2024 ▶ 7:44 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
Ben Allal: Small models can generate synthetic data by rephrasing web pages
“The interesting thing in this approach is that you can use a model that is small Because it doesn't, rewriting doesn't require knowledge. It's just rewriting a page into a different style. So the model doesn't need to have like knowledge that is like extensive…”
Loubna Ben Allal Dec 24, 2024 ▶ 9:39 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
LATENT SPACE Assertion Supported
Ben Allal: Pre-training on rewritten C4 web data outperforms raw C4
“They rewrite some samples from C four into Q and A into Wikipedia, and they find that doing this works better than training just on C four.”
Loubna Ben Allal Dec 24, 2024 ▶ 9:59 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
Ben Allal: Pooling multiple teacher models produces superior synthetic datasets
“Synthetic data, it doesn't have to come from a single model. And because we have so many good models now, you could like pull these models together and get like a dataset that's over really high quality and that's diverse and that's covers all your needs.”
Loubna Ben Allal Dec 24, 2024 ▶ 16:26 Best of 2024: Synthetic Data / Smol Models, Loubna Ben Allal, HuggingFace [LS Live! @ NeurIPS 2024]
BG2 Insight
Patel: Synthetic data generation enables continued AI scaling despite data limits
“You can create data out of thin air almost, right? In certain domains, right? And so this is the whole, the debate around scaling laws is how can we create data?”
Dylan Patel Dec 23, 2024 ▶ 25:25 AI Semiconductor Landscape feat. Dylan Patel | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
BG2 Assertion Not checkable as stated
Patel: AI industry is in early days of synthetic data
“Where have we gone on synthetic data? Oh, we're still like very early days, right? We've spent tens of millions of dollars maybe on synthetic data.”
Dylan Patel Dec 23, 2024 ▶ 28:08 AI Semiconductor Landscape feat. Dylan Patel | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
BG2 Insight
Patel: Synthetic training only works in functionally verifiable domains like math
“We can't teach it what good art is. Because we have no way to functionally prove what good art is. We can teach it to write really good software. We can teach it how to do mathematical proofs. We can teach it how to engineer systems, because there are, while t…”
Dylan Patel Dec 23, 2024 ▶ 28:58 AI Semiconductor Landscape feat. Dylan Patel | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
20VC Assertion Not checkable as stated
Hoffman: LLM scaling is not out of data thanks to synthetic data
“It's like, well, actually, in fact, we can create synthetic data, and there's a ton of data that's out there that's not part of the standard internet training corpus. So, the scale game is still playing.”
Reid Hoffman Dec 16, 2024 ▶ 49:57 Reid Hoffman, LinkedIn & Paypal Founder: Trump Administration, Elon Musk and DOGE | E1239 · 20VC with Harry Stebbings
BIG TECHNOLOGY Assertion Not checkable as stated
Wang: Pure synthetic AI training data has underperformed industry expectations
“One of the things that we've seen over the past few past year in particular is that synthetic data has not worked as well as I think everybody had hoped. You know, pure synthetic data, just using data generated from the models to try to train future models, th…”
Alexandr Wang Dec 11, 2024 ▶ 44:46 AI Predictions for 2025: Geopolitics, Agents, and Data Scaling — With Alexandr Wang
Pre-training on human data is hitting limits; synthetic data is required
“So I think on the data side, we're approaching the limit and the only data to increase that is synthetic generated data.”
Lin Qiao Nov 25, 2024 ▶ 35:05 Why Compound AI + Open Source will beat Closed AI — with Lin Qiao, CEO of Fireworks AI
Gomez: Synthetic data struggles outside verifiable domains like math
“Synthetic data Probably doesn't get us out of that, that issue. I actually, I don't know if synthetic data outside of easily verifiable domains like math, it's hard to use synthetic data to drive outcomes.”
Aidan Gomez Oct 30, 2024 ▶ 26:51 The Next Gen AI Models: Reliable, Consistent, Trustworthy — With Cohere CEO Aidan Gomez
BIG TECHNOLOGY Disclosure
Gomez: Synthetic Data Makes Up a Growing Portion of Cohere's Training
“More and more synthetic data is becoming a huge chunk of the data that we train on.”
Aidan Gomez Oct 30, 2024 ▶ 27:54 The Next Gen AI Models: Reliable, Consistent, Trustworthy — With Cohere CEO Aidan Gomez
LATENT SPACE Assertion Supported
Most Open Vision Models Rely on Synthetic Data From Proprietary Models
“Most VLMs are distillations of proprietary closed source models, right? So if you need to generate synthetic data, like most open weight models rely heavily on synthetic data from private models.”
Vibhu Sapra Oct 13, 2024 ▶ 1:13 [Paper Club] Molmo + Pixmo + Whisper 3 Turbo - with Vibhu Sapra, Nathan Lambert, Amgadoz
20VC Insight
Kant: Synthetic data only improves AI when evaluated by an objective oracle
“If you have something that can determine an oracle of truth that can help say, this is better and this is worse, or this is correct and this is wrong, that's when you can actually use synthetic data.”
Eiso Kant Oct 7, 2024 ▶ 13:43 Eiso Kant, CTO @Poolside: Raising $600M To Compete in the Race for AGI | E1211 · 20VC with Harry Stebbings
a16z Assertion Not checkable as stated
Jeff Schmidt: Hermes pioneered synthetic data training before it was standard
“So Hermes was very early to the idea that you could have synthetic data, which is that you could actually make, you could make a better model by taking an AI model, having it generate words and text, and then training a new a model on that output. This is now …”
Jeff Schmidt Oct 1, 2024 ▶ 17:58 The Quest for Community-Trained Open Source AI Models
NO PRIORS Prediction Not checkable as stated
Karpathy: AI will not run out of training data due to synthetic data
“So I think basically synthetic data is absolutely the future. We're not going to run out of data, is my impression. I just think you have to be careful.”
Andrej Karpathy Sep 5, 2024 ▶ 20:21 No Priors Ep. 80 | With Andrej Karpathy from OpenAI and Tesla
THE PITCH Opinion
Desai: Synthetic healthcare data fails to mimic true patients
“A lot of solutions out there, especially in healthcare, use synthetic data. Synthetic data does not mimic a true patient.”
Davina Desai Sep 4, 2024 ▶ 13:03 Will VCs Bet $1.5M on This AI That Prevents Hospital Falls?
ALL-IN Prediction Not checkable as stated
Synthetic data will make individual media training data irrelevant
“Don't try to hold out for money on the training side of things, because, you know, we're going to create synthetic data, we're going to do all kinds of other things that are going to mean that no one's particular data is really going to matter.”
Reid Hoffman Aug 30, 2024 ▶ 28:10 In conversation with Reid Hoffman & Robert F. Kennedy Jr.

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.