“My prediction is that reg will be much cheaper than long contacts going forward.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Tengyu Ma
PredictionNot checkable as stated
Ma: AI Is Running Out of Compute and Training Data
“My vision is that in the future the efficiency is very important because we are running out of data and compute. So we have to either use the data much better and use the compute much better.”
Tengyu MaJun 6, 2024▶ 1:23No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
AssertionSupported
Ma: Sophia Optimizer Improves LLM Pre-Training Efficiency by 2x
“One of the paper we wrote last year was Sophia which we found, where we found that we have a nutrient optimizer, which can improve the training efficiency by two X for pre-training.”
Tengyu MaJun 6, 2024▶ 2:59No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
AssertionNot checkable as stated
Ma: Meta Achieved 1.6x Training Efficiency Gain Using Sophia Optimizer
“And recently I think one of the Facebook friends actually used this in their large scale multimodal training. And they found that in on that scale, I don't know exactly how many parameters there are, but I think, I assume it's kind of more than a hundred billi…”
Tengyu MaJun 6, 2024▶ 3:54No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
Insight
Ma: Fine-Tuning Often Fails Due to Data Demands and Hallucinations
“Fine tuning in many cases doesn't work because you need a lot of data to see the results and there are still hallucinations even after fine tuning.”
Tengyu MaJun 6, 2024▶ 12:06No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
“So right now, if they have a hundred million tokens, I don't think they can use long context transformers at all because it's way too expensive.”
Tengyu MaJun 6, 2024▶ 17:43No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
PredictionNot checkable as stated
Ma: Iterative Retrieval Will Diminish as Embedding Models Improve
“However, in the long run, my suspicion is that iterative retrieval will be useful, but it will be a bit less useful as the If the embedding models becomes more and more clever, right? So once the embedding models are more clever, then maybe one run or two runs…”
Tengyu MaJun 6, 2024▶ 20:01No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 100 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.