“So we produce embeddings that is like a three X, you know, four X smaller dimension than some of the competitors.”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Tengyu Ma
PredictionHeld up
Ma: RAG Will Remain Much Cheaper Than Long-Context Windows
“My prediction is that reg will be much cheaper than long contacts going forward.”
Tengyu MaJun 6, 2024▶ 14:11No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
PredictionNot checkable as stated
Ma: AI Is Running Out of Compute and Training Data
“My vision is that in the future the efficiency is very important because we are running out of data and compute. So we have to either use the data much better and use the compute much better.”
Tengyu MaJun 6, 2024▶ 1:23No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
AssertionSupported
Ma: Sophia Optimizer Improves LLM Pre-Training Efficiency by 2x
“One of the paper we wrote last year was Sophia which we found, where we found that we have a nutrient optimizer, which can improve the training efficiency by two X for pre-training.”
Tengyu MaJun 6, 2024▶ 2:59No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
AssertionNot checkable as stated
Ma: Meta Achieved 1.6x Training Efficiency Gain Using Sophia Optimizer
“And recently I think one of the Facebook friends actually used this in their large scale multimodal training. And they found that in on that scale, I don't know exactly how many parameters there are, but I think, I assume it's kind of more than a hundred billi…”
Tengyu MaJun 6, 2024▶ 3:54No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
Insight
Ma: Fine-Tuning Often Fails Due to Data Demands and Hallucinations
“Fine tuning in many cases doesn't work because you need a lot of data to see the results and there are still hallucinations even after fine tuning.”
Tengyu MaJun 6, 2024▶ 12:06No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
“So right now, if they have a hundred million tokens, I don't think they can use long context transformers at all because it's way too expensive.”
Tengyu MaJun 6, 2024▶ 17:43No Priors Ep. 67 | With Voyage AI Co-Founder and CEO
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 100 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.