The Ledger, every show

Every statement that passed quotation and attribution checks, across all 44 shows. Pick shows below, then mix any filter with any other.

shows every show 44 of 44
every show
clear all ✕
MAD Assertion Not checkable as stated
Memory tuning eliminates hallucinations and enables near-perfect task performance
“Been able with memory tuning, which is what I've been working on to remove those hallucinations, to remove that and actually get these models from, you know, not necessarily being general for everything. And instead of being pretty good at everything, but perf…”
Sharon Zhou Jul 25, 2024 ▶ 10:02 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Assertion Supported
The enterprise GPU shortage has eased at the company level
“Today, actually, I'm seeing the GPU shortage go away at the level, at the company level, meaning companies are able to procure enough compute enough is a strong word, but they're able to procure compute at some level to work with, to fine tune and run heavy in…”
Sharon Zhou Jul 25, 2024 ▶ 16:01 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Assertion Supported
Re-engineering LLM decoders can guarantee absolute schema accuracy for structured outputs
“So that's something else we offer through our inference service to actually make it a hundred percent by re-engineering the decoder of any LLM.”
Sharon Zhou Jul 25, 2024 ▶ 25:22 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Prediction Not checkable as stated
Future AI models will deliver 100B parameter intelligence at 1B speeds
“I even think there's a future where these models can be a hundred billion parameters, but at, you know, have that intelligence of a hundred billion parameters, but then have the speed, latency, and cost of something that's still one billion or seven billion pa…”
Sharon Zhou Jul 25, 2024 ▶ 28:15 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Prediction Not checkable as stated
Combining MoE and LoRA will eliminate big versus small model trade-offs
“And I do think that's the future so we can get something that is incredibly smart, incredibly huge, but with the latency cost and speed of something, something tiny. So no more big model versus small model paradigm. It's potentially one in the same.”
Sharon Zhou Jul 25, 2024 ▶ 32:24 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
SAASTR Assertion Not checkable as stated
Zhou: Public training data for LLMs is running out
“Yeah, I think even just zooming back out at a technical level, I think actually public data is running out for all, all of what LLMs can take advantage of.”
Sharon Zhou Dec 8, 2023 ▶ 4:38 The Where, When, and How of AI with Theory Ventures, Open AI, MotherDuck and Lamini
SAASTR Prediction Not checkable as stated
Zhou: Domain experts, not AI researchers, will drive top models
“However, I believe that, and this is based on my experience training these models, it's actually the domain experts will be driving the best models out there. It won't be people like me who can actually do all the model training, et cetera.”
Sharon Zhou Dec 8, 2023 ▶ 6:23 The Where, When, and How of AI with Theory Ventures, Open AI, MotherDuck and Lamini
MAD Assertion Not checkable as stated
Zhou: Lamini cuts LLM fine-tuning time from months to milliseconds
“And by efficiency, I mean, you know, it's instead of something that might take weeks or even months that's bringing it down to even like the millisecond level.”
Sharon Zhou Nov 8, 2023 ▶ 9:56 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Contradicted
Zhou: Lamini is the only platform running LLMs on AMD GPUs
“We are the only folks who can actually run your language models on top of AMD AMD GPUs.”
Sharon Zhou Nov 8, 2023 ▶ 12:37 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Contradicted
Lamini's AMD support unlocks 20,000 GPUs, enough to train GPT-4
“So that unlocks, what that means is that unlocks about 20,000 GPUs readily available today for enterprises to be able to use, and to get a sense of what that means you can train GBD-IV.”
Sharon Zhou Nov 8, 2023 ▶ 12:55 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Not checkable as stated
Zhou: Outsourcing specialized medical AI data labeling to Scale AI failed
“We tried outsourcing actually to like scale AI, et cetera. None of that worked. It had to basically be me.”
Sharon Zhou Nov 8, 2023 ▶ 21:38 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Not checkable as stated
Zhou: LLMs can reach 99% accuracy today with narrow scoping
“I think we can get to that performance today, but it's based on how you scope out the problem. So if it's a very narrow scope, of course you can get that.”
Sharon Zhou Nov 8, 2023 ▶ 22:33 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Not checkable as stated
Lamini has achieved software parity on AMD GPUs with CUDA
“We have reached software parity with essentially CUDA.”
Sharon Zhou Nov 8, 2023 ▶ 37:03 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Partly supported
NVIDIA A100s and AMD MI300s are readily available, but H100s remain scarce
“A 100 in particular are pretty available. Obviously the AMD chips that we also agnostically work with the MI 300 and MI two fifties, those are available. H 100 still kind of. A little bit harder to get, but you can get started very easily with any of those oth…”
Sharon Zhou Jul 25, 2024 ▶ 17:47 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
SAASTR Prediction Not checkable as stated
Zhou: Best LLMs of the next wave will be enterprise models
“So actually the next frontier for LLMs is in enterprises, and I believe the best LLMs for this next, next wave essentially will be enterprise LLMs.”
Sharon Zhou Dec 8, 2023 ▶ 4:49 The Where, When, and How of AI with Theory Ventures, Open AI, MotherDuck and Lamini
MAD Assertion Open · timeframe Nov 2026
Lamini switches across 1,000 fine-tuned models in three milliseconds
“With the technology that we've used with parameter efficient fine tuning and just like efficiency, different efficiency methods, that time to switch across a thousand models is three milliseconds”
Sharon Zhou Nov 8, 2023 ▶ 27:59 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
MAD Assertion Partly supported
Fine-tuning a GPT-3 class model with LoRA provides a 10,000x efficiency boost
“I think for something like a GPD three level model, it's a 10,000 X speed up in efficiency while losing nearly not much at all in accuracy.”
Sharon Zhou Jul 25, 2024 ▶ 30:19 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Prediction Not checkable as stated
Future AI models will undergo continuous fine-tuning as easily as prompt engineering
“I believe in a future where we're continuously fine tuning these models where it's as easy as prompt engineering and you know, these models continually improve.”
Sharon Zhou Jul 25, 2024 ▶ 40:32 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
MAD Assertion Supported
Zhou: Fine-tuning transformed GPT-3 into ChatGPT
“Fine tuning is the technology that got from a research project in 2020 called GPT-III and turned that into ChatGPT, a billion dollar app, right?”
Sharon Zhou Nov 8, 2023 ▶ 4:30 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
SAASTR Assertion Not checkable as stated
Zhou: Adding images drove a 10x increase in marketing engagement
“I think one thing that, you know, we were analyzing our data when it came to all our marketing blog posts or like tweets, everything going out, and there was a 10 X increase in anything with an image, right?”
Sharon Zhou Dec 8, 2023 ▶ 23:02 The Where, When, and How of AI with Theory Ventures, Open AI, MotherDuck and Lamini
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.