Lamini

includes Lamini Memory Tuning

12 statements across 2 episodes · 11 bullish · 1 bearish · 1 people on the record · first statement Nov 8, 2023 by Sharon Zhou · said 45 times in 2 episodes since 2023 · across every show →

Mentions by year, the whole family

brought up most by Matt Turck (38), Sharon Zhou (7)

tap a year for its mentions
0013125120232024episodesmentions
01120232024episodes it came up in
00130.525120232024episodesmentions per episode

every mention, scene by scene, with the transcript →

Everything said about Lamini, oldest first

Nov 8, 2023 positive
Disclosure
Lamini's hosted service ran exclusively on AMD GPUs for a year
“The Lamini hosted service over the past year has been running on AMD GPUs only. We haven't been running on NVIDIA chips.”
Sharon Zhou Nov 8, 2023 ▶ 33:01 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 positive
Assertion Open · timeframe Nov 2026
Lamini switches across 1,000 fine-tuned models in three milliseconds
“With the technology that we've used with parameter efficient fine tuning and just like efficiency, different efficiency methods, that time to switch across a thousand models is three milliseconds”
Sharon Zhou Nov 8, 2023 ▶ 27:59 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 bullish
Assertion Not checkable as stated
Lamini has achieved software parity on AMD GPUs with CUDA
“We have reached software parity with essentially CUDA.”
Sharon Zhou Nov 8, 2023 ▶ 37:03 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 bullish
Assertion Contradicted
Lamini's AMD support unlocks 20,000 GPUs, enough to train GPT-4
“So that unlocks, what that means is that unlocks about 20,000 GPUs readily available today for enterprises to be able to use, and to get a sense of what that means you can train GBD-IV.”
Sharon Zhou Nov 8, 2023 ▶ 12:55 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 positive
Assertion Not checkable as stated
Zhou: Lamini cuts LLM fine-tuning time from months to milliseconds
“And by efficiency, I mean, you know, it's instead of something that might take weeks or even months that's bringing it down to even like the millisecond level.”
Sharon Zhou Nov 8, 2023 ▶ 9:56 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 positive
Assertion Contradicted
Zhou: Lamini is the only platform running LLMs on AMD GPUs
“We are the only folks who can actually run your language models on top of AMD AMD GPUs.”
Sharon Zhou Nov 8, 2023 ▶ 12:37 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Nov 8, 2023 positive
Disclosure
Zhou: Lamini can train models up to 100 billion parameters
“We can train up to a hundred billion parameters.”
Sharon Zhou Nov 8, 2023 ▶ 24:34 Custom LLMs at Scale: Lamini CEO Sharon Zhou’s Playbook for Enterprise AI
Jul 25, 2024 bullish
Assertion Not checkable as stated
Memory tuning eliminates hallucinations and enables near-perfect task performance
“Been able with memory tuning, which is what I've been working on to remove those hallucinations, to remove that and actually get these models from, you know, not necessarily being general for everything. And instead of being pretty good at everything, but perf…”
Sharon Zhou Jul 25, 2024 ▶ 10:02 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
Jul 25, 2024 negative
Disclosure
In mid-2023, multi-billion dollar companies could not obtain AWS GPU nodes
“Last year was, at this time, was absolutely insane. That's why we threw up our own cloud, because there was just like, large companies with multi-billion revenue numbers could not get a node from AWS, despite their accounts being tens of millions or hundreds o…”
Sharon Zhou Jul 25, 2024 ▶ 17:06 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
Jul 25, 2024 bullish
Disclosure
Zhou: Lamini deploys enterprise LLMs on-premise in air-gapped environments
“We're an integrated inference and fine tuning platform for enterprises to be able to run factual LLMs. So essentially LLMs that don't hallucinate on their proprietary data within their secure walls. So we can deploy on premise air gapped, no internet sites. So…”
Sharon Zhou Jul 25, 2024 ▶ 10:58 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
Jul 25, 2024 positive
Assertion Supported
Re-engineering LLM decoders can guarantee absolute schema accuracy for structured outputs
“So that's something else we offer through our inference service to actually make it a hundred percent by re-engineering the decoder of any LLM.”
Sharon Zhou Jul 25, 2024 ▶ 25:22 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
Jul 25, 2024 positive
Prediction Not checkable as stated
Future AI models will undergo continuous fine-tuning as easily as prompt engineering
“I believe in a future where we're continuously fine tuning these models where it's as easy as prompt engineering and you know, these models continually improve.”
Sharon Zhou Jul 25, 2024 ▶ 40:32 Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.