LLM Inference Performance
topic on 1 show · 1 statements across 1 episodes
1 statements about LLM Inference Performance, every show
Srivastava: LLM inference performance will commoditize and increasingly run locally
“We think over time, It will get somewhat commoditized, the performance, especially, especially for language models, to be honest. I think, you know, more and more of that stuff should run locally to some degree, I think”