SGLang, every mention
1 scene · ← back to SGLang
tap a year for its mentions
Verbatim, from the transcripts: the passages where SGLang comes up
Multi-GPU Kernels, Intelligence per Watt, Heterogeneous Inference, and More | YC Paper Club · Y Combinator
- ▶ 45:47 unnamed speaker Um, similarly, if you want to deploy these kernels in a real inference engine, well, inference engines, like, let's say, out of the box, if you were to use something like SGLang or VLM, if it's just loading DeepSeq before the first…