inference-time compute

also referred to as: inference time compute

5 statements across 4 episodes · 3 bullish · 0 bearish · 4 people on the record · first statement Apr 25, 2023 by Noam Brown · across every show →

Everything said about inference-time compute, oldest first

Apr 25, 2023 bullish
Insight
Inference-time compute is the missing scaling dimension for AI reasoning
“This is why I'm interested in the reasoning direction, because I think there's this whole other dimension. That people are not scaling right now, which is the amount of compute at inference time.”
Noam Brown Apr 25, 2023 ▶ 17:12 No Priors Ep. 1 | With Noam Brown, Research Scientist at Meta
Aug 30, 2024 neutral
Insight
Steinberger: AI performance requires trading off training compute against inference compute
“Well, so you can think of model performance as some function of training compute times some function of inference time compute. Now those are specific functions that are just scaling law things that you can like model, but the general Way to think about it is …”
Eric Steinberger Aug 30, 2024 ▶ 7:32 No Priors Ep. 79 | With Magic.dev CEO and Co-Founder Eric Steinberger
Nov 14, 2024 bullish
Opinion
Mehta: AlphaProof's RL scaling and test-time compute generalize across domains
“Some of the sort of tech we developed here of like, you know, like scaling RL and like figuring out how to spend a lot of inference time compute stuff like this feels like it's Quite generally applicable to many other problems.”
Rishi Mehta Nov 14, 2024 ▶ 21:49 No Priors Ep. 90 | With Google's DeepMind's AlphaProof Team
Nov 21, 2024
Assertion Not checkable as stated
Inference-Time Compute Does Not Require Densely Interconnected Supercomputers
“If we have a new avenue, which is inference time compute, That doesn't require this densely interconnected supercomputer. It's fine to have nodes. You can do a lot more locally and less distributed.”
Aidan Gomez Nov 21, 2024 ▶ 29:28 No Priors Ep. 91 | With Cohere Co-Founder and CEO Aidan Gomez
Nov 21, 2024 bullish
Insight
Inference-Time Compute Lets Labs Scale Intelligence Without Doubling Supercomputers
“I don't need to go double the size of my supercomputer to hit a requisite intelligence threshold. I can just double the amount of inference time compute that my customers pay for.”
Aidan Gomez Nov 21, 2024 ▶ 28:12 No Priors Ep. 91 | With Cohere Co-Founder and CEO Aidan Gomez
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.