multi-agent RL

1 statements across 1 episodes · 1 bullish · 0 bearish · 1 people on the record · first statement May 1, 2025 by Brandon McKinzie · across every show →

Everything said about multi-agent RL, oldest first

May 1, 2025 positive
Insight
McKinzie: Multi-agent RL is a good baseline for human collaboration
“There's no reason you can't scale all this up so that models are trained to be really good at cooperating with each other. I mean, there's a lot of already existing literature on multi-agent RL and yeah, if you want the model to be good at something like colla…”
Brandon McKinzie May 1, 2025 ▶ 26:38 No Priors Ep. 113 | With OpenAI's Eric Mitchell and Brandon McKinzie
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 100 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.