Assertion Partly supported AI assessment confidence: 85% certainty 4/5 debate potential 1/5

Zaremba: RL models need three years of real-time play to learn games

Wojciech Zaremba · An AI Primer with Wojciech Zaremba · Y Combinator · May 17, 2017 · at 6:24

OpenAI cofounder Wojciech Zaremba explains the sample efficiency gap between deep reinforcement learning models and human learners playing simple video games.

0:00 / 0:18exact quote · 18.1s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“Well, for instance, in terms of real-time execution, it takes something around three, three, three years of play to learn to play simple games . I mean it, it can be hugely parallelized, therefore it takes a few days to train it on current computers”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Wojciech Zaremba

Opinion
Zaremba: Universal basic income is the only viable path as automation grows
“So, I believe that we'll have to offer to people a basic income. I super strongly believe that actually that's the only way.”
Wojciech Zaremba May 17, 2017 ▶ 52:06 An AI Primer with Wojciech Zaremba · Y Combinator
Insight
Zaremba: Most business problems can be solved via supervised learning
“Majority of business problems can be framed as supervised learning, and therefore they can be solved with current techniques, as long as we have sufficient number of input examples and what we want to predict”
Wojciech Zaremba May 17, 2017 ▶ 46:59 An AI Primer with Wojciech Zaremba · Y Combinator
Assertion Partly supported
Zaremba: OpenAI has secured $1 billion in total investment
“In total we gather an investment of one billion dollar in the group.”
Wojciech Zaremba May 17, 2017 ▶ 1:34 An AI Primer with Wojciech Zaremba · Y Combinator
Insight
Zaremba: Reinforcement learning struggles in reality due to reward and reset assumptions
“The assumption underlying reinforcement learning Is that then there is some environment, and environment, you are an agent, and you are acting in environment by executing actions and getting rewards from the environment. And the rewards might be taught as, let…”
Wojciech Zaremba May 17, 2017 ▶ 7:12 An AI Primer with Wojciech Zaremba · Y Combinator
Opinion
Zaremba: ImageNet is the essential dataset that enabled deep learning
“That's the essential data set that made deep learning happen.”
Wojciech Zaremba May 17, 2017 ▶ 34:36 An AI Primer with Wojciech Zaremba · Y Combinator
Assertion Supported
Zaremba: ImageNet error dropped to 3%, achieving superhuman vision performance
“Within several years, people got down, I believe, to three percent error, and that's essentially superhuman performance.”
Wojciech Zaremba May 17, 2017 ▶ 37:40 An AI Primer with Wojciech Zaremba · Y Combinator
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.