Towards Data Science
Towards Data Science
The TDS team
106. Yang Gao - Sample-efficient AI
49 minutes Posted Dec 8, 2021 at 3:26 pm.
Intro
Yang’s background
MuZero’s activity
MuZero to EfficiantZero
Sample efficiency comparison
Leveraging algorithmic tweaks
Importance of evolution to human brains and AI systems
Human-level sample efficiency
Existential risk from AI in China
Evolution and language
Wrap-up
0:00
49:53
Download MP3
Show notes
Historically, AI systems have been slow learners. For example, a computer vision model often needs to see tens of thousands of hand-written digits before it can tell a 1 apart from a 3. Even game-playing AIs like DeepMind’s AlphaGo, or its more recent descendant MuZero, need far more experience than humans do to master a given game.
So when someone develops an algorithm that can reach human-level performance at anything as fast as a human can, it’s a big deal. And that’s exactly why I asked Yang Gao to join me on this episode of the podcast. Yang is an AI researcher with affiliations at Berkeley and Tsinghua University, who recently co-authored a paper introducing EfficientZero: a reinforcement learning system that learned to play Atari games at the human-level after just two hours of in-game experience. It’s a tremendous breakthrough in sample-efficiency, and a major milestone in the development of more general and flexible AI systems.
--- 
Intro music:
➞ Artist: Ron Gelinas
➞ Track Title: Daybreak Chill Blend (original mix)
➞ Link to Track: https://youtu.be/d8Y2sKIgFWc
---
Chapters: 
-
-
-
-
-
-
-
-
-
-
-