Google DeepMind's Deep Q-Learning & Superhuman Atari Gameplays | Two Minute Papers #27
Skills:
RL Foundations90%
Key Takeaways
Explains Google DeepMind's Deep Q-Learning technique for playing Atari games at a superhuman level
Full Transcript
dear fellow Scholars this is 2minute papers with Caro here this one is going to be huge certainly one of my favorites this work is a combination of several techniques that we have talked about earlier if you don't know some of these terms it's perfectly okay you can remedy this by clicking on the popups or checking the description box but you'll get the idea even watching only this episode so first we have a convolutional neural network this helps processing images and understanding what is depicted on an image and a reinforcement learning algorithm this helps creating strategies or to be more exact it decides what the next action we make should be what buttons we push on a joystick so this technique mixes together these two concepts and we call it deep Q learning and it is able to learn to play games the same way as a human would it is not exposed to any additional information in the code all it sees is the screen and the current score when it starts learning to play an old game Atari Breakout at first the algorithm loses all of its lives without any signs of intelligent action if we wait a bit it becomes better at playing the game roughly matching the skill level of an adapt player but here's the C catch if we wait for longer we get something absolutely spectacular it finds out that the best way to win the game is digging a tunnel through the bricks and hit them from behind I really didn't know this and this is an incredible moment I can use my computer this box next to me that is able to create new knowledge find out new things I haven't known before this is completely absurd science fiction is not the future it is already here it also plays many other games the percentages show the relation of the game scores compared to a human player above 70% means it's great and above 100% it's superum as a follow-up work scientists at Deep Mind started experimenting with 3D games and after a few days of training it could learn to drive on idea racing lines and pass others with ease I've had my driving license for a while now but I still don't always get the idea racing lines right Bravo I have heard the complaint that this is not real intelligence because it doesn't know the concept of a ball or what it is exactly doing edar dyra once said the question of whether machines can think is about as relevant as the question of whether submarines can swim beyond the fact that rigorously defining intelligence leans more into the domain of philosophy than science I I'd like to add that I am perfectly happy with effective algorithms we use these techniques to accomplish different tasks and they are really good problem solvers in the Breakout game you as a person learn the concept of a ball in order to be able to use this knowledge as a Machinery to perform better if this is not the case whoever knows a lot but can't use it to achieve anything useful is not an intelligent being but an encyclopedia what about the future there are two major unexplored directions the algorithm doesn't have long-term memory and even if it had it wouldn't be able to generalize its knowledge to other similar tasks super exciting directions for future work thanks for watching and for your generous support and I'll see you next time
Original Description
Google DeepMind implemented an artificial intelligence program using deep reinforcement learning that plays Atari games and improves itself to a superhuman level. The technique is called deep Q-learning, it uses a combination of deep neural networks and reinforcement learning, and it is capable of playing many Atari games as good or better than humans. After presenting their initial results with the algorithm, Google almost immediately acquired the company for several hundred million dollars, hence the name Google DeepMind. I am sure that this is one of the biggest triumphs of deep learning, especially given the fact that now the first few successful experiments for 3D games are out there!
________________________
The Nature paper "Human-level control through deep reinforcement learning" is available here:
http://www.nature.com/nature/journal/v518/n7540/full/nature14236.html
http://www.cs.swarthmore.edu/~meeden/cs63/s15/nature15b.pdf
The code is available here:
https://sites.google.com/a/deepmind.com/dqn/
Ilya Kuzovkin's fork with visualization:
https://github.com/kuz/DeepMind-Atari-Deep-Q-Learner
This configuration file will run Ilya Kuzovkin's version with less than 1GB of VRAM:
http://cg.tuwien.ac.at/~zsolnai/wp/wp-content/uploads/2015/03/run_gpu
Recommended for you:
Artificial Neural Networks and Deep Learning - https://www.youtube.com/watch?v=rCWTOOgVXyE&list=PLujxSBD-JXgnqDD1n-V30pKtp6Q886x7e&index=13
Recurrent Neural Network Writes Sentences About Images - https://www.youtube.com/watch?v=e-WB4lfg30M&list=PLujxSBD-JXgnqDD1n-V30pKtp6Q886x7e&index=15
Deep Neural Network Learns Van Gogh's Art - https://www.youtube.com/watch?v=-R9bJGNHltQ&list=PLujxSBD-JXgnqDD1n-V30pKtp6Q886x7e&index=22
Terrain Traversal with Reinforcement Learning - https://www.youtube.com/watch?v=_yjHPu1aYCY&list=PLujxSBD-JXgnqDD1n-V30pKtp6Q886x7e&index=9
Subscribe if you would like to see more of these! - http://www.youtube.com/subscription_center?add_user=keeroyz
The thumbnail was made
Watch on YouTube ↗
(saves to browser)
Sign in to unlock AI tutor explanation · ⚡30
Playlist
Uploads from Two Minute Papers · Two Minute Papers · 31 of 60
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
▶
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
Fluid Simulations with Blender and Wavelet Turbulence | Two Minute Papers #1
Two Minute Papers
Capturing Waves of Light With Femto-photography | Two Minute Papers #2
Two Minute Papers
Artificial Neural Networks and Deep Learning | Two Minute Papers #3
Two Minute Papers
Blender Rendering - Top 7 LuxRender Features
Two Minute Papers
Simulating Breaking Glass | Two Minute Papers #4
Two Minute Papers
Time Lapse Videos From Community Photos | Two Minute Papers #5
Two Minute Papers
AI Learns Van Gogh's Art
Two Minute Papers
Hydrographic Printing | Two Minute Papers #7
Two Minute Papers
Announcing LuxRender 1.5
Two Minute Papers
Digital Creatures Learn To Walk | Two Minute Papers #8
Two Minute Papers
Manipulating Photorealistic Renderings | Two Minute Papers #9
Two Minute Papers
Adaptive Fluid Simulations | Two Minute Papers #10
Two Minute Papers
Building Bridges With Flying Machines | Two Minute Papers #11
Two Minute Papers
Reconstructing Sound From Vibrations | Two Minute Papers #12
Two Minute Papers
Creating Photographs Using Deep Learning | Two Minute Papers #13
Two Minute Papers
Adaptive Cloth Simulations | Two Minute Papers #14
Two Minute Papers
Synthesizing Sound From Collisions | Two Minute Papers #15
Two Minute Papers
Metropolis Light Transport | Two Minute Papers #16
Two Minute Papers
3D Printing a Glockenspiel | Two Minute Papers #17
Two Minute Papers
Modeling Colliding and Merging Fluids | Two Minute Papers #18
Two Minute Papers
Recurrent Neural Network Writes Music and Shakespeare Novels | Two Minute Papers #19
Two Minute Papers
Gradients, Poisson's Equation and Light Transport | Two Minute Papers #20
Two Minute Papers
Real-Time Facial Expression Transfer | Two Minute Papers #21
Two Minute Papers
Automatic Lecture Notes From Videos | Two Minute Papers #22
Two Minute Papers
Be a Part of Two Minute Papers on Patreon!
Two Minute Papers
Recurrent Neural Network Writes Sentences About Images | Two Minute Papers #23
Two Minute Papers
How Does Deep Learning Work? | Two Minute Papers #24
Two Minute Papers
Cryptography, Perfect Secrecy and One Time Pads | Two Minute Papers #25
Two Minute Papers
Terrain Traversal with Reinforcement Learning | Two Minute Papers #26
Two Minute Papers
Multiple-Scattering Microfacet BSDFs with the Smith Model
Two Minute Papers
Google DeepMind's Deep Q-Learning & Superhuman Atari Gameplays | Two Minute Papers #27
Two Minute Papers
Are We Living In a Computer Simulation? | Two Minute Papers #28
Two Minute Papers
Artificial Superintelligence [Audio only] | Two Minute Papers #29
Two Minute Papers
Automatic Parameter Control for Metropolis Light Transport | Two Minute Papers #30
Two Minute Papers
Randomness and Bell's Inequality [Audio only] | Two Minute Papers #31
Two Minute Papers
OpenAI - Non-profit AI company by Elon Musk and Sam Altman
Two Minute Papers
How Do Genetic Algorithms Work? | Two Minute Papers #32
Two Minute Papers
Painting with Fluid Simulations | Two Minute Papers #33
Two Minute Papers
Peer Review #1 [Audio only] | Two Minute Papers
Two Minute Papers
Neural Programmer-Interpreters Learn To Write Programs | Two Minute Papers #34
Two Minute Papers
9 Cool Deep Learning Applications | Two Minute Papers #35
Two Minute Papers
Designing Cities and Furnitures With Machine Learning | Two Minute Papers #36
Two Minute Papers
Designing 3D Printable Robotic Creatures | Two Minute Papers #37
Two Minute Papers
3D Printing Objects With Caustics | Two Minute Papers #38
Two Minute Papers
Interactive Editing of Subsurface Scattering | Two Minute Papers #39
Two Minute Papers
Simulating Viscosity and Melting Fluids | Two Minute Papers #40
Two Minute Papers
What Do Virtual Objects Sound Like? | Two Minute Papers #41
Two Minute Papers
How DeepMind Conquered Go With Deep Learning (AlphaGo) | Two Minute Papers #42
Two Minute Papers
Breaking Deep Learning Systems With Adversarial Examples | Two Minute Papers #43
Two Minute Papers
Extrapolations and Crowdfunded Research (Experiment) | Two Minute Papers #44
Two Minute Papers
Biophysical Skin Aging Simulations | Two Minute Papers #45
Two Minute Papers
What is Impostor Syndrome? | Two Minute Papers #46
Two Minute Papers
Should You Take the Stairs at Work? (For Weight Loss) | Two Minute Papers #47
Two Minute Papers
Artistic Manipulation of Caustics | Two Minute Papers #48
Two Minute Papers
Deep Learning Program Learns to Paint | Two Minute Papers #49
Two Minute Papers
Interactive Photo Recoloring | Two Minute Papers #50
Two Minute Papers
How To Get Started With Machine Learning? | Two Minute Papers #51
Two Minute Papers
Awesome Research For Everyone! - Two Minute Papers Channel Trailer
Two Minute Papers
10 More Cool Deep Learning Applications | Two Minute Papers #52
Two Minute Papers
How DeepMind's AlphaGo Defeated Lee Sedol | Two Minute Papers #53
Two Minute Papers
More on: RL Foundations
View skill →Related Reads
📰
📰
📰
📰
Help Choosing Neural Network Architecture for Matrix Classification
Reddit r/deeplearning
How to Choose the Best Deep Learning Model for Medical Imaging
Medium · Deep Learning
Another Way to Read Neural Geometry
Medium · Data Science
Another Way to Read Neural Geometry
Medium · Deep Learning
🎓
Tutor Explanation
DeepCamp AI