Terrain Traversal with Reinforcement Learning | Two Minute Papers #26
Key Takeaways
Applies reinforcement learning to train a digital dog for terrain traversal
Full Transcript
dear fellow Scholars this is two-minute papers with Caro here reinforcement learning is a technique that can learn how to play computer games or any kind of activity that requires a sequence of actions we are not interested in figuring out what we see on an image because the answer is one thing we are always interested in a sequence of actions the input for reinforcement learning is a state that describes where we are and how the world looks around us and the algorithm outputs optimal next action to take in this case we would like a digital dog to run and leap over and onto obstacles by choosing the optimal next action it is quite difficult as there are a lot of body parts to control in harmony the algorithm has to be able to decide how to control leg forces spine curvature angles for the shoulder elbow hip and knees and what is really amazing is that if it has learned everything properly it will come up with exactly the same movements as we'd expect animals to do in real life so this is how reinforcement learning works if you do well you get a reward and if you don't you get some kind of punishment these rewards and punishments are usually encoded in a score if your score is increasing you know you've done something right and you try to self-reflect and analyze the last few actions to find out which of them were responsible for this positive change the score would be for instance how far the dog could run on the map without falling and at the same time it also makes sense to minimize the amount of effort to make it happen so reinforcement learning in a nutshell it is very similar to how a real world animal or even a human would learn if you're not doing well try something new and if you're succeeding remember what you did that led to your success and keep doing that in this technique dogs were used to demonstrate the concept but it's worth noting that it also works with bads reinforcement learning is typically used in many control situations that are extremely difficult to solve otherwise like controlling a quadrocopter properly it's quite delightful to see such a cool work especially given that there are not so many uses of reinforcement learning in computer Graphics yet I wonder why that is is it that not so many graphical tasks require a sequence of actions or maybe we just need to shift our mindset and get used to the idea of formalizing problems in a different way so we can use such powerful techniques to solve them it is definitely worth the effort thanks for watching and for your generous support and I'll see you next time
Original Description
Reinforcement learning is a technique that can learn how to play computer games, or any kind of activity that requires a sequence of actions. In this case, we would like a digital dog to run, and leap over and onto obstacles by choosing the optimal next action. It is quite difficult as there are a lot of body parts to control in harmony. And what is really amazing is that if it has learned everything properly, it will come up with exactly the same movements as we'd expect animals to do in real life! In this technique, dogs were used to demonstrate that reinforcement learning works well in this context, but it's worth noting that it also works with bipeds.
_____________________________
The paper "Dynamic Terrain Traversal Skills Using Reinforcement Learning " is available here:
http://www.cs.ubc.ca/~van/papers/2015-TOG-terrainRL/
Recommended for you:
Digital Creatures Learn To Walk - https://www.youtube.com/watch?v=kQ2bqz3HPJE
Subscribe if you would like to see more of these! - http://www.youtube.com/subscription_center?add_user=keeroyz
Thumbnail image by localpups (CC BY 2.0). It was slightly edited (flipped, color adjustments, content aware filling) - https://flic.kr/p/wXfFt1
Splash screen/thumbnail design: Felícia Fehér - http://felicia.hu
Károly Zsolnai-Fehér's links:
Patreon → https://www.patreon.com/TwoMinutePapers
Facebook → https://www.facebook.com/TwoMinutePapers/
Twitter → https://twitter.com/karoly_zsolnai
Web → https://cg.tuwien.ac.at/~zsolnai/
Watch on YouTube ↗
(saves to browser)
Sign in to unlock AI tutor explanation · ⚡30
Playlist
Uploads from Two Minute Papers · Two Minute Papers · 29 of 60
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
▶
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
Fluid Simulations with Blender and Wavelet Turbulence | Two Minute Papers #1
Two Minute Papers
Capturing Waves of Light With Femto-photography | Two Minute Papers #2
Two Minute Papers
Artificial Neural Networks and Deep Learning | Two Minute Papers #3
Two Minute Papers
Blender Rendering - Top 7 LuxRender Features
Two Minute Papers
Simulating Breaking Glass | Two Minute Papers #4
Two Minute Papers
Time Lapse Videos From Community Photos | Two Minute Papers #5
Two Minute Papers
AI Learns Van Gogh's Art
Two Minute Papers
Hydrographic Printing | Two Minute Papers #7
Two Minute Papers
Announcing LuxRender 1.5
Two Minute Papers
Digital Creatures Learn To Walk | Two Minute Papers #8
Two Minute Papers
Manipulating Photorealistic Renderings | Two Minute Papers #9
Two Minute Papers
Adaptive Fluid Simulations | Two Minute Papers #10
Two Minute Papers
Building Bridges With Flying Machines | Two Minute Papers #11
Two Minute Papers
Reconstructing Sound From Vibrations | Two Minute Papers #12
Two Minute Papers
Creating Photographs Using Deep Learning | Two Minute Papers #13
Two Minute Papers
Adaptive Cloth Simulations | Two Minute Papers #14
Two Minute Papers
Synthesizing Sound From Collisions | Two Minute Papers #15
Two Minute Papers
Metropolis Light Transport | Two Minute Papers #16
Two Minute Papers
3D Printing a Glockenspiel | Two Minute Papers #17
Two Minute Papers
Modeling Colliding and Merging Fluids | Two Minute Papers #18
Two Minute Papers
Recurrent Neural Network Writes Music and Shakespeare Novels | Two Minute Papers #19
Two Minute Papers
Gradients, Poisson's Equation and Light Transport | Two Minute Papers #20
Two Minute Papers
Real-Time Facial Expression Transfer | Two Minute Papers #21
Two Minute Papers
Automatic Lecture Notes From Videos | Two Minute Papers #22
Two Minute Papers
Be a Part of Two Minute Papers on Patreon!
Two Minute Papers
Recurrent Neural Network Writes Sentences About Images | Two Minute Papers #23
Two Minute Papers
How Does Deep Learning Work? | Two Minute Papers #24
Two Minute Papers
Cryptography, Perfect Secrecy and One Time Pads | Two Minute Papers #25
Two Minute Papers
Terrain Traversal with Reinforcement Learning | Two Minute Papers #26
Two Minute Papers
Multiple-Scattering Microfacet BSDFs with the Smith Model
Two Minute Papers
Google DeepMind's Deep Q-Learning & Superhuman Atari Gameplays | Two Minute Papers #27
Two Minute Papers
Are We Living In a Computer Simulation? | Two Minute Papers #28
Two Minute Papers
Artificial Superintelligence [Audio only] | Two Minute Papers #29
Two Minute Papers
Automatic Parameter Control for Metropolis Light Transport | Two Minute Papers #30
Two Minute Papers
Randomness and Bell's Inequality [Audio only] | Two Minute Papers #31
Two Minute Papers
OpenAI - Non-profit AI company by Elon Musk and Sam Altman
Two Minute Papers
How Do Genetic Algorithms Work? | Two Minute Papers #32
Two Minute Papers
Painting with Fluid Simulations | Two Minute Papers #33
Two Minute Papers
Peer Review #1 [Audio only] | Two Minute Papers
Two Minute Papers
Neural Programmer-Interpreters Learn To Write Programs | Two Minute Papers #34
Two Minute Papers
9 Cool Deep Learning Applications | Two Minute Papers #35
Two Minute Papers
Designing Cities and Furnitures With Machine Learning | Two Minute Papers #36
Two Minute Papers
Designing 3D Printable Robotic Creatures | Two Minute Papers #37
Two Minute Papers
3D Printing Objects With Caustics | Two Minute Papers #38
Two Minute Papers
Interactive Editing of Subsurface Scattering | Two Minute Papers #39
Two Minute Papers
Simulating Viscosity and Melting Fluids | Two Minute Papers #40
Two Minute Papers
What Do Virtual Objects Sound Like? | Two Minute Papers #41
Two Minute Papers
How DeepMind Conquered Go With Deep Learning (AlphaGo) | Two Minute Papers #42
Two Minute Papers
Breaking Deep Learning Systems With Adversarial Examples | Two Minute Papers #43
Two Minute Papers
Extrapolations and Crowdfunded Research (Experiment) | Two Minute Papers #44
Two Minute Papers
Biophysical Skin Aging Simulations | Two Minute Papers #45
Two Minute Papers
What is Impostor Syndrome? | Two Minute Papers #46
Two Minute Papers
Should You Take the Stairs at Work? (For Weight Loss) | Two Minute Papers #47
Two Minute Papers
Artistic Manipulation of Caustics | Two Minute Papers #48
Two Minute Papers
Deep Learning Program Learns to Paint | Two Minute Papers #49
Two Minute Papers
Interactive Photo Recoloring | Two Minute Papers #50
Two Minute Papers
How To Get Started With Machine Learning? | Two Minute Papers #51
Two Minute Papers
Awesome Research For Everyone! - Two Minute Papers Channel Trailer
Two Minute Papers
10 More Cool Deep Learning Applications | Two Minute Papers #52
Two Minute Papers
How DeepMind's AlphaGo Defeated Lee Sedol | Two Minute Papers #53
Two Minute Papers
More on: RL Foundations
View skill →Related Reads
📰
📰
📰
📰
A lightweight workflow for keeping up with AI conference papers
Dev.to · Daniel
Why CitedEvidence Believes Great Researchers Read Less Than You Think
Medium · AI
How to Write a Literature Review That Actually Argues Something
Medium · Machine Learning
I Built a Personal Paper Engine to Stop Losing Research Papers
Dev.to · Ethan
🎓
Tutor Explanation
DeepCamp AI