video thumbnail 11:27
Reinforcement Learning: Crash Course AI#9

2019-10-11

[public] 70.3K views, 4.10K likes, 35.0 dislikes audio only

channel thumbCrashCourse

Reinforcement learning is particularly useful in situations where we want to train AIs to have certain skills we don’t fully understand ourselves. Unlike some of the techniques we’ve discussed so far, reinforcement learning generally only looks at how an AI performs a task AFTER it has completed it. And when an AI completes that task figuring out when and how to reward an AI, called credit assignment, is one of the hardest parts of reinforcement learning. So today, we’re going to explore these ideas, introduce a ton of new terms like value, policy, agent, environment, actions, and states and we’ll show you how we can use strategies like exploration and exploitation to train John Green Bot to find things more efficiently next time.

Crash Course AI is produced in association with PBS Digital Studios:

https://www.youtube.com/user/pbsdigitalstudios/videos

Crash Course is on Patreon! You can support us directly by signing up at http://www.patreon.com/crashcourse

Thanks to the following patrons for their generous monthly contributions that help keep Crash Course free for everyone forever:

Eric Prestemon, Sam Buck, Mark Brouwer, Indika Siriwardena, Avi Yashchin, Timothy J Kwist, Brian Thomas Gossett, Haixiang N/A Liu, Jonathan Zbikowski, Siobhan Sabino, Zach Van Stanley, Jennifer Killen, Nathan Catchings, Brandon Westmoreland, dorsey, Kenneth F Penttinen, Trevin Beattie, Erika & Alexa Saur, Justin Zingsheim, Jessica Wode, Tom Trval, Jason Saslow, Nathan Taylor, Khaled El Shalakany, SR Foxley, Sam Ferguson, Yasenia Cruz, Eric Koslow, Caleb Weeks, Tim Curwick, David Noe, Shawn Arnold, William McGraw, Andrei Krishkevich, Rachel Bright, Jirat, Ian Dundore

--

Want to find Crash Course elsewhere on the internet?

Facebook - http://www.facebook.com/YouTubeCrashCourse

Twitter - http://www.twitter.com/TheCrashCourse

Tumblr - http://thecrashcourse.tumblr.com

Support Crash Course on Patreon: http://patreon.com/crashcourse

CC Kids: http://www.youtube.com/crashcoursekids

#CrashCourse #ArtificialIntelligence #MachineLearning


Intro
/youtube/video/nIgIv4IfJ6s?t=0
REINFORCEMENT LEARNING
/youtube/video/nIgIv4IfJ6s?t=33
REWARD
/youtube/video/nIgIv4IfJ6s?t=104
CREDIT ASSIGNMENT
/youtube/video/nIgIv4IfJ6s?t=206
EXPLORATION
/youtube/video/nIgIv4IfJ6s?t=455
VALUE FUNCTION
/youtube/video/nIgIv4IfJ6s?t=606
CrashCourse At Crash Course, we believe that high-quality educational videos should be available to everyone for free! Subscribe for weekly videos from our current courses! Right now, we're producing Climate & Energy. The Crash Course team has produced more than 45 courses on a wide variety of subjects, including organic chemistry, literature, world history, biology, philosophy, theater, ecology, and many more! We also recently teamed up with Arizona State University to bring you more courses on the Study Hall channel. Help support Crash Course at Patreon.com/CrashCourse.
/youtube/channel/UCX6b17PVsYBQ0ip5gyeme-Q
Thought Café our animators
youtube.com/channel/UCwTZ-JLF5FQ3EmQ2nPaS-lg
Computer Science by CrashCourse
/youtube/video/tpIctyqH29Q