
Reinforcement Learning: Crash Course AI#9
video description
Date: 2022-04-04
Related videos
Comments and reviews: 10
Dojo
Black & White and Black & White: Creature Island used reinforcement learning. The creature you commanded could learn incredibly complex routines, such as planting a sapling, water the tree with the water miracle, then pick it up and throw it into the resource center and repeat. With enough training.
I really hope that we'll see more games exploring that kind of relationship with a computer character. Imagine a game where you're personally teaching a group of monsters how to hunt and then guiding their instincts by reinforcing or punishing a particular set of circumstances, until they conquer their world.
reply
Black & White and Black & White: Creature Island used reinforcement learning. The creature you commanded could learn incredibly complex routines, such as planting a sapling, water the tree with the water miracle, then pick it up and throw it into the resource center and repeat. With enough training.
I really hope that we'll see more games exploring that kind of relationship with a computer character. Imagine a game where you're personally teaching a group of monsters how to hunt and then guiding their instincts by reinforcing or punishing a particular set of circumstances, until they conquer their world.
reply
Matt
Is there a better reason than consolidating the total amount of stored data the reason we only store a single value per square? Why not store 4 values per square so you can store a value per direction you could go from the current spot. That way you could find/exploit the near black hole shortcut that the current algorithm is too scared to find.
reply
Is there a better reason than consolidating the total amount of stored data the reason we only store a single value per square? Why not store 4 values per square so you can store a value per direction you could go from the current spot. That way you could find/exploit the near black hole shortcut that the current algorithm is too scared to find.
reply
Trent
In the john green bot example, is the objective to find the shortest path or get the most points? What would getting more points even do, I feel like in that case exploration is best so that you can find the shortest path, exploiting only when racing another bot
reply
In the john green bot example, is the objective to find the shortest path or get the most points? What would getting more points even do, I feel like in that case exploration is best so that you can find the shortest path, exploiting only when racing another bot
reply
PinkPonyOfPrey
0: 17 That cookie looked completely. edible hahaha! What brand is that? :D
Screw Reinforcement Learning. I'm now officially hungry!
[back from the kitchen]
Reinforcement Learning is actually extremely interesting! :D
reply
0: 17 That cookie looked completely. edible hahaha! What brand is that? :D
Screw Reinforcement Learning. I'm now officially hungry!
[back from the kitchen]
Reinforcement Learning is actually extremely interesting! :D
reply
Shashank
yay, this was actually better than most of the explanatory videos i have seen. thanks for providing us always with informative content, crash course. looking forward for more of these videos
reply
yay, this was actually better than most of the explanatory videos i have seen. thanks for providing us always with informative content, crash course. looking forward for more of these videos
reply
Kurt
I don't agree with the bagel/donut choice example. Why choose the option of two bagels or donuts vs. the greater risk of more donuts (6) or a guaranteed single donut?
reply
I don't agree with the bagel/donut choice example. Why choose the option of two bagels or donuts vs. the greater risk of more donuts (6) or a guaranteed single donut?
reply
Ian
Not sure the kitchen metaphor works for me. Why is the bag more likely to contain donuts than the box? It sure looked like the kind of box that donuts come in to me.
reply
Not sure the kitchen metaphor works for me. Why is the bag more likely to contain donuts than the box? It sure looked like the kind of box that donuts come in to me.
reply
Jimmy
Yo as a black guy with dreadlocks who likes coding, it-s really cool to listen to another black guy with dreadlocks who likes coding
reply
Yo as a black guy with dreadlocks who likes coding, it-s really cool to listen to another black guy with dreadlocks who likes coding
reply
Raj
Are you going to use openai for rl and keras when we come to deep reinforcement learning
When will this playlist be finished.
reply
Are you going to use openai for rl and keras when we come to deep reinforcement learning
When will this playlist be finished.
reply
Antti
Why would JohnGreenBot in that battery example only go in straight lines? Would it not be better to go in a diagonal path?
reply
Why would JohnGreenBot in that battery example only go in straight lines? Would it not be better to go in a diagonal path?
reply
Add a review, comment
Other channel videos















