4 ms·
This passage from the article might help answer that: "DeepMind learned to play video games by randomly taking any action it could. This may be fine for video
by cjfont 11y ago
This passage from the article might help answer that:
"DeepMind learned to play video games by randomly taking any action it could. This may be fine for video games, but in the real world it could be expensive, time-consuming, and even deadly to have a robot trying to learn by trying every possible action to see what generated a “reward” and what didn’t. So Osaro also wrote algorithms that helps computers learn by mimicking humans."