Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kywch
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
Less is more: When agents learn not because but despite my doings
(kywch.github.io)
2 points
by
kywch
8mo ago
|
1 comments
2.
▲
by
kywch
9mo ago
I think Go-Explore ( https://arxiv.org/abs/1901.10995 ) is promising. It'll provide automatic scaffolding and prevent catastrophic forgetting. If one can frame the problem into a competition, then self-play has been
3.
▲
by
kywch
9mo ago
You can watch these agents play live, and you can also intervene * 2048: https://kywch.github.io/games/2048.html * Tetris: https://kywch.github.io/games/tetris.html
4.
▲
by
kywch
9mo ago
Pufferlib already had a pretty good model before: https://puffer.ai/ocean.html?env=tetris