Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
qwert7890
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
4th Place Solution to Python Speed Coding Challenge
(medium.com)
1 points
by
qwert7890
6y ago
|
0 comments
2.
▲
Remote Control (Learning from Others’ Mistakes)
(blog.ipinfo.io)
1 points
by
qwert7890
6y ago
|
0 comments
3.
▲
by
qwert7890
8y ago
Thank you Ajedi32, I updated the post https://medium.com/@stelmaszczykadam/do-openai-five-dota-2-b... .
4.
▲
by
qwert7890
8y ago
> They managed to coordinate simply by looking at what the others were doing. This bit seems incorrect, https://medium.com/@stelmaszczykadam/do-openai-five-dota-2-b... .
5.
▲
Do OpenAI Five Dota 2 bots communicate?
(medium.com)
1 points
by
qwert7890
8y ago
|
0 comments
6.
▲
by
qwert7890
8y ago
You broke the LGPL license, you didn't state changes: https://github.com/snowkylin/ntm/blob/master/LICENSE Moreover, in the paper 5 times you write: "Our implementation" You also don'
7.
▲
A Deep Dive into Reinforcement Learning
(toptal.com)
14 points
by
qwert7890
8y ago
|
0 comments
8.
▲
by
qwert7890
9y ago
Simplest RL algorithm (Q-learning) achieves 100m in QWOP: https://www.youtube.com/watch?v=e27TUmMkOA0 Although it found and exploited a local maximum of "knee scraping" technique (which humans can replicate) :)