Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
eric_gu
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
Training AI to Play Super Smash Bros. Melee
(ericyuegu.com)
2 points
by
eric_gu
2y ago
|
0 comments
2.
▲
by
eric_gu
2y ago
Ah understood. Makes sense now!
3.
▲
by
eric_gu
2y ago
Thanks for the reply! Re question #3: I'm not sure I understand why you need to vary the base model or how doing so would allow LSR to take advantage? Isn't your LSR technique used on the activations of the evaluator model? As a n
4.
▲
by
eric_gu
2y ago
This is cool! I have not read up on evaluation techniques that use LLM-as-a-Judge, so I hadn't heard of the term "evaluator LLM" before. Questions that came to mind: - How are you deciding on the positive/negative concep