Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
d_legs
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
d_legs
2y ago
Totally agree that there are better ways to build a program to beat tic tac toe. I'd expect an LLM could probably write the code itself with a few turns of the crank (it wrote most of the code I used to test this). The point here is to
2.
▲
by
d_legs
2y ago
Agreed. I tried asking the models to outline a strategy and they can produce a decent output although not as robust as I expected. I'm sure you could fine tune an LLM to be good at Tic-Tac-Toe too but the surprising thing to me was how
3.
▲
by
d_legs
2y ago
Hi HN - I decided to try to compare LLMs by having them play Tic-Tac-Toe. The results were surprisingly bad considering all the talk about how LLMs have "saturated" benchmarks. Have you run into any tasks where LLMs are way worse
4.
▲
LLMs are really bad at Tic-Tac-Toe
(gensx.com)
4 points
by
d_legs
2y ago
|
8 comments