Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jazarwil
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
jazarwil
9mo ago
I'd love to run more games, just very expensive unfortunately.
2.
▲
by
jazarwil
9mo ago
Wdym exactly? I ran 163 games, are you suggesting more games or something else?
3.
▲
by
jazarwil
9mo ago
There are a few reasons that come to mind, such as winning larger pots on average, and also playing more hands by virtue of not getting knocked out as frequently.
4.
▲
by
jazarwil
9mo ago
Oh to be clear, there are ~21k hands here, and far more decisions than that.
5.
▲
by
jazarwil
9mo ago
I didn't incorporate any open weights/source models just to limit the number of API providers I had to juggle, but it is just a config change if somebody wants to try a run with them.
6.
▲
by
jazarwil
9mo ago
Not at the moment, do you have something in mind?
7.
▲
by
jazarwil
9mo ago
I've seen some theories tossed around but I don't think I'm qualified to offer an authoritative answer. Gemini 3 Pro specifically seems to be consistently "tighter" and more passive than Flash.
8.
▲
by
jazarwil
9mo ago
It greatly depends on the models. The 6-handed setup with Opus and Pro cost about $30/game. The 4-handed setup with just small models was $6/game. I'd love to run more but I already spent quite a bit as it is.
9.
▲
Show HN: Watch LLMs play 21,000 hands of Poker
(pokerbench.adfontes.io)
36 points
by
jazarwil
9mo ago
|
19 comments
10.
▲
by
jazarwil
3y ago
You cannot compare GPT 4 to Gemini Pro. They are different classes of models.
11.
▲
by
jazarwil
3y ago
What exactly do you think you saw? Bard is not trained on any data of that nature, unless it is already publicly available.