Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
chaeronanaut
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
chaeronanaut
2y ago
> The words that are coming out of the model are generated to optimize for RLHF and closeness to the training data, that's it! This is false, reasoning models are rewarded/punished based on performance at verifiable tasks, not
2.
▲
by
chaeronanaut
3y ago
BT2 is old news, we have BT4 now
3.
▲
by
chaeronanaut
3y ago
An excellent explanation of Magic Bitboards can be found here: https://analog-hors.github.io/site/magic-bitboards/
4.
▲
by
chaeronanaut
4y ago
this pretty much summarises my opinion - one nitpick - i assume you meant " omit bounds and other checks", not "emit bounds and other checks" which seems to mean the opposite of what you're intending