Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
farisallafi
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
farisallafi
3mo ago
thanks a lot for the support!! yeah for the next training run im planning on testing that out extensively before putting into training
2.
▲
by
farisallafi
3mo ago
haha yeah I wrote this fully by hand the only LLM's input into this was as a peer reviewer I guess I have spent a bit too much time with the models these days (for the record I mostly use Fable 5)
3.
▲
by
farisallafi
3mo ago
Hey! Thanks so much for the thoughtful reply... totally agree with your points on benchmarking and research also while the website ui was supported by an LLM, all the actual writing was by me(a human) whether it looks somewhat AI generated(
4.
▲
by
farisallafi
3mo ago
Author here. This is hr-diffuse-1-nano: bidirectional Mamba-2 + LLaDA-style masked diffusion at 288M params, cross- arch distilled from SmolLM-135M, trained on 1xh100 for ~$500. The honest headline results: 14% infill recovery where autoreg
5.
▲
Show HN: I trained a language model that thinks the capital of Japan is Paris
(hamiltonianresearch.xyz)
15 points
by
farisallafi
3mo ago
|
8 comments