Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
brrrrrm
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
brrrrrm
10d ago
only sorts numbers? wouldn't radix be much better?
2.
▲
Show HN: Serverside.chat
(github.com)
1 points
by
brrrrrm
12d ago
|
0 comments
3.
▲
by
brrrrrm
16d ago
this is basically the only thing pre-training teams work on in labs. compute efficiency is the metric, the assumption that scaling = intelligence is considered a given.
4.
▲
by
brrrrrm
16d ago
they say they're looking at base models, so I think it's fairly compared as written.
5.
▲
by
brrrrrm
24d ago
perhaps its unfair to say this in hindsight, but it's a fairly straightforward application of little's law that's been around for some time https://arxiv.org/html/2401.09670v2
6.
▲
by
brrrrrm
24d ago
this is a nice and concise writeup. what's striking to me is that these techniques really have not changed in /years/. sure, precision has become slightly lower, spec decoding acceptance has gotten slightly better and the c
7.
▲
by
brrrrrm
2mo ago
this is Qwen3.8 max, right? https://qwen.ai/blog?id=qwen3.8
8.
▲
by
brrrrrm
2mo ago
are these all uniform quantization? or mixed and matched by layer (can't tell from the naming scheme)
9.
▲
by
brrrrrm
2mo ago
this is cool but like, are we just vibe coding NAND burners at this point? these decode times don't really tell the whole story, because prefill becomes the bottleneck. half an hour to process 10k tokens on an M5 seems... not great
10.
▲
by
brrrrrm
2mo ago
there's wifi ah, which runs on 900mhz band and has the same 30dbm limitation I've used it with some raspberry pis to create hi-fidelity walkie talkies it's quite pleasant.
11.
▲
by
brrrrrm
2mo ago
Models are increasingly showing their ability to extrapolate into the human unexplored (math proofs being the most apparent). What gives you confidence the absurdity of life is uniquely difficult for models to source?
12.
▲
by
brrrrrm
2mo ago
> you are basically looking at a whole system prompt just describing the new language whats wrong with this? You may be over-indexing on the need for large quantities of examples. These days self-play through RL is far more effective and
13.
▲
by
brrrrrm
3mo ago
it's very much an in-domain term for folks in machine learning. heavily used when pipeline parallelism caught on in training https://alband.github.io/doc_view/pipeline.html
14.
▲
by
brrrrrm
4mo ago
it has paddle shifters - what are those for?
15.
▲
by
brrrrrm
5mo ago
what's MRT?
16.
▲
by
brrrrrm
5mo ago
same can be said for a lot of things tho. e.g. nature used to be fun but then we discovered it all :’( I miss when ships literally sailed into the unknown and found surprising and novel things like hot peppers and pineapples
17.
▲
by
brrrrrm
5mo ago
I agree fully. Hyundai has a mockup that starts to get there (different era, but same concept) called the N vision 74[1], but I doubt we'll see it in market anytime soon. The unfortunate reality IIUC is that modern cars (electric veh
18.
▲
by
brrrrrm
6mo ago
meta.ai in instant mode gets it first try too (I think?) ``` 2x + y = \operatorname{eml}\Big(1,\; \operatorname{eml}\big(\operatorname{eml}(1,\; \operatorname{eml}(\operatorname{eml}(1,\; \operatorname{eml}(\operatorname{eml}(L_2 + L_x, 1),
19.
▲
by
brrrrrm
6mo ago
you're right, this is actually correctly placed! I was confusing the orientation. I live right around there and recognize the M&T bank in the photo on the left, so it can't be down by 9th
20.
▲
by
brrrrrm
6mo ago
I checked 3 spots I'm familiar with and 1 is wrong https://www.oldnyc.org/#707133f-a this is supposed to be here https://www.oldnyc.org/#702487f-a also, if folks are interested in these old depictions
21.
▲
How to Fail as an Organization in 2026
(jott.live)
2 points
by
brrrrrm
8mo ago
|
0 comments
22.
▲
Internal Combustion Engine Acoustic Synthesis
(jott.live)
2 points
by
brrrrrm
9mo ago
|
0 comments
23.
▲
by
brrrrrm
9mo ago
looks cool! one bit of feedback: make your demo gif get to the point faster. either practice typing a bit quicker or speed it up 2x for the typing section
24.
▲
Show HN: Binfer, an experimental LLM inference engine in TypeScript and CUDA
(github.com)
1 points
by
brrrrrm
10mo ago
|
0 comments
25.
▲
by
brrrrrm
10mo ago
on Bun's website, the runtime section features HTTP, networking, storage -- all are very web-focused. any plans to start expanding into native ML support? (e.g. GPUs, RDMA-type networking, cluster management, NFS)
26.
▲
Five Times Faster
(jott.live)
2 points
by
brrrrrm
10mo ago
|
0 comments
27.
▲
Bitwise Consistent On-Policy Reinforcement Learning with VLLM and TorchTitan
(blog.vllm.ai)
1 points
by
brrrrrm
11mo ago
|
0 comments
28.
▲
by
brrrrrm
11mo ago
we've discovered some kind of differentiable computer[1] and as with all computers, people have their own interests and hobbies they use them for. but unlike computers, everyone pitches their interest or hobby as being the only one t
29.
▲
by
brrrrrm
11mo ago
a recent wave of interest in bitwise equivalent execution had a lot of kernels this level get pumped out. new attention mechanisms also often need new kernels to run at any reasonable rate theres definitely a breed of frontend-only ML dev t
30.
▲
Should we apply old-school multi-core scheduling to GPUs?
(jott.live)
4 points
by
brrrrrm
11mo ago
|
0 comments
More ›