Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sosodev
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
12 ms
·
121.
▲
by
sosodev
8mo ago
I think there are multiple ways these infinite loops can occur. It can be an inference engine bug because the engine doesn't recognize the specific format of tags/tokens the model generates to delineate the different types of toke
122.
▲
by
sosodev
8mo ago
I was testing the 4-bit Qwen3 Coder Next on my 395+ board last night. IIRC it was maintaining around 30 tokens a second even with a large context window. I haven't tried Minimax M2.5 yet. How do its capabilities compare to Qwen3 Coder
123.
▲
by
sosodev
8mo ago
Isn't it just the usual feedback loop that happens with popular podcasters? They have connections and get a few highly popular guests on. As long as their demeanor is agreeable and they keep the conversation interesting other high prof
124.
▲
by
sosodev
8mo ago
Do you mean agents dating other agents for their own sake or on behalf of their owners?
125.
▲
by
sosodev
8mo ago
An excellent quote, but I'm curious, how do you think it applies here?
126.
▲
by
sosodev
8mo ago
The knee-jerk reaction reaction to Moltbook is almost certainly "what a waste of compute" or "a security disaster waiting to happen". Both of those thoughts have merit and are worth considering, but we must acknowledge t
127.
▲
by
sosodev
9mo ago
The browser works shockingly well considering it was created in 72 hours. It can render Wikipedia well enough to read and browse articles. With some basic form handling and browser standards (url bar, history, bookmarks, etc) it would be a
128.
▲
by
sosodev
9mo ago
> The security hazards of artisanal hosting are more real than ever How could this possibly be true? It's not at all rocket science to create a static blog and serve it via a production grade web server (nginx, etc). > The UX of
129.
▲
by
sosodev
9mo ago
The data does not say that 53% have a conviction or charge. It says that 27% do. The 26% you miscategorized are people with pending charges. Everyone is innocent until proven guilty.
130.
▲
by
sosodev
9mo ago
I'm not saying that pre-trained only models are useless. They've clearly extracted a ton of knowledge from the corpus. The interface may seem strange because it's not what we're accustom to but they still prove valuable.
131.
▲
by
sosodev
9mo ago
Except it's not an impossible request. If my manager told me "fix this code with no questions asked" I would produce a similar result. If you want it to push back, you can just ask it to do that or at least not forbid it to.
132.
▲
by
sosodev
9mo ago
Are you sure about that? There's a lot of slop on the internet. Imagine I ask you to predict the next token after reading an excerpt from a blog on tortoises. Would you have predicted that it's part of an ad for boner pills? Proba
133.
▲
by
sosodev
9mo ago
I think you're confused about the training steps for LLMs. What the industry generally calls pre-training is when the LLM learns the job of predicting the most probable next token given a huge volume of data. A large percentage of that
134.
▲
by
sosodev
9mo ago
In theory NPUs are a cheap, efficient alternative to the GPU for getting good speeds out of larger neural nets. In practice they're rarely used because for simple tasks like blurring, speech to text, noise cancellation, etc you can get
135.
▲
by
sosodev
9mo ago
Nope. Pretraining runs have been moving forward with internet snapshots that include plenty of LLM content.
136.
▲
by
sosodev
9mo ago
Hallucinations generally don't matter at scale. Unless you're feeding back 100% synthetic data into your training loop it's just noise like everything else. Is the average human 100% correct with everything they write on the
137.
▲
by
sosodev
9mo ago
That idea is called model collapse https://en.wikipedia.org/wiki/Model_collapse Some studies have shown that direct feedback loops do cause collapse but many researchers argue that it’s not a risk with real world data
138.
▲
by
sosodev
9mo ago
He asked the models to fix the problem without commentary and then… praised the models that returned commentary. GPT-5 did exactly what he asked. It doesn’t matter if it’s right or not. It’s the essence of garbage in and garbage out.
139.
▲
by
sosodev
9mo ago
I think you're overestimating how much people care about quality.
140.
▲
by
sosodev
9mo ago
It already exists. Tailwind has had GitHub sponsorships enabled for years but only 5 people have ever given them money that way.
141.
▲
by
sosodev
9mo ago
The paid products Adam mentions are the pre-made components and templates, right? It seems like the bigger issue isn't reduced traffic but just that AI largely eliminates the need for such things. While I understand that this has been
142.
▲
by
sosodev
10mo ago
I’ve spent a little bit of time testing Minimax M2. It’s quite good given the small size but it did make some odd mistakes and struggle with precise instructions.
143.
▲
by
sosodev
10mo ago
I’ve spent some time thinking about privacy and LLMs. I developed the impression that encryption isn’t meaningful in this space. It seems like end to end encryption only truly works when both ends are outside of the system and can manage th
144.
▲
by
sosodev
10mo ago
I'm something of a data scientist for a community college. Most of the problems are social not technical but I am still writing code often enough.
145.
▲
by
sosodev
10mo ago
It’s refreshing to a see single optimistic take in this thread
146.
▲
by
sosodev
10mo ago
Prosecution for insider trading in 2026? I highly doubt that
147.
▲
by
sosodev
10mo ago
I suspect that 2026 will be the year we see a big breakthrough in the use of LLM agent systems. I don’t know what that will look like but I suspect the agents will be doing meaningful research (probably on AI).
148.
▲
by
sosodev
10mo ago
You do know this reads the same as every pessimistic commentary on technology ever, right? So many people were convinced that television was going to fry our brains.
149.
▲
by
sosodev
10mo ago
Nvidia released Nemotron 3 nano recently and I think it fits your requirements for an OSS model: https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B... It's extremely fast on good hardware, quite smart, a
150.
▲
by
sosodev
10mo ago
Did you see this HN submission? https://news.ycombinator.com/item?id=46242838 It seems similar to what you're describing.
More ›