Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sosodev
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
61.
▲
by
sosodev
4mo ago
Meh. My server can run these models for neglible power draw (like ~130W fully maxed out). That's with ~30 tok/s which isn't that bad. I do agree that they're still nowhere near as good as the frontier models though. I do
62.
▲
by
sosodev
4mo ago
Any artifacts or blogs I can check out? I'm curious how you manage to make them all useful in parallel. I have a hard enough time getting one instance of Qwen3.6-27B being useful full time haha.
63.
▲
by
sosodev
4mo ago
I think this is overselling their capabilities. I've used Gemma 4 and Qwen 3.6 quite a bit on my strix halo home server. They're great models and the dense variants are significantly better, but they're still very far behind
64.
▲
by
sosodev
4mo ago
My strix halo board is feeling more useful and less toylike with the recent performance gains combined from MTP, better quantization, and generalized performance improvements across the stack. For example, I can run Unsloth's Gemma4-31
65.
▲
by
sosodev
4mo ago
The problem with this question is that it encompasses a huge spectrum of capabilities and expectations. If you can only run an 8B model and expect it to be good at vibe coding / one shotting things you're going to have a bad time.
66.
▲
by
sosodev
4mo ago
I think it heavily depends on what you're asking the model to do. Qwen3.6, both 27B and 35B-A3B, do agentic tool use very well. Their decision making is sus, but the dense model is decent in that way. A 4-bit quant for either of those
67.
▲
by
sosodev
4mo ago
I don’t know why you’re getting downvoted. It’s true. Averaged across a wide variety of benchmarks Fable is the only Anthropic model that performs better than GPT 5.5 xhigh.
68.
▲
by
sosodev
4mo ago
I wonder if model distillation will continue to work as well as it has. Given hidden reasoning, the ever expanding number of expected capabilities, a serious compute shortage, the looming possibility of model collapse, and dramatically high
69.
▲
by
sosodev
4mo ago
Do you have any resources to share regarding independent expert training? I was under the impression that it's not feasible.
70.
▲
Software Has Long Been Beyond Our Understanding
(kylemcgough.com)
3 points
by
sosodev
4mo ago
|
1 comments
71.
▲
by
sosodev
4mo ago
Support requests have always been the weakest link in the security chain for big corps. I've had accounts of mine turned over with 2FA disabled by humans before. I guess we shouldn't be surprised that the LLMs are doing the same t
72.
▲
by
sosodev
4mo ago
Most of the examples they've chosen seem.. not good? What an odd mix of bad game engine and AI slop. I can't imagine that this stuff makes good training data for real-world applications.
73.
▲
by
sosodev
4mo ago
I’m not sure I understand your question. Every interaction you have with a model in a web page does the same thing in the backend. It feeds the whole conversation history, perhaps with a bit of processing, into the model so it can process t
74.
▲
by
sosodev
4mo ago
Isn’t this contradictory to your point? They dropped it, collected data, and then reverted when the evidence suggested they made the wrong choice.
75.
▲
by
sosodev
5mo ago
I also cut off JetBrains recently after a long relationship with their tools. I agree with the points made by the author. The tools are clunky resource hogs for seemingly no reason. I was really excited when JetBrains announced Fleet and pr
76.
▲
by
sosodev
5mo ago
Why is nobody on HN talking about this? Unless this is a very advanced fake, it seems like the first proof that humanoid robots are actually capable of real labor.
77.
▲
by
sosodev
5mo ago
If Wall Street was so wise they would only reward meaningful layoffs. Laying off 10% of a company by stack ranking every team accomplishes nothing. Particularly if the company just hires the same number of cut people next quarter. If a tree
78.
▲
by
sosodev
5mo ago
Well put. I too am optimistic that, in the long term, good will prevail and we'll be stronger because of the suffering. I also agree that there's happiness and meaning to be found in presence and local life. However, it feels quit
79.
▲
by
sosodev
5mo ago
Is the US one of the best places for career growth and income? I'm 30. I've been in the tech industry for several years. During the COVID tech boom I would have agreed. I made insane amounts of money for a new grad. Then I was lai
80.
▲
by
sosodev
5mo ago
Does anybody have more insight into the demand for electricians during data center construction? This article is really light on the details. I was researching it recently and got the impression that the majority of electricians hired durin
81.
▲
by
sosodev
5mo ago
I sense that the frustration you feel is that professors are able to make choices based on their values, but the average person is not. That is broadly speaking, of course. I think it is a great shame that we live in a modern world where we
82.
▲
by
sosodev
6mo ago
Yeah, but don’t you agree that less tokens to accomplish the same goal is a sign of increasing intelligence?
83.
▲
by
sosodev
6mo ago
I hope the industry starts competing more on highest scores with lowest tokens like this. It's a win for everybody. It means the model is more intelligent, is more efficient to inference, and costs less for the end user. So much bench-
84.
▲
by
sosodev
6mo ago
I've seen plenty of people look at those metrics and they certainly do tell a story of growing inequality and instability. To me, it seems more obvious that those issues are largely unaddressed by the people in power because they'
85.
▲
by
sosodev
6mo ago
TIL LPCAMM2 exists. What an awesome solution to allow memory replacements while meeting all of the other requirements for laptops.
86.
▲
by
sosodev
6mo ago
I was surprised by the bit about Costco selling the outlet-tier trash. I don't currently have a membership, but I've generally understood their position to be quality at cost.
87.
▲
by
sosodev
6mo ago
Oh, that’s interesting. Thanks for the correction. I didn’t know such heavily post trained models could still do good ol fashion autocomplete.
88.
▲
by
sosodev
6mo ago
These are not autocomplete models. It’s built to be used with an agentic coding harness like Pi or OpenCode.
89.
▲
by
sosodev
6mo ago
I don't know if Next.js, TanStack, etc are more abstract than Rails, Django, etc. They're undoubtedly more complex though. I also find it hard to believe that it's some sort of conspiracy by management to make developers more
90.
▲
by
sosodev
6mo ago
I think the unfortunate truth is the simplest. Web development has long been detached from rationality. People are drawn to complexity like moths to a flame.
More ›