Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
XCSme
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
121.
▲
by
XCSme
2mo ago
Yes, I was surprised to see doing it as well as Qwen 3.7 27b. Even though that model is already "old", qwen was way ahead everyone else in that size category before this Meta model. Also, probably for non-Chinese usage, using a no
122.
▲
by
XCSme
2mo ago
I should add a F.a.q. for this question. The suite is across many categories, not only coding, and most of the tasks are low-horizon (or what the opposite of long-horizon is), where the max thinking time is around 10 minutes. Gemini models
123.
▲
by
XCSme
2mo ago
I don't think it matters, if it's for local/on-device usage. The cost is similar vram footprint I guess (?)
124.
▲
by
XCSme
2mo ago
The new Meta 30B models seems A LOT better: https://aibenchy.com/compare/meta-muse-glimmer-30b-xhigh/nvi...
125.
▲
by
XCSme
2mo ago
As a solo dev, this gives me hope. I feel like I have an advantage over big companies, if I can use the best models on a subscription and not worry about costs much, when they can't do the same as outlined in the article.
126.
▲
by
XCSme
2mo ago
My question was more about more complex problems, which no seem to be multi-turn somehow, or maybe just the harnesses make it look that way. I am curious what the drop in thoughput is for multi-turn answers, instead of one-shot. More in lin
127.
▲
by
XCSme
2mo ago
Yeah, makes sense, if it's good for very small models only, then there's no point, as those van already run on cheap consumer hardware. Yet, maybe it can work well enough, so that as a manufacturer, you don't pay $50 for a PI
128.
▲
by
XCSme
2mo ago
I asked a LLM after posting my comment, to see if I had a genius idea or not,just for it to tell me the same as you, that's now they work already...
129.
▲
by
XCSme
2mo ago
My concern is that reasoning could involve some sequential steps that instant models don't. Not sure if modern models "think" only by outputting <thinking> blocks, or there is a more complex mechanism at play.
130.
▲
by
XCSme
2mo ago
It was just a random example, you could think of it as being a lot more complex (detect which type of food it is, what detergent to use, how much water, remember patterns, learn over time, adapt, etc.)
131.
▲
by
XCSme
2mo ago
So local personalized ads? Not sure if that's better or worse than online personalizaed ads...
132.
▲
by
XCSme
2mo ago
Why not have some a device/hardware that programs itself on-boot. Sort of a FPGA, that (electrically) arranges the connections on-boot, and then it's like a static inference chip.
133.
▲
by
XCSme
2mo ago
I think this would make sense for consumer hardware, not for AI companies. AI companies constantly update/change stuff, new models come out, new requirements, etc. But if you ship an "ai-powered" dishwasher, it can come with
134.
▲
by
XCSme
2mo ago
Wait, is it even thinking? Or is it an instant model?
135.
▲
by
XCSme
2mo ago
Wow, that's instant, crazy.
136.
▲
by
XCSme
2mo ago
I am surprised that they keep going with it, seeing how fast it improves and basically soon running themselves too out of business. What's even their end goal? Open source models make sense, if profit is not the target, but for OpenAI
137.
▲
by
XCSme
2mo ago
If they trained it well, and can do computer use, it will be a new era. Companies can keep PCs, put Qwen 3.8 27b on it and get rid of the employees, lol...
138.
▲
by
XCSme
2mo ago
Can't really use it now, without giving away your data: > Trains: this provider may use prompts for training and may retain prompt data.
139.
▲
by
XCSme
2mo ago
Or maybe not: https://news.ycombinator.com/item?id=49119559
140.
▲
by
XCSme
2mo ago
I think that with this change, DeepSeek v4 Flash has finally been dethroned.
141.
▲
by
XCSme
2mo ago
Wasn't one of the main original points of LLMs to be creative? To create stories, creative writing?
142.
▲
by
XCSme
2mo ago
Is this like Runescape free armour trimming? You send your gpu, and get it back with 2x memory?
143.
▲
by
XCSme
2mo ago
Depends on whether the models report the correct amount of tokens. 5.5 Sol repors 10x fewer reasoning tokens than Kimi k3. If it is correct, than it unlikely has those doubt issues. At the same time, I feel like their reporting is incorect
144.
▲
by
XCSme
2mo ago
Yeah, they thing forever and doubt everything "wait but" for 200k tokens for almost any question.
145.
▲
by
XCSme
2mo ago
So nowadays the hardware and hosting providers must be in an optimization race, whoever can make the model just a bit smaller or more efficient (to fit on fewer/less powerful cards) will have a huge advantage and can make a lot of mone
146.
▲
by
XCSme
3mo ago
One of the best hamsters [0]. Again, their "none" version costs more than "low", and says zero reasoning tokens, makes no sense[1]. As always, the "low" version seems to be the best price/perf ratio for fa
147.
▲
by
XCSme
3mo ago
Twice the cost for 4% more intelligence, is it worth it?
148.
▲
by
XCSme
3mo ago
I have a spare 3090 that I want to use to off-load some tasks from Claude to a local model (probably Qwen 3.6 27b), any success with that? Is it good enough to follow some tasks, coding requirements or browser usage?
149.
▲
by
XCSme
3mo ago
I thought login-protected apps are not allowed on Show HN.
150.
▲
by
XCSme
3mo ago
Shouldn't OpenAI be legally responsible for "hacking" another system/company? If I ask GPT-5.6 Sol to hack a website using their work/servers features, who is responsible? Maybe my request was accidental, or it was
More ›