Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mft_
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
91.
▲
by
mft_
3mo ago
Thanks for flagging. From the few benchmarks I can find, it looks there or thereabouts with Qwen 3.6-35B-A3B, or maybe a touch below. I'm interested to compare a model that is a big jump larger with pretty impressive benchmarks, but
92.
▲
by
mft_
3mo ago
Looks impressive, and this size fits achievable home hardware. That said, if someone would kindly quantise this down for the 64GB paupers, that would be appreciated. (I know there’s likely degradation, but some people reported good results
93.
▲
by
mft_
3mo ago
Alternatively, if someone wants cameras (say, for monitoring of a property) but without a subscription or privacy concerns, does anyone have a recommendation?
94.
▲
by
mft_
3mo ago
I think the premise is wrong: I don't think Google is more hated than all of the organisations you list. Within the tech world, I believe people generally dislike meta, Palantir, Oracle, and Microsoft more than Google. Out in the rea
95.
▲
by
mft_
3mo ago
There is a Gemma 4 model with 12B parameters which might be worth trying. e.g. https://huggingface.co/mlx-community/gemma-4-12B-it-qat-4bit That said, your computer will still get hot!
96.
▲
by
mft_
3mo ago
Not from a single data point. You have no way of knowing whether this is, on the one hand, a single reasonable request or, on the other, the start of a pattern of bureaucratic pain.
97.
▲
by
mft_
3mo ago
If your entire business decision is driven solely by an emotional response to Hetzner having a policy that needs ID, maybe you’re not the right person to be making these sort of decisions?
98.
▲
by
mft_
3mo ago
I think everyone is hoping this! It would be great if they'd release an MoE model somewhere between the 35B size of 3.6 and the 122B version of 3.5 - it could be a great balance of speed and ability for people with reasonably powerful
99.
▲
by
mft_
3mo ago
I tried this on my Mac soon after launch and it was consuming a significant amount of processor cycles even just sitting idle in the menu bar. (From memory, ~20% of an M1 Max.) It may have been an early issue but with no obvious way to inte
100.
▲
by
mft_
3mo ago
I’m really glad to hear this! A while ago, I auditioned about 10 different STT apps on my Mac, with this realtime/streaming transcription as a goal. I failed to find that feature in an app I was happy with, but settled on Handy as the
101.
▲
by
mft_
3mo ago
I had the same problem with (I think) a Hero 3 over a decade ago; given snow sports is a key use case, this is poor.
102.
▲
by
mft_
3mo ago
Same here, in Germany. Some categories are especially exaggerated: when needing a number of new wardrobes a few years ago, we struggled to find anything that was remotely similar in price or value to IKEA. The big retail park furniture sto
103.
▲
by
mft_
3mo ago
> You just need a sufficiently self-interested actor that sees open ecosystems as a necessary part of reducing their own risk profile, relative to the alternative of complete reliance of a third-party business that can take an exorbitant
104.
▲
by
mft_
3mo ago
Open models are probably also comparatively astronomically expensive to train - just less so than the frontier models because they’re somewhat smaller, +/- the creators are more incentivised to focus on getting more from less compute b
105.
▲
by
mft_
3mo ago
Yes. If we reduce back to the “LLMs are next word prediction algorithms” and they have a huge training corpus including positive and negative human interactions, it’s not crazy to think they’ll be influenced by the flow of those learned int
106.
▲
by
mft_
3mo ago
Great to see more competition in this space, and especially from Europe, but... it's a shame when the benchmarks don't include the current best comparable models. They shows results against Qwen 3.5 and Gemma 3, but Qwen 3.6 and
107.
▲
by
mft_
3mo ago
Really nice visuals! Could you give a little more information about the 'stack' you're using to display the clock? From this: > This minicomputer will power the clock. The Pi 3B+ seems just enough to render some simple ani
108.
▲
by
mft_
3mo ago
"No generation, no estimates - just token counts:" "Authoritative: it is the same count Anthropic bills against." "This reframes a headline that looked like good news."
109.
▲
by
mft_
3mo ago
> Would we do a trip like this again? It's certain a lot of travel. We weren't very spontaneous - most of the trip was planned out way in advance, along with hotels. Having 2-4 days in each place is like taking a series of mini
110.
▲
by
mft_
3mo ago
A tangent, but for full clean removals of apps in MacOS (because I'm old-school and despite having plenty of GBs of storage, I hate the idea of dregs lying around) I've had success with AppCleaner[0] and Pearcleaner[1]. Pearcleane
111.
▲
by
mft_
3mo ago
Agree. It's a fairly minimal list with very few extras added. Current /context on a fresh session (compare to that above) is: Opus 4.8 15.8k/1m tokens (2%) System prompt: 4.5k tokens (0.4%) System tools: 7.9k token
112.
▲
by
mft_
3mo ago
The main ones missed immediately were web access/search. Then the to-do list features (it was a nice surprise to try OpenCode and see this working immediately.). There were a couple of other niggles but it was a few months ago. Also,
113.
▲
by
mft_
3mo ago
Apologies, you're right - I used imprecise terminology. The entire initial JSON structure that was sent from Claude Claude to the LLM at the start of a session was 162k. This included the system prompt together with a list of tools (
114.
▲
by
mft_
3mo ago
> It's easy to add using plugins. Sure, but you have to add almost everything, no? It deliberately only comes with read, write, edit, and bash. My point wasn't that you can't add stuff, but that I'd just rather use
115.
▲
by
mft_
3mo ago
Would you mind sharing?
116.
▲
by
mft_
3mo ago
Maybe related to this minimalism, Pi doesn't come with most of the tools an LLM needs to function efficiently or effectively. I get that a blank slate is the paradigm, and you can add whatever you want, but it's too blank IMO.
117.
▲
by
mft_
3mo ago
Early on in experimenting with local models, I found that hooking them up to Claude Code worked very well, but it was also really slow. I used mitmproxy (setup assisted by Claude, natch) to capture Claude Code's entire initial system
118.
▲
by
mft_
3mo ago
I’ve been wondering about something similar - a system that enforces (or does the heavy lifting) of dividing a large task into smaller sub-tasks so that it’s easy to run/check/test each one independently - even on a fresh model in
119.
▲
by
mft_
3mo ago
Let me have one more go. I'm 100% not trying to trigger or upset you. This is the chain of logic I'm trying to argue for, in very clear blunt terms. 1. Germany's economy and future is threatened by its demographics, and uni
120.
▲
by
mft_
3mo ago
This post can essentially be distilled down to: yes, Fable's classifier (which is meant to downgrade cybersecurity, biology, or jailbreak attempts to Opus 4.8) is definitely overly sensitive to the point of uselessness. e.g. a colleagu
More ›