Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
3abiton
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
31.
▲
by
3abiton
3mo ago
It would be so cool as a feature to allow people on HN to make some of their favorited posts/comments public.
32.
▲
by
3abiton
3mo ago
Another crucial difference is: ROCm is open source, while cuda isn't. Yes it's tough to port things to newer gpus, but in theory people did it (therock spear heading the ROCm patch for strix halo before the offical support)
33.
▲
by
3abiton
3mo ago
Honestly, root detection is a cat and mouse game. So many ways to spoof it. Same with Play Integrity. If the goal is to prevent bad actors, it will never work really. It will be just a big headache for the citizens. Look at how gatekeeping
34.
▲
by
3abiton
3mo ago
This was a very enjoyable read, and loved the distribution inclusion, rather than point statistics. I hope there are more tests on different factors like GPUs (nvidia vs amd vs intel), screens, mice, and ultimately windows to see if there i
35.
▲
by
3abiton
3mo ago
Funnily the current high end Mac Studio are not suited for current LLMs. M3 Ultra is "quite an old" chip for AI, despite its bandwidth. The issue for running local models (especially LLMs), you need few things to align really well
36.
▲
by
3abiton
3mo ago
This is technically impressive, but is it usable in practice?
37.
▲
by
3abiton
3mo ago
> I just don't think that I can ever trust an xAI model knowing that they are actively trying to shape its replies to fit a political narrative. How can you trust their models to be reliable in a business setting with the foreknowle
38.
▲
by
3abiton
3mo ago
I still don't fully get the additional value over tmux, beside notification regarding the agent status?
39.
▲
by
3abiton
3mo ago
I run OpenWrt on 2 of my routers. It's really amazing the level of control. Although, I am now building a better approach: a mini PC as a managed linux router replacing my ISP (no wifi). Then my 2 wifi routers for wifi.
40.
▲
by
3abiton
3mo ago
The target audience is different. Coding is mainly a trade of the tech savvy, who like many on r/localllama users do not hesitate to deply on 16GB Vram gpus. Even if so, it is estimated that within 2 years we will be able to run Claude
41.
▲
by
3abiton
3mo ago
The "trick" is well documented in their Deepseek-OCR paper, that builds on plenty of other work. It's just not simple to just switch a commonly used LLM architecture to a new one, but I don't doubt most frontier labs are
42.
▲
by
3abiton
3mo ago
They are heavily bogged down by bandwidth unfortunately. The macs are on another level. If Apple decides to release AI dedicated hardware, it would dominate this space (consumer AI).
43.
▲
by
3abiton
4mo ago
No wonder why deers are seen as snobbish. All the illiterate ones were shot.
44.
▲
by
3abiton
4mo ago
Honestly it's just a hierarchy difference between the two countries. In the US, tech/fin/military companies have the upper hand compared to the government (fragmented between 2 parties). Despite the sharades with Anthropic, T
45.
▲
by
3abiton
4mo ago
> One day, maybe not far from now, a breakthrough will allow huge LLMs (say 200B in size) to run well on an old 5 year old Dell desktop. I think there will be specialized hardware (beside GPUs) that would be custom made for LLMs. Yes TPU
46.
▲
by
3abiton
4mo ago
And a big thing that's missing is ... the harness comparison. Ot plays a very big role. I use forge, and I have been inpressed with what it can do given all the limitations of local models.
47.
▲
by
3abiton
4mo ago
I think nearly everyone mentioned Qwen, so my turn I guess. Qwen 3.6 35B Q8 (MTP), on a Strix Halo, with llama.cpp. Around 40-50 t/s. Really great pefromance, I get always suprised by its capability. I used with forge-code directly in
48.
▲
by
3abiton
4mo ago
I have the same. The difference is, if you do email verification, you will "verified" status. If not, you can still add the company to your linkedin, just unverified, which is not a label.
49.
▲
by
3abiton
4mo ago
Are there evidence that this approach helps maintain "accuracy" performance when quantized? It sounds a bit like mxfp4 with gpt-oss, which was a confusing model upon release.
50.
▲
by
3abiton
4mo ago
Podman has lots of underappreciated features, and it's fully open-source!
51.
▲
by
3abiton
5mo ago
Not to mention the competition: chinese open-weight models and open-source harnesses. Qwen3.6-(27B and 35B) have proven to be worthy and capable of running locally. I am confident more SMEs would look into this as a solution given the ballo
52.
▲
by
3abiton
5mo ago
> * That can still yield useful "discoveries" in certain fields, absent the discovery of new mechanics that exist outside said training data One can argue, new knowledge is just restructured data. I think the main concerns abou
53.
▲
Reverse engineering Android malware from popular Chinese projectors
(zanestjohn.com)
89 points
by
3abiton
5mo ago
|
19 comments
54.
▲
by
3abiton
5mo ago
Step 1: Have a workshop space Step 2: ? Step 3: Profit
55.
▲
by
3abiton
5mo ago
> Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. I second this! Using the Unsloth Q6 (I forgot the exact name). Currently using it with forgecode (with zsh), on my Strix Halo, and it'
56.
▲
by
3abiton
6mo ago
They patched the "non-existent" issue it seems. And totally denied it happened in the first place. Honestly, someone should do a dump of redacted client documents to teach them a lesson. Short of a class action lawsuit would be an
57.
▲
by
3abiton
6mo ago
Even though all can be replaced by a decent mini pc with beefy memory, with lots of VMs.
58.
▲
by
3abiton
6mo ago
> So, the lessons for all other countries in the world is pretty clear: grow yourselves some mountains, dig yourselves a big river, and dam, baby, dam !! You're forgetting corruption. Many countries can easily go 100% renewable, but
59.
▲
by
3abiton
6mo ago
Benchmarking has been already known to be far from a signal of quality for LLMs, but it's the "best" standardized way so far. Few exists like the food truck and the svg test. At the end of the day, there is only 1 way: havin
60.
▲
by
3abiton
6mo ago
> Dr. B is the king of slop, with 84 extensions published, all of them vibe coded. > How do I know? Most of their extensions has a README.md in them describing their process of getting these through addon review, and mention Grok 3. A
More ›