Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
BoredomIsFun
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
BoredomIsFun
7d ago
31B
2.
▲
by
BoredomIsFun
7d ago
Not sure if GPT based LLMs are polynomial time.
3.
▲
by
BoredomIsFun
7d ago
> It's impossible for finite number of LLMs to solve all theorems. This would imply that the busy beaver sequence is computable which implies the halting problem is decidable LLMs use RNG for sampling, so they are not pure computers
4.
▲
by
BoredomIsFun
7d ago
exactly. Hindsight is 20/20.
5.
▲
by
BoredomIsFun
7d ago
> t still have the same AI uniformity and smell as the crappy restaurant poster you see on the street You have really try hard to see that. It is good enough to not give immediate annoyed reflex of seeing something in the same ubiquitous
6.
▲
by
BoredomIsFun
13d ago
ehhh....to be pedantic, LLMs, like any NNs derive power from _not_ being pure matrix multiplication - they are "halved" matrix algebra, due to ReLU
7.
▲
by
BoredomIsFun
13d ago
> The models have not plateaued, and they are not even mildly close to any sort of ceiling. Depends on defnition of "plateaued" and "ceiling". I am not impressed with 2026 consumer models at all. > This is also com
8.
▲
by
BoredomIsFun
15d ago
It does not if you switch swapp off and use zram instead. I am typing right now on such a setup wityh 16 GiB ram and it occasionally, once a week or so, kills my firefox due to oom. If you are you using disk swap - not sure why would if you
9.
▲
by
BoredomIsFun
16d ago
Apple mnitors for whatever reason use ancient 450nm backlight, which strains my eyes like no tomorrow. All modern 5k and 6k use 455-460nm backlight. Otherwise, yes, APples are good.
10.
▲
by
BoredomIsFun
16d ago
There is a plenty of Chinese lower quality 28 3:2s
11.
▲
by
BoredomIsFun
16d ago
> Built-in hardware calibration is non-negotiable for color critical work No it is mostly a convenience gimmick. Support for external hardware calibration is important though. But far more important feature, implemented properly only by
12.
▲
by
BoredomIsFun
17d ago
> Could you explain how you define hallucination in the context of creativity Mostly as you've mentioned, "lack of consistency", which manifests in variety of ways - presence of cellphones in historical settings (I'd
13.
▲
by
BoredomIsFun
17d ago
You can "great learning experience" about the human nature from a good fiction book too. Classics have another function - being cultural landmarks, one can refer to in non-fictional contexts as well.
14.
▲
by
BoredomIsFun
17d ago
LLMisms are almost universally results of using a narrow set of mainstream offerings from Anthropic and Openai. Even Muse Spark has style much less "sloppy" than Claude et al, let alone Chinese LLMs. I mean yes, they have their ow
15.
▲
by
BoredomIsFun
17d ago
I tried at it creative writing - and, with thinking off, it was considerably better than Mercury 2 and generally good in fact, not very sloppy. Now with thinking on, it got worse, began hallucinating things; this is something I've noti
16.
▲
by
BoredomIsFun
18d ago
Netflix is just a single point. I'd never run current in production.
17.
▲
by
BoredomIsFun
18d ago
You can generate images on 5060ti, it'd take less than minute per image. Trivial environmental footprint.
18.
▲
by
BoredomIsFun
18d ago
> Also, real businesses use -CURRENT, everybody knows that. No, not really.
19.
▲
by
BoredomIsFun
18d ago
> seen as the inhibitors of progress I become the government institution, the inhibitor of progress. What a load of delusion.
20.
▲
by
BoredomIsFun
18d ago
> Mistral Small 4 is way worse than Gemma 4 26B A4B Depends for what purpose? I found large Mistrals are massively better than Gemma 4 at creative writing: have more natural tone, better consistency than 26B as it is MoE.
21.
▲
by
BoredomIsFun
18d ago
Mistral is odd. They have made mostly flops, boring models (Ministral 3, Mistral Small 4, Small 3, Small 3.1) together with a classic masterpiece Mistral Nemo and very good Mistral Large 2407, Mistral Small 22b, Mistral Small 3.2.
22.
▲
by
BoredomIsFun
18d ago
> They are regulationmaxxing instead of benchmaxxing, that's my problem with them. To those who is in know (r/localllama, r/sillytavernai), is well aware that Mistral models - at least the small, <=24b ones - are the le
23.
▲
by
BoredomIsFun
19d ago
> that almost all Amish use one. There are small pedal powered ones. I am sure you can rig a horse driven washing machine...
24.
▲
by
BoredomIsFun
19d ago
> short story writers esp. among sci-fi, as sci-fi is more about concept than execution,
25.
▲
by
BoredomIsFun
19d ago
Yes surely. I like ML, >D>S and such but lacking in stats background I wishh I had.
26.
▲
by
BoredomIsFun
22d ago
Yep, an old idea, that has long, long been known in local LLM community - it was achieved by "self-merging". One of the latest, most succesful examples is a self-merge of Microsoft Phi4-14b into Phi4-25b. Some people at r/Loc
27.
▲
by
BoredomIsFun
22d ago
Looped transformers are an old idea, has long, long been known in local LLM community - it was achieved by "self-merging". One of the latest, most succesful examples is a self-merge of Microsoft Phi4-14b into Phi4-25b. Some people
28.
▲
by
BoredomIsFun
23d ago
> (not even 3.8...) 3.6 is better for non-coding tasks, noticeably so.
29.
▲
by
BoredomIsFun
23d ago
It does not matter really - stiffness does lower up to T=0.7, then platoes; even at high temperatures tics/slop-patterns are still there.
30.
▲
by
BoredomIsFun
23d ago
> Try tinkering with a base model and you'll be surprised how diverse it is. Have you tried? I have. Not much different from RLHFed; full of tics and slop, similar but slightly different from intsruction posttrains.
More ›