Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Systemerror7A69
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
Systemerror7A69
7d ago
This is just speculation on my part, but LLMs work best when they get immediate, verifiable feedback on their task, and the kind of physical optimizations they mean might not give that to LLMs.
2.
▲
by
Systemerror7A69
8d ago
I'd really love to see more studies about effectiveness of AI in general. As in, what works best and how to use it and such. Because I feel that the technology and space is - so - hyped and fast moving that a lot of cultish feeling rit
3.
▲
by
Systemerror7A69
8d ago
I have to say, I am starting to hate this line of reasoning. Yes, LLMs move extremely fast and a lot of improvements are done in a short amount of time. And there might be a point to these arguments, vaguely. However: There never seems to b
4.
▲
by
Systemerror7A69
8d ago
AMD 7900 XTX with Vulkan here as well, wasn't faster on my test either. Might be much different on Nvidia though. I assume the limit for me is memory bandwith, as the 7900 XTX has the same bandwith as the 3090 from what I can gather an
5.
▲
by
Systemerror7A69
8d ago
My approach to this problem is to just...not try them all. As long as the model you're using solves the problems you have to your satisfaction, there is no need to try any other models, except for financial reasons maybe. So I start wi
6.
▲
by
Systemerror7A69
9d ago
So, alignment does need to be taking seriously, you're right. But keep in mind this is a report from OpenAI about OpenAI, who have a financial incentive to present this in a certain light. Take these things with a grain of salt. This d
7.
▲
by
Systemerror7A69
15d ago
It's absolutely amazing to see this harness. I've been on the lookout for something like this for a while now. I've used both Pi and Maki in the past but was unhappy with certain aspects for both. Pi is not respecting XDG and
8.
▲
by
Systemerror7A69
18d ago
And my wallet is thought to contain a billion dollars, so long as I don't open it. Seriously "thought to be" is such a baseless statement. Thought to be by whom? And on what basis?
9.
▲
by
Systemerror7A69
22d ago
To be honest, I believe I get the point the article is trying to make, and to an extent I agree, but I also think the point is not really made very well. The core of the argument as I understood it is that LLMs aren't just using existi
10.
▲
by
Systemerror7A69
25d ago
Even as someone using AI on the regular I'm starting to hate the "You didn't actually use this exact most expensive model so your point is invalid" argument. This is fair to say if someones last experience with AI was co
11.
▲
by
Systemerror7A69
1mo ago
I think the assumption here that might not hold is simply that increases in efficiency and smaller size will be achieved by linearly just training smaller models better. You are absolutely right that there is a physical limit about these th
12.
▲
by
Systemerror7A69
1mo ago
One of the big problems is that no one actually does read these scripts. You could say "Oh but it's their own fault, duh" but theres a very legitimate argument to be made users going the path of least resistance and that you
13.
▲
by
Systemerror7A69
1mo ago
Yes, but since we are specifically talking about claude.md, Anthropic themselves claim on their website these files "serve[s] multiple purposes: providing architectural context, ..." ( https://claude.com/blog/
14.
▲
by
Systemerror7A69
1mo ago
Theres a study from earlier this year which suggests the opposite: https://arxiv.org/abs/2602.11988 Most of the agents.md and what people use it for / write into is does, in fact, not make a difference. Now, sure,
15.
▲
by
Systemerror7A69
1mo ago
Since it seems like this not only improved sizes but also performance I can't wait for some benchmarks and comparisons. If you don't have a separate GPU for inference, every single GB matters so a comparison between specific Q4 Qu
16.
▲
by
Systemerror7A69
1mo ago
From what I understood in the ticket ( I haven't confirmed independently) it's the later. You can set an environment variable to move the .pi directory ( into your XDG Config dir) but it's both the cache and config combined.
17.
▲
by
Systemerror7A69
1mo ago
Oh sorry - basically just a github ticket to an explanation about the .config folder. Pi doesn't respect XDG_BASE_DIR specification, and even if you use the env var it combines cache and config. The ticket is closed and this will not b
18.
▲
by
Systemerror7A69
1mo ago
Sure but I'm really not that invested in one single harness. I tried out pi because people were recommending it so much. It turns out I personally have some things which annoy me, so I try out others now as well. If I don't find a
19.
▲
by
Systemerror7A69
1mo ago
This has been frustrating me for a while and is part of why I explore other coding agents. As many advantages as pi has in some areas, there are definitely areas where I believe the hype to be a bit overstated. While the config folder is, u
20.
▲
by
Systemerror7A69
1mo ago
Am I wrong or are these evaluations, while interesting, not really meaningful for anyone doing serious development work? I'm asking because I personally only use AI with specific and detailed instructions, building my projects piece-by
21.
▲
by
Systemerror7A69
2mo ago
I'm in a similar situation is you are, having ADHD, and my observations have been similar. I've been more hestiant to use AI at all for a while at the beginning but even now I only ever use it for one thing at a time. It helps me
22.
▲
by
Systemerror7A69
2mo ago
It's also relying on the assumption that the checking LLM only ever corrects wrong statements and never incorrectly "corrects" an already correct statement, which might not always be the case as well.
23.
▲
by
Systemerror7A69
2mo ago
It's not thinking. Not in the way she probably meant. It can "think" that fast the same way a calculator can "think" that fast (kind of). Because it's not human and not "thinking", it's a mathema
24.
▲
by
Systemerror7A69
2mo ago
If they already require your constant supervision the reason is money.
25.
▲
by
Systemerror7A69
2mo ago
I just wish it wouldn't ask me to pipe an install script directly to shell to install. Yes, I can probably inspect that but I do think installing through package managers is the best practice. It looks better than pi with XDG and not b
26.
▲
by
Systemerror7A69
2mo ago
I've absolutely had the same experience. Pi has been praised a lot, people speaking so highly of it's code quality. And then I installed it and found all the problems you mention. Not only that but I read Github Issues about the X
27.
▲
by
Systemerror7A69
2mo ago
I don't think that example applies at all here. The quote you quoted itself said it - "subjects for high art". Theres a difference between treating the banalities of life as SUBJECTS for your art and making human, non-mass pr
28.
▲
by
Systemerror7A69
2mo ago
No, Open weights US models would not break as well - this isn't related to China or USA, it's about Open Weights and the fact that you can download the models.
29.
▲
by
Systemerror7A69
3mo ago
I've heard amazing things about pi and it's effectivenes but when I tried installing it I quickly found out it doesn't respect XDG_BASE_DIRECTORY at all, you need to set some environment variables and the author rejected bot
30.
▲
by
Systemerror7A69
3mo ago
GLM 5.2 is monstrous in size, no one is running that on their own hardware. But it's a very competitively priced model other providers can offer (since it's open) so it's a much cheaper alternative than claude in practice. I
More ›