Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
logicprog
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
logicprog
4mo ago
That's not what I'm saying. What I'm saying is that if the criticism is referring to a broad set of metrics like bugs per release and number of commits that were made by Claude, then it's correct to look at precisely tho
62.
▲
by
logicprog
4mo ago
I mean, you can literally clone my repo, run the Python that rebuilds the database and does the whole data analysis and to end from scratch, and verify that the numbers are accurate. I made the code for this analysis public for that exact r
63.
▲
by
logicprog
4mo ago
"Placement" as in where the Claude-driven releases exist within the existing distribution of bugs per 100 commits. If they're not OOD, then nothing is unusual. Also, it wasn't written by Claude FWIW, GLM 5.1.
64.
▲
by
logicprog
4mo ago
Some notes on this: - I used GLM 5.1 to help with the coding and math for this. - However, I explicitly dictated where the data should be pulled from (GitHub, Bugzilla, mailing list), how it should be tagged and grouped, and what data to lo
65.
▲
Did Claude increase bugs in rsync?
(alexispurslane.github.io)
512 points
by
logicprog
4mo ago
|
567 comments
66.
▲
by
logicprog
5mo ago
My understanding is that that's because they were trying to do a structurally homologous port from Zig to Rust, precisely to keep their mental model and not change "too much" at once, and then they plan to refactor to make it
67.
▲
by
logicprog
5mo ago
Well yes, I'm well aware. They're both obnoxiously stupid and annoying, lol.
68.
▲
by
logicprog
5mo ago
Thank you . Finally, someone (albeit with heavy LLM editing...) says it.
69.
▲
by
logicprog
5mo ago
https://en.wikipedia.org/wiki/Moravec%27s_paradox
70.
▲
by
logicprog
5mo ago
That's so obnoxiously stupid and annoying, wow.
71.
▲
by
logicprog
5mo ago
OK, so let me get this straight. First, you ask for evidence of someone who isn't a VIP doing a similarly difficult problem using an LLM, to show that it isn't just VIPs being given special models. And then, when I provide that ex
72.
▲
by
logicprog
5mo ago
@simonw explains how hilariously misguided that paper is in one of the top comments, and how it doesn't apply remotely to a real agent harness. Plus it's not even clearly relevant here, because the model isn't trying to regur
73.
▲
by
logicprog
5mo ago
> Where's that? https://archive.ph/2w4fi
74.
▲
by
logicprog
5mo ago
Looks like he did the maintainability performance and test suite checks and made his decision :)
75.
▲
by
logicprog
5mo ago
> It is always Gowers, Tao and Lichtman (math.ínc startup) who are pushing these technologies. In your mind does this mean that they are lying, or driven by motivated reasoning and cognitive bias, or whatever you'd like to say? Beca
76.
▲
by
logicprog
5mo ago
Yeah. People (Gary Marcus) have been claiming that AI will hit a wall or is hitting a wall or already has hit a wall since 2023, basically. And yet every time they proclaim that the AI industry found new ways of training their AI's, ne
77.
▲
by
logicprog
5mo ago
You can make even lighter weight and just as keyboard driven GUIs. The only downside, as you say, is them not integrating with Tmux.
78.
▲
by
logicprog
5mo ago
Just use something like Tk or wxWidgets.
79.
▲
by
logicprog
5mo ago
In my experience, deep seek models are massively overrated in terms of how good they actually are at agantic usage, coding and writing, just because they are kind of the first open source entrant and the name a lot of people know. Try GLM 5
80.
▲
by
logicprog
5mo ago
RESPONSE TO EDIT: You still haven't even answered my question. Why are you so concerned by there being AI code in the editor that you need this level of trust? The point I have been attempting to make is that needing this level of
81.
▲
by
logicprog
5mo ago
You still haven't explained literally anything. Yeah, okay, if there's a switch, you can't be sure that every single AI related code path is fully disabled. But if you flip the switch and there isn't any AI integration v
82.
▲
by
logicprog
5mo ago
They have a single switch that will remove all AI features from the interface. Why do you need more than that? This is not a rhetorical question. I genuinely don't understand it — if you can get all of those features completely out of
83.
▲
by
logicprog
6mo ago
You're not arguing in good faith here, but just to make this apparent to everyone else: the disclaimers talk about the general case of Erdos problems as a whole. The article explicitly acknowledges them, but then says that the disclaim
84.
▲
by
logicprog
6mo ago
They literally have a quote from Tao in the article saying it was a novel approach humans hadn't tried, and that the problem hadn't been solved even after a lot of professional attention.
85.
▲
by
logicprog
6mo ago
They explicitly say many of these disclaimers don't apply in the article.
86.
▲
by
logicprog
6mo ago
The built in agent supports MCPs, as do the agents you can use through ACP that do already.
87.
▲
by
logicprog
6mo ago
Ah ok, we probably do agree then.
88.
▲
by
logicprog
6mo ago
I don't think saying "as a general rule, data centers don't use remotely enough water to be any kind of significant threat, when you see through the accounting games, media hype, and look at things in a proper context" i
89.
▲
by
logicprog
6mo ago
> Others are clearly exaggerating or making mistakes. I'd be interested to hear a specific example, so I can get a sense for what you mean. > But he jumps from that to the incorrect conclusion that “data centers don’t use water”.
90.
▲
by
logicprog
6mo ago
Would you care to explain in what way his claims have been misleading? Because I have read all of his articles and attacked his math and his sources, and so on, and I haven't found them misleading at all. The biggest way I've seen
More ›