Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
logicprog
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
logicprog
4mo ago
Edited that claim, and made several clarifications elsewhere. The whole point of this analysis is that outrage is unjustified on the basis of two totally statistically unremarkable releases that no one would have remarked on pre-AI (my furt
32.
▲
by
logicprog
4mo ago
Fair point. Let me edit (if I still can) to tone it down.
33.
▲
by
logicprog
4mo ago
> This analysis showed that there is indeed an absence of evidence, but it concludes there is evidence of absence. I tried pretty hard to avoid saying that, can you point me at how to rephrase? The point I'm trying to make is just t
34.
▲
by
logicprog
4mo ago
Depends on the methods you use. If you're trying to fit curves and so on, yes. The methods I use were designed for very low amounts of data, and are generally okay for that, specifically and especially when you're just trying to s
35.
▲
by
logicprog
4mo ago
To quote Tridge: > As to all the people saying “I’m going to package openrsync for platform XXX and we’ll use that!”. I find that rather amusing. If you do decide to go down that path I’d suggest you try the new rsync test suite on openr
36.
▲
by
logicprog
4mo ago
And anti-AI people accuse people who use AI of being intellectually lazy. First of all, it's long because it's expanded to respond to all the criticisms. It seems that either something can be short, and dismissed as incomplete, or
37.
▲
by
logicprog
4mo ago
Okay, so you didn't respond to any of my rebuttals — like the double standard between anti-AI and pro-AI claims, one of which gets to make claims based on cherry-picked anecdotes, and the other which must produce rigorous studies — you
38.
▲
by
logicprog
4mo ago
> Also if you write a paper where you get statistical conclusions out of whole 2 datapoints you'd be laughed out of the room I'm using methods appropriate to that low amount of data, first of all. Second of all, since I'm
39.
▲
by
logicprog
4mo ago
That's sort of the point. There isn't enough data to extrapolate, and yet that's exactly what those outraged about AI were doing, and when you do do the very minimal types of analyses (permutation tests, and looking at dist
40.
▲
by
logicprog
4mo ago
Thank you!
41.
▲
by
logicprog
4mo ago
I've now resolved this. The new version, which should be live on GH Pages soon, uses — what I think is — a pretty good methodology for assigning severity to each bug, normalizes it to 0.0-1.0, sums that, and treats that as the total se
42.
▲
by
logicprog
4mo ago
> EDIT: Found it! it is in the (untitled) discussion section (after the results). I also paraphrase Tridge himself explicitly saying that this is why commits/releases have increased: > Essentially, this isn't a "Claude&
43.
▲
by
logicprog
4mo ago
Another update: did an automated severity analysis on each bug report (~2000 of them!) using an LLM at temp=0 with a very strict rubric (and I checked to make sure that it rated things in a consistent, stable way using it). The rubric, LLM
44.
▲
by
logicprog
4mo ago
Great, so I rewrite everything in my own prose, and now it's still "obvious AI writing," just because I'm literate.
45.
▲
by
logicprog
4mo ago
> .. what are those em-dashes doing there though? You're literally doing exactly the bullying I was trying to avoid, even while denouncing it. I like em-dashes . I have AuDHD, and they help me represent how I think.
46.
▲
by
logicprog
4mo ago
I rewrote all the AI prose several hours ago with purely my own. I like em-dashes, and specifically use them with spaces as a habit. I don't know what to tell you.
47.
▲
by
logicprog
4mo ago
I link to it multiple times in TFA and quote the specific thing I'm talking about here in there to explain that possible confounder. I think I've done more than the work I'm obligated to it.do to make all of the relevant inf
48.
▲
by
logicprog
4mo ago
Your first and second points seem to contradict each other because if all of the bugs for 3.4.1 should be attributed to 3.4.0, that pushes the timetable back even further that unattributed LLM commits would have to have been being committed
49.
▲
by
logicprog
4mo ago
I have done so! that was a misremembering on my part. first mention of Lobsters is now here: > On Lobste.rs, in response to the Medium essay Tridge himself posted in response, finally some users like boramalper begin to actually ask for
50.
▲
by
logicprog
4mo ago
Yes, but we know why there was an "extraordinary spike," and it has nothing to do with rsync being "vibe coded." The maintained has directly addressed this.
51.
▲
by
logicprog
4mo ago
> The presence of "The Outlier Nobody Noticed" proves nothing and deserves no more than a passing mention. A random release introduced way more bugs than the Claude-containing releases. That provides evidence that Claude doesn&
52.
▲
by
logicprog
4mo ago
That's interesting; IME, most people get equally angry and are as likely to disengage with a superior tone over my autism-infodump verbose essay prose as with LLM output.
53.
▲
by
logicprog
4mo ago
Thank you for your constructive input, you're one of only a few others here who had any. I'll definitely do that. I didn't think, since the output was templated directly from the numbers generated by a reproducible python scr
54.
▲
by
logicprog
4mo ago
Alright, I'll do that. Although, sadly, I already posted it here, so I won't be able to post it again — I'll be stuck with this trash comments section that doesn't deal with any of the actual claims, just the aesthetics.
55.
▲
by
logicprog
4mo ago
This seems fair. Of course, now that I've posted this here once, I doubt it'll get constructive engagement again, but I can at least improve this for the future
56.
▲
by
logicprog
4mo ago
Okay, I really have to point out to everyone: the numbers and report cards are TEMPLATED IN BY A SCRIPT . Hallucinations are a moot point. https://github.com/alexispurslane/rsync-analysis/blob/main/s...
57.
▲
by
logicprog
4mo ago
Do you think it would help if I went through and manually rewrote all of the prose? If it would get people to listen, I'd be totally willing to do it. It's not like I don't like writing. I just was focused on something else w
58.
▲
by
logicprog
4mo ago
Did you face any actual bugs or regressions? Or are you doing this just because of the bandwagon that's going around right now? Because until you can actually present an argument for why this release is worse than any of the others, wh
59.
▲
by
logicprog
4mo ago
It's not defending itself here, both because I used GLM 5.1, not Claude, and because I was the one who decided to do this analysis, iterated through six or seven different methodologies to try to find the one that was most honest with
60.
▲
by
logicprog
4mo ago
No, I didn't write the text itself. I'm typically significantly more verbose and elliptical, and more than that, the numbers and methodology changed often enough over the course of the last couple days I was working on this becaus
More ›