4 ms·
I'd %100 prefer an opus 4.8 rewrite over %99 of the time. Unless Fabrice Bellard is rewriting the stuff I need, I'd prefer AI over a human coder.
by bozdemir 3mo ago
I'd %100 prefer an opus 4.8 rewrite over %99 of the time. Unless Fabrice Bellard is rewriting the stuff I need, I'd prefer AI over a human coder.
- raincole 3mo agoOr, you know, you can use Postgres. It's right there for you.
- bozdemir 3mo agowhy? if a rewrite is better/faster/secure, why not? (I'm not saying PGrust is better, I didnt even install it, my perspective is in general)
- OtomotO 3mo agoAI is an average coder. It was trained on all code the code that could be found. Not just code written by genius programmers like Carmack and Bellard. Given that it's average, I'd prefer a human coder above average :)
- piker 3mo agoWhich you will necessarily have if they’ve completed a Rust rewrite.
- bigupthewhole 3mo agoYou haven't been using AI extensively I presume... I've been programming a long time and considered myself among the top in my domain and AI agents using like GPT 5.5 etc. are much better than me.
- OtomotO 3mo ago> You haven't been using AI extensively I presume... Ex falso quodlibet > I've been programming a long time and considered myself among the top in my domain I am not trying to attack you, but you considered yourself that... I don't know whether you actually were and frankly I don't care.
- witx 3mo ago[flagged]
- rytill 3mo agoLLMs learn a distribution during pre-training, not only an average. Then, by giving them context or by post-training, you can make them sample non-average parts of the distribution they learned.
- OtomotO 3mo ago> Then, by giving them context or by post-training, you can make them sample non-average parts of the distribution they learned. How do you derive that something is "below average" or "average" or "above average"?
- rytill 3mo agoWell, it’s up to the user or post-trainer of the LLM what they believe to be above average. Then they can design around that. In the case of real world LLMs and post-training, what is above average is defined roughly as: labeled good by expert humans, and scoring high on RL environments related to coding like debugging, passing tests, or running efficiently and verifiably correctly.
- nextaccountic 3mo ago> How do you derive that something is "below average" or "average" or "above average"? One technique is RLHF: have an human expert assess it.
- OtomotO 3mo agoMhm, I just wonder how many samples they get and how much time they have to come to the conclusion. Like a short example is easier to grade, but not in the same ballpark as a whole codebase.
- u8080 3mo ago>How do you derive that something is "below average" or "average" or "above average"? How do you? I mean, that was your point basis.
- bozdemir 3mo agoI dont think Opus 4.8 is an average coder, with my own experience (I have coded 20 + years before even llms existed) it is anything but average. I don't think training data alone determines the success of these models, there are lots of reinforncement learning principles and fine tuning takes place, a crappy code in the dataset doesnt hold those llms scoring high in benchmarks, I dont think an average programmer can score 70% (opus 4.8) in SWE Bench Pro, which is a good one.
- jeltz 3mo agoI think Opus 8.4 is a below average developer, but maybe I have just worked with good developers and have a skewed perspective of what the average is.
- hoppp 3mo agoI would say it's an average coder when it comes to writing functions because it keeps using regex. It might pass a benchmark but doesn't pass the smell test.
- esafak 3mo agoThat is not how it works. IF it was condemned to be average models wouldn't be constantly improving, given that humans aren't getting better.