3 ms·
People are generally bad at judging themselves, and for some reason even worse about judging their agents' output. I've heard a million people say their LLM wo
by airport_barfly 3mo ago
People are generally bad at judging themselves, and for some reason even worse about judging their agents' output.
I've heard a million people say their LLM workflow produces amazing code. But I have never actually seen the amazing code allegedly being produced. Where are the people saying "I love when my coworkers send me AI-generate PRs"?
- purple-leafy 3mo ago“Amazing code”… who cares? that’s like saying a book is only worth reading if each sentence is beautiful To me I couldn’t care less as long as the overall story is coherent, and I care more about the idea density in the book. Lord of the rings is a great book because of the ideas, not because Tolkien is the best writer ever. Same with AI. I couldn’t care less if some of its code is crap, if it is dense with ideas and generally good execution. If it writes crap code sections I can “beautify” myself. Plus LLMs are only going to improve… they have improved so much since ChatGPT 3. It’s insane. And the arguments saying they haven’t can’t be taken in good faith