6 ms·
I think this is more a symptom of the problem than the actual problem. The issue is that we can generate tons of code using AI, but then are blocked on having
by dlevine 2mo ago
I think this is more a symptom of the problem than the actual problem.
The issue is that we can generate tons of code using AI, but then are blocked on having humans review all of it.
I don’t think we should auto-approve all of this code without human review - that clearly doesn’t work either.
What I do think we need is probably at least two-fold
1) better ways to explain these big PRs to human reviewers.
2) better ways to verify the functionality of a piece of code. Things like auto generating walkthrough videos
I’m not sure that even this is enough. I’m sure there will be agents that try to solve this problem.
- bigstrat2003 2mo agoOr people could stop using LLMs to generate huge swaths of code, and companies could discipline those who refuse to stop. It's providing negative value at this point.
- Arainach 2mo ago> Things like auto generating walkthrough videos I'm trying to write this with respect, but please explain your thought process here. If your PR description is a video instead of a written explanation, I'm rejecting it without even reading the code.
- dlevine 2mo agoIt’s not an either/or. It’s doing both. The video makes it easier to understand what the code does. We have used Looms on PRs since before AI, and agents are beginning to be able to do this on PRs they generate.
- ArnaudDebray 1mo agoAgree with you. Going even further, I'm wondering whether the Diff alone is still the right review artifact... On your 1), I see Entire.io is trying to build some useful building blocks (capturing all agent sessions and prompts to extract the human intent). With a friend, we're currently working on a tool to ease the review using entire's agent session