3 ms·
The replies are all a variation of: "You're using it wrong"
by kykat 1y ago
The replies are all a variation of: "You're using it wrong"
- motorest 1y ago> The replies are all a variation of: "You're using it wrong" I don't know what you are trying to say with your post. I mean, if two persons feed their prompts to an agent and while one is able to reach their goals the other fails to achieve anything, would it be outlandish to suggest one of them is using it right whereas the other is using it wrong? Or do you expect the output to not reflect the input at all?
- ares623 1y agoI expect the $500 billion magic machine to be magic. Especially after all the explicit threats to me and my friends livelihoods.
- motorest 1y ago> I expect the $500 billion magic machine to be magic. Especially after all the explicit threats to me and my friends livelihoods. That's a problem you are creating for yourself by believing in magical nonsense. Meanwhile, the rest of the world is gradually learning how to use the tool to simplify their work, being it helping onboard onto projects, doing ad-hoc code reviews, serving as sparring partners, helping with design work, and yes even creating complete projects from scratch.
- theshrike79 1y agoIt's giving "I paid $200k for this RV, the cruise control should keep the car on the road while I go in the back to make coffee".
- kykat 1y agoOf course the output reflects the input, that's why it's a bad idea to let the LLM run in a loop without constraints, it's simple maths, if something is 99% accurate, after 5 times is 95% accurate, after 10 steps it's about 90% accurate, after 100 times it's about 36% accurate. For LLMs to be effective, you (or something else) needs to constantly find the errors and fix it.
- ninetyninenine 1y agoI’ve seen LLM catch and fix their own mistakes and literally tell me they were wrong and that they are fixing their self made wrong mistake. This analogy is therefore not accurate as error rate can actually decrease over time.
- kykat 1y agoIf we assume that each action has 99% success rate, and when it fails, it has 20% chance of recovery, and if the math here by gemini 2.5 pro is correct, that means the system will tend towards 95% chance of success. === In equilibrium, the probability of leaving the Success state must equal the probability of entering it. (Probability of being in S) * (Chance of leaving S) = (Probability of being in F) * (Chance of leaving F) Let P(S) be the probability of being in Success and P(F) be the probability of being in Failure. P(S) * 0.01 = P(F) * 0.20 Since P(S) + P(F) = 1, we can say P(F) = 1 - P(S). Substituting that in: P(S) * 0.01 = (1 - P(S)) * 0.20 0.01 * P(S) = 0.20 - 0.20 * P(S) 0.21 * P(S) = 0.20 P(S) = 0.20 / 0.21 ≈ 0.95238
- ninetyninenine 1y agoThat’s math based off of arbitrary initial assumptions. There are numbers that work. All this math is useless. Use your brain. The entire point I’m communicating is that it’s not a given that it must become less accurate. There are multiple open possibilities here and scenarios that can occur. Doing random math here as if you’re dropping the mic is just pointless. It doesn’t do anything. It’s like making up a cosmological constant and saying the universe is collapsing look at my math.
- darkwater 1y agoI saw them too. And after that, slip in another mistake.
- magicalhippo 1y agoI've had good experience getting a different LLM perform a technical review, then feed that back to the primary LLM but tell it to evaluate the feedback rather than just blindly accepting it. You still have to have a hand on the wheel, but it helps a fair bit.
- Alex_L_Wood 1y agoAnd yours is also "you are using it wrong" in the spirit. Are they doing the same thing? Are they trying to achieve the same goals, but fail because one is lacking some skill? One person may be someone who needs a very basic thing like creating a script to batch-rename his files, another one may be trying to do a massive refactoring. And while the former succeeds, the latter fails. Is it only because someone doesn't know how to use agentic AI, or because agentic AI is simply lacking?
- berkes 1y agoAnd some more variations that, in my anecdotal experience make or break the agentic experience: * strictness of the result - a personal blog entry vs a complex migration to reform a production database of a large, critical system * team constraints - style guides, peer review, linting, test requirements, TDD, etc * language, frameworks - quick node-js app vs a java monolyth e.g. * legacy - a 12+ year Django app vs a greenfield rust microservice * context - complex, historical, nonsensical business constraints and flows vs a simple crud action * example body - a simple crud TODO in PHP or JS, done a million times vs a event-sourced, hexagonal architecrtured, cryptographical signing system for govt data.
- throwawayb2025 1y agoI had both good and bad experience. Bad with regex or things involving recursion. Also bad at integration between modules. I have not tried to solve this yet, by giving documentation of both the modules. Also model used impacted. To understand java code it was great. I first ask it to generate the detailed prompt by giving a high level prompt. Then use the detailed prompt to execute the task. Java code size Upto 20k loc is fine. Other wise context becomes big. So you have to do module by module. I believe to have a discussion maybe someone has to take an open source code example and then say it doesn't work. Other people can then discuss and decide. Overall happy with gpt5 and claude code. Edit:updated bad at integration
- motorest 1y ago> Also bad at integration between modules. I have not tried to solve this yet, by giving documentation of both the modules. You should first draft the interface and roll out coverage with automated tests, and then prompt your way into filling in the implementation. If you just post a vague prompt on how you want multiple modules workinh together, odds are the output might not met implicit constraints.
- leptons 1y agoIn my experience it depends on which way the wind is blowing, random chance, and a lot of luck. For example, I was working on the same kind of change across a few dozen files. The prompt input didn't change, the work didn't change, but the "AI" got it wrong as often as it got it right. So was I "using it wrong" or was the "AI" doing it wrong half the time? I tried several "AI" offerings and they all had similar results. Ultimately, the "AI" wasted as much time as it saved me.
- l1ng0 1y ago[dead]
- timschmidt 1y agoI've certainly gotten a lot of value from adapting my development practices to play to LLM's strengths and investing my effort where they have weaknesses. "You're using it wrong" and "It could work better than it does now" can be true at the same time, sometimes for the same reason.
- thundoe 1y agoWhich is true. Like launching a Ferrari at 200mph without steering doesn’t take anyone anywhere, it’s just a very painful waste of money
- ZeWaka 1y agoI find it quite funny that one of the users actually posted a fully AI-generated reply (dramatically different grammar and structure than their other posts).
- gwd 1y agoExactly one of two things is true: 1. The tool is capable of doing more than OP has been able to make it do 2. The tool is not capable of doing more than OP has been able to make it do. If #1 is true, then... he must be using it wrong. OP specifically said: > Please pour in your responses please. I really want to see how many people believe in agentic and are using it successfully So, he's specifically asking people to tell him how to use it "right".
- geldedus 1y agoYes. Because it is the correct answer.