6 ms·
Hate to pull the skill issue card here, but that is a trivial problem that can be one shotted with almost any model with
by timacles 7mo ago
Hate to pull the skill issue card here, but that is a trivial problem that can be one shotted with almost any model with
- smj-edison 7mo agotbh that's not a helpful thing to say. I think a more productive thing would be to ask "What model are you using?" "Are you using it in chat mode or as a dedicated agent?" "Do you have an AGENTS.md or CLAUDE.md?" I've also been underwhelmed with its ability to iterate, as it tends to pile on hacks. So another useful question is "did you try having it write again with what you/it learned?"
- ErroneousBosh 7mo ago> I think a more productive thing would be to ask "What model are you using?" "Are you using it in chat mode or as a dedicated agent?" "Do you have an AGENTS.md or CLAUDE.md?" In my case I'd have to say "Don't know, whatever VS Code's bot uses", and "no idea what those are or why I have to care".
- smj-edison 7mo ago> Don't know, whatever VS Code's bot uses The reason I ask about what model is I initially dismissed AI generated code because I was not impressed with the models I was trying. I decided if I was going to evaluate it fairly though, I would need to try a paid product. I ended up using Claude Sonnet 4.5, which is much better than the quick-n-cheap models. I still don't use Claude for large stuff, but it's pretty good at one-off scripts and providing advice. Chances are VS Code is using a crappy model by default. > no idea what those are or why I have to care For the difference between chat mode and agent mode, chat mode is the online interface where you can ask it questions, but you have to copy the code back and forth. Agent mode is where it's running an interface layer on your computer, so the LLM can view files, run commands, save files, etc. I use Claude in agent mode via Claude Code, though I still check and approve every command it runs. It also won't change any files without your permission by default. AGENTS.md and CLAUDE.md are pretty much a file that the LLM agent reads every time it starts up. It's where you put your style guide in, and also where you have suggestions to correct things it consistently messes up on. It's not as important at the beginning, but it's helpful for me to have it be consistent about its style (well, as consistent as I can get it). Here's an example from a project I'm currently working on: https://github.com/smj-edison/zicl/blob/main/CLAUDE.md https://github.com/smj-edison/zicl/blob/main/CLAUDE.md I know there's lots of other things you can do, like create custom tools, things to run every time, subagents, plan mode, etc. I haven't ever really tried using them, because chances are a lot of them will be obsolete/not useful, and I'd rather get stuff done. I'm still not convinced they speed up most tasks, but it's been really useful to have it track down memory leaks and silly bugs.
- johnnyanmac 7mo ago>I decided if I was going to evaluate it fairly though, I would need to try a paid product. Okay. Get me a job and I'll pay for any model of your choosing. Until then, finances are very slim.
- smj-edison 7mo ago> Get me a job Heh, I'm a college student, so I can't help with that... You could also try Gemini 3 pro with Gemini's CLI which is free, though it's not as good at using tools. But, it sounds like you're not interested, which is fine! Just please don't continue to argue with finer points if you're not interested. I've done my best to engage with your points, but I get the sense that it doesn't matter what I say. I am curious though, why do you feel so strongly about LLM products?
- johnnyanmac 7mo agoI should note that I'm not the same person that you were talking to you the chain. So I hope we're not mixing conversations and people. I don't think I've said that much in this chain, so I can't answer much. But sure: >why do you feel so strongly about LLM products? Personally, I work in games. So pretty much everything in the discourse of LLMs and Gen AI has been amplified 5x for me. The layoffs, the gamers' reaction to stuff utilizing AI, the impact on hardware prices, the politics, etc. Theres a war of consumers and executives, and I'm trapped in the middle taking heat from both. It's tiring and it's clear who to blame for all of this. I want all of this to pop so the true innovation can rise out, instead of all the gold rush going on right now. Also,game code is very performance sensitive. It's not like a website or app where I can just "add 5 seconds to a load time" unless I'm working on a simple 2D game, nor throw more hardware to improve performance. Even if LLMs could code up the game, I'd spend more time optimizing what it makes than it saved. It simply doesn't help for the kind of software I work with.
- smj-edison 7mo agoCrap, you're right. I swear, tiny usernames is both a boon and a curse... > Personally, I work in games. So pretty much everything in the discourse of LLMs and Gen AI has been amplified 5x for me. The layoffs, the gamers' reaction to stuff utilizing AI, the impact on hardware prices, the politics, etc. > Theres a war of consumers and executives, and I'm trapped in the middle taking heat from both. It's tiring and it's clear who to blame for all of this. I want all of this to pop so the true innovation can rise out, instead of all the gold rush going on right now. That makes a lot of sense. I've been pretty fed up with the hyperbole and sliminess, and I can't imagine how difficult it is to be squeezed between angry gamers and naive and dense executives. When you say "true innovation", is that in terms of non-AI innovation, or non-slimy AI innovation? I guess I personally still believe that LLMs are useful, but only as another tool amongst many others. I'm also a big believer in human centered UX design, and it's kinda sad that the dominant experience is all textual. > Also,game code is very performance sensitive It does seem like game programming is the last bastion of performance, at least in terms of normal hardware, since the game has to go to the consumer's hardware. The "silver bullet" mentality drives me a little crazy because it clearly doesn't work in all situations. Anyways, I don't know if this response really has a point, but I wanted to at least acknowledge your experience.
- timacles 7mo agoAgreed was a bit rough. Yes they are not great at iterating and keeping long contexts, but you look at what he’s describing and you have to agree that’s exactly the type of problem llm excel at Shouldn’t have to baby step through the basics when the author is clearly not interested in learning himself
- smj-edison 7mo ago> Shouldn’t have to baby step through the basics when the author is clearly not interested in learning himself I'd rather assume good faith, because when I first started using LLMs I was incredibly confused what was going on, and all the tutorials were grating on me because the people making the tutorials were clearly overhyping it. It was precisely the measured and detailed HN comments that I read that convinced me to finally try out Claude, so I do my best to pay it forward :)
- timacles 7mo agoI totally agree, and myself have gone through that cycle. But the guy is being adversarial and antagonistic. Its a 2 way street, sometimes you have to call people out on their BS because I'm not seeing someone argue in good faith, but rather pretending some superior knowledge because hes working on a esoteric protocol like people here don't know how packet headers work
- smj-edison 7mo agoI don't read it as superiority, perhaps bitterness would be the closest word to what I'm reading. > sometimes you have to call people out on their BS That's true, but I think that it's often much later than what some people would consider enough. Someone can be bitter, and still have good points. It's very dangerous to preemptively dismiss points, because it means that I won't listen to anyone who disagrees with me. I'm willing to put in the work to interpret someone's response in a productive light because there's often something to find. There's a framework that I work within when I'm in a discussion. There's three elements: arguments, values, and assumptions. An argument is the face value statements. But those statements come from the values and assumptions of the person. Values are what people consider most important. In most cases, our values are the same, which is good! The biggest difference is assumptions. For example, one assumption I have is that free markets are the best method we have to lift individuals out of poverty. This colors how I talk about AI. Another person might assume that free markets have failed, and we need to use a different approach. This colors how they would view AI. So we'll completely talk past each other when arguing AI, because it's more of a proxy war of our assumptions.
- ErroneousBosh 7mo agoOkay, tell you what then. Help me learn. The problem is that I want something that listens on a TCP connection for GD92 packets, and when they arrive send appropriate handshaking to the other end and parse them into Go structs that can be stuffed into a channel to be dealt with elsewhere. And, of course, something to encode them and send them again. How would I do that with whatever AI you choose? I'm pretty certain you can't solve this with AI because there is literally no published example of code to do it that it can copy from.
- timacles 7mo agoGD92 packets? No idea what you’re talking about but if it has a spec then it doesn’t matter if it’s trained on it. Break the problem down into small enough chunks. Give it examples of expected input and output then any llm can reason about it. Use a planning mode and keep the context small and focused on each segment of the process. You’re describing a basic tcp exchange, learn more about the domain and how packets are structured and the problem will become easier by itself. Llms struggle with large code bases which pollute the context not straightforward apps like this
- smj-edison 7mo agoOne other thing, it might be worthwhile having the spec fresh in the LLM's context by downloading it and pointing the agent at it. I've heard that that's a fruitful way to get it to refresh its memory.
- timacles 7mo agoYep you can even extract the relevant parts and put them into local files the llm can scan
- ErroneousBosh 7mo ago> GD92 packets? No idea what you’re talking about but if it has a spec then it doesn’t matter if it’s trained on it. Okay, so you're running into the same problem that LLMs are. > Break the problem down into small enough chunks. Give it examples of expected input and output then any llm can reason about it. So I have to do lots of grunt work? > You’re describing a basic tcp exchange, learn more about the domain and how packets are structured and the problem will become easier by itself I've written dozens of things that deal with TCP. I already have a fully-working example of what I want. The idea was to test if I could recreate it using LLMs. How is it supposed to work? How does it put in the code I already know I want?