5 ms·
Am I the only one that feels like Claude is clearly winning code generation, and Gemini in general LLM? I just don’t feel like OpenAI has a legitimate shot at
by jader201 6mo ago
Am I the only one that feels like Claude is clearly winning code generation, and Gemini in general LLM?
I just don’t feel like OpenAI has a legitimate shot at winning any of the AI battles.
Therefore, I feel like “Sam Altman may control our future” is a far stretch.
- dominotw 6mo agohow is gemini winning in general llm. what is general llm .
- SwellJoe 6mo agoGeneral LLM is what Apple is paying Google for.
- tartoran 6mo agoI noticed that Apple speech to text has gotten pretty good lately. Is that because they’re paying Google? Not sure I use other AI features from Apple as I have my Siri turned off.
- laserlight 6mo ago> Is that because they’re paying Google? No, the Google deal hasn't shipped yet.
- guelo 6mo agoWell I just canceled my Claude Pro subscription because of the mysterious limits that I don't experience with codex, even after paying for "extra usage". If Anthropic can't figure out their capacity problems they are in trouble.
- chrisjj 6mo agoI doubt Anthropic see this as their capacity problem. They like "extra usage", and users who don't, well its their capacity problem.
- gambiting 6mo ago>>and Gemini in general LLM? You might be. Or at least I feel like Gemini is actually dumber than a house of bricks - I have multiple examples, just from last week, where following its advice would have lead to damage to equipment and could have hurt someone. That's just trying to work on an electronics project and askin Gemini for advice based on pictures and schematics - it just confidently states stuff that is 100000% bullshit, and I'm so glad that I have at least a basic understanding of how this stuff works or I would have easily hurt myself. It's somewhat decent at putting together meal plans for me every week, but it just doesn't follow instructions and keeps repeating itself. It hardly feels worth any money right now, like it's some kind of giant joke that all these companies are playing on us, spending billions of these talking boxes that don't seem that intelligent. I also use claude at work, and for C++ programming it behaves like someone who read a C++ book once and knows all the keywords, but has never actually written anything in C++ - the code it produces is barely usable, and only in very very small portions. Edit: I just remembered another one that made me incredibly angry. I've been reading the Neuromancer on and off, and I got back into it, but to remind myself of the plot I asked Gemini to summarise the plot only up to chapter 14, and I specifically included the instruction that it should double check it's not spoiling anything from the rest of the book. Lo and behold, it just printed out the summary of the ending and how the characters actions up to chapter 14 relate to it. And that was in the "Pro" setting too. Absolute travesty. If a real life person did that I'd stop being friends with them, but somehow I'm paying money for this. Maybe I'm the clown here.
- staticman2 6mo agoI'm curious: did you give Gemini the entire text of Neuromancer or did you expect it to use search results for chapters 1 to 14? I would have just fed it the text of chapters 1 to 14 from a non drm copy.
- gambiting 6mo agoI just asked like I said, give me plot summary until chapter 14, don't spoil the rest of the book. And of course when I told it what it just did it was like oh I'm sorry, here's a summary without the spoilers for the ending. So clearly it could do it without additional context.
- dartharva 6mo agoThe preference rankings keep fluctuating on every release for me. A year ago it was Gemini dominating coding tasks, then it was Claude, now it is the latest Codex again. With the next point release(s) the cycle will continue.