5 ms·
I don’t understand good experiences people are having with Gemini. It’s the only model that sometimes loses/forgets context in literally next message. Plus feed
by galkk 17d ago
I don’t understand good experiences people are having with Gemini. It’s the only model that sometimes loses/forgets context in literally next message. Plus feeding unasked product links to responses.
- travthedev 17d agoDefinitely true in some cases, but I find Gemini to be less biased about certain topics - which is quite nice, especially when I'm just trying to hear the facts.
- solenoid0937 17d agoI wonder how much life DeepMind has left in it, especially after Hassabis's departure. Google execs must be having discussions about simply throwing their weight behind Anthropic since they already own so much of the company.
- mianos 17d agoI agree. Lot's of people really like it. For me it often just forgets all context and starts showing random slop. It's super clear as, when I ask it what happened to some element earlier in the conversation it tells me it does not have that. It might be good if it told me, but randomly lose the plot is quite frustrating. Claude does it occasionally but it's a more a soft landing earlier context seems to be compacted, not completely lose the plot. I just cancelled my pro subscription. I really wanted it to be good but not yet.
- Gareth321 17d agoI strongly agree. I suspect it's people who have not yet used the paid models from OpenAI and Anthropic. Gemini is comparable to free models from other providers, but not in the same universe as paid models. This is frustrating because when I discuss AI with laypeople they think it's still incapable of counting the number of Rs in "strawberry." They believe it to be essentially useless and incapable of basic tasks. Which, to be fair, is the case with the free models.
- KeplerBoy 17d ago[dead]
- otabdeveloper4 17d ago> you're just not using the latest model, bro Pro tip, ChatGPT is the normiest of all normie websites right now. You're not part of the cognoscenti just because you learned how to type prompts into one of the most popular websites in the world. P.S. You're probably not using OpenAI models for complex or non-standard tasks. It shits the best just as often as Qwen when you need precision and detail in a non-obvious problem.
- belowavgiq 17d agoWhere did he claim otherwise? Take your meds.
- Gareth321 17d ago> You're probably not using OpenAI models for complex or non-standard tasks. It shits the best just as often as Qwen when you need precision and detail in a non-obvious problem. The benchmarks clearly show otherwise. This is your cue to tell me the benchmarks are made by the Illuminati and only your superior and subjective methods of evaluation are correct.
- vintermann 17d agoThe tasks you actually need to do trump benchmarks, yes. I haven't tried out the ridiculously expensive models besides the latest Gemini, and it gave from equal to slightly worse results than latest DeepSeek, at a far higher price. It does also seems Gemini's main problem wasn't that it was stupid, but that it was good at doing slightly different things than what I asked it to, very well. Which might well also have to do with me being better at wrangling DeepSeek's quirks than Gemini. Still, at that price tag, it's not worth it.
- Gareth321 17d ago
- whateveracct 17d agoi have programmed a dozen scripts with google search AI lol they'll live in my rc for decades
- TacticalCoder 17d agoAh yup for quick one-off one-liners or small Bash script, Gemini works perfectly fine too. For longer scripts I use another model.
- TomGarden 17d agoI love it as a variation from the others. 3.8 flash is the best back-and-forth model for iterating imo, but would not use for long horizon
- galkk 16d agoAs I mentioned - I find it awful for iterating. My recent example: I was researching shoes for toddler, wide, boa or similar mechanism instead of laces/velcro. I did same prompt in ChatGPT and flash. After several back and forth responses/clarifications flash completely lost track of what I’m looking for and started suggesting nonsense.
- TomGarden 16d agoI see! I've only really used it for coding. You should design velcrobench!
- w4yai 17d agoHave you tried Gemini 3.8 Flash recently ? I was like you before, Gemini was the worst model to me. Then 3.8 came out. At first I was sceptical, but this model *is* able to do useful things ! Complex things. Of course, it is NOT perfect. But for things like small/medium complex tasks subagents, it's perfect. Now, is it worth the money vs Astra ? I don't think so. But still my point remain relevant.
- glub 17d agoThey do a lot of weird things with the context in their user facing products like Gemini. Google seems to always have trouble with their harnesses
- atraac 17d agoI also had that issue happen to me, surprisingly only when I used Polish, not English. But other than that I actually like Gemini, I check stuff against it all the time, especially when walking my dog. It also works in AndroidAuto for me, but I use very basic stuff like changing Spotify music. I was on iOS before and there's no comparison to old Siri that was just garbage. I use Gemini practically every day, it's fine for the most part IMO. They really do need to polish integrations though, app connectors barely every work outside of Google's own apps. Using Oppo's app Mind Place through Gemini f.e. is just bad experience and almost never works. For example, I had pleasant experience of Gemini finding me places for a walk/hike on vacation in Tirol, where I specifically required not too much of an ascent and an asphalt road for a stroller.
- Cameri 17d agoI’ve found this to be true as well. It’s got really poor attention and will derail into world-building fast. I’m convinced Google is just shipping it to capture market share but they know Gemini isn’t ready for serious use.
- bayindirh 17d agoFor information retrieval and collation related tasks (reading lists, deep dives on subjects, etc.), Gemini is way better w.r.t. other models in my experience. This difference is probably due to Google's web knowledge and free pass to YouTube. However, I'm happy what I got from it so far. When I ask the question once in a blue moon, it can generally one-shot the answer, even.
- epistasis 17d agoIt is by far the least accurate of any model I have used too. For a company that was started to organize the world's information, it has by far the most misinformation I've encountered. I couldn't even get it to tell me how to pay for antigravity, it sent me on some fruitless paths and eventually said "you shouldn't pay for this, it's too hard to figure it out."
- irregularbowels 16d ago[dead]
- inquirerGeneral 16d ago[dead]