4 ms·
Gemini 2.5 Flash is an impressive model for its price. However, I don't understand why Gemini 2.0 Flash is still popular. From OpenRouter last week: * xAI: Gr
by Liwink 1y ago
Gemini 2.5 Flash is an impressive model for its price. However, I don't understand why Gemini 2.0 Flash is still popular.
From OpenRouter last week:
* xAI: Grok Code Fast 1: 1.15T
* Anthropic: Claude Sonnet 4: 586B
* Google: Gemini 2.5 Flash: 325B
* Sonoma Sky Alpha: 227B
* Google: Gemini 2.0 Flash: 187B
* DeepSeek: DeepSeek V3.1 (free): 180B
* xAI: Grok 4 Fast (free): 158B
* OpenAI: GPT-4.1 Mini: 157B
* DeepSeek: DeepSeek V3 0324: 142B
- crazysim 1y agoMaybe the same reason why they kept the name for the 2.5 Flash update. People are lazy at pointing to the latest name.
- koakuma-chan 1y agoWhy is Grok so popular
- coder543 1y agoI think it has been free in some editor plugins, which is probably a significant factor. I would rather use a model that is good than a model that is free, but different people have different priorities.
- YetAnotherNick 1y agoNon free has double usage than free. Free one uses your data for training.
- Imustaskforhelp 1y agoI mean, I can kinda roll through a lot of iterations with this model without worrying about any AI limits. Y'know with all these latest models, the lines are kinda blurry actually. The definition of "good" is being foggy. So it might as well be free as the definition of money is clear as crystal. I also used it for some time to test on something really really niche like building telegram bot in cloudflare workers and grok-4-fast was kinda decent on that for the most part actually. So that's nice.
- davey48016 1y agoI think it's very cheap right now.
- keeeba 1y agoIt came from nowhere to 1T tokens per week, seems… suspect.
- riku_iki 1y agoI think it is included for free into some coding product
- BoredPositron 1y agoThey had a lot of free promos with coding apps. It's okay and cheap so I bet some sticked with it.
- NitpickLawyer 1y agoIt's pretty good and fast af. At backend stuff is ~ gpt5-mini in capabilities, writes ok code, and works good with agentic extensions like roo/kilo. My colleagues said it handles frontend creation so-so, but it's so fast that you can "roll" a couple of tries and choose the one you want. Also cheap enough to not really matter.
- SR2Z 1y agoYeah, the speed and price are why I use it. I find that any LLM is garbage at writing code unless it gets constant high-entropy feedback (e.g. an MCP tool reporting lint errors, a test, etc.) and the quality of the final code depends a lot more on how well the LLM was guided than the quality of the model. A bad model with good automated tooling and prompts will beat a good model without them, and if your goal is to build good tooling and prompts you need a tighter iteration loop.
- nwienert 1y agoThis is so far off my experience. Grok 4 fast is straight trash, it literally isn’t even close to decent code for what I tried. Meanwhile Sonnet is miles better - but even still, Opus while I guess technically being only slightly better, in practice is so much better that I find it hard to use Sonnet at all.
- minimaxir 1y agoGrok Code Fast 1 usage is driven almost entirely by Kilo Code and Cline: https://openrouter.ai/x-ai/grok-code-fast-1/apps https://openrouter.ai/x-ai/grok-code-fast-1/apps Both apps have offered usage for free for a limited time: https://blog.kilocode.ai/p/grok-code-fast-get-this-frontier-ai-model-free https://blog.kilocode.ai/p/grok-code-fast-get-this-frontier-... https://cline.bot/blog/grok-code-fast https://cline.bot/blog/grok-code-fast
- ewoodrich 1y agoYep Kilo (and Cline/Roo more recently) push these free trial of the week models really hard, partially as incentive to register an account with their cloud offering. I began using Cline and Roo before "cloud" features were even a thing and still haven't bothered to register, but I do play with the free Kilo models when I see them since I'm already signed in (they got me with some kind of register and spend $5 to get $X model credits deal) and hey, it's free (I really don't care about my random personal projects being used for training). If xAI in particular is in the mood to light cash on fire promoting their new model, you'll see it everywhere during the promo period, so not surprised that heavily boosts xAI stats. The mystery codename models of the week are a bit easier to miss.
- Simon321 1y agoit was free
- frde_me 1y agoI know we have a lot of workloads at my company on older models no one has bothered to upgrade yet
- koakuma-chan 1y agoHell yeah, GPT 35 Turbo
- kilroy123 1y agoThere are cheaper models. Could cut the bill in half or more.
- koakuma-chan 1y agodavinci-001 xd
- tiahura 1y agoPrimarily classification or something else?
- YetAnotherNick 1y agoGemini 2.0 Flash is the best fast non reasoning model by quite a margin. Lot of things doesn't require any reasoning.
- mistic92 1y agoPrice, 2.0 Flash is cheaper than 2.5 Flash but still very good model.
- nextos 1y agoAPI usage of Flash 2.0 is free, at least till you hit a very generous bound. It's not simply a trial period. You don't even need to register any payment details to get an API key. This might be a reason for its popularity. AFAIK only some Mistral offerings have a similar free tier?
- FergusArgyll 1y agoYeah, that's my use case. When you want to test some program / script that utilizes an llm in the middle and you just want to make sure everything non-llm related is working. It's free! just try again and again till it "compiles" and then switch to 2.5
- indigodaddy 1y agowow this would be great for a webapp/site that just needs a basic/performant LLM for some basic tasks.
- nextos 1y agoYou might hit some throttling limits. During certain periods of the day, at least in my location, some requests are not served. It might not be OK for that kind of usecase, or might breach ToS. But it's still great. Even my premium Perplexity account doesn't give me free API access.
- PetrBrzyBrzek 1y agoIt’s cheaper and faster. What’s not to understand?
- testycool 1y agoYou can get it to be unhinged as well. It's awesome.
- simonw 1y agoMy one big problem with OpenRouter is that, as far as I can tell, they don't provide any indication of how many companies are using each model. For all I know there are a couple of enormous whales on there who, should they decide to switch from one model to another, will instantly impact those overall ratings. I'd love to have a bit more transparency about volume so I can tell if that's what is happening or not.
- minimaxir 1y agoGranted, due to OpenRouter's 5.5% surcharge, any enormous whales have a strong financial incentive to use the provider's API directly. A "weekly active API Keys" faceted by models/app would be a useful data point to measure real-world popularity though.
- eli 1y agoThey kinda have that already, no? https://openrouter.ai/apps?url=https%3A%2F%2Faider.chat%2F https://openrouter.ai/apps?url=https%3A%2F%2Faider.chat%2F
- minimaxir 1y agoAggregating by tokens causes the problem simonw mentions in that one poweruser can skew the chart too much.
- simonw 1y agoRight, that chart shows App usage based on the user-agent header but doesn't tell you if there is a single individual user of an app that skews the results.
- __mharrison__ 1y agoI was skewing the Gemini starts with my Aider usage. Basically the only model in using with openrouter, until I recently started running qwen3-next locally. 2.5 is probably the best balance for tools like Aider.
- rohansood15 1y ago2.0 Flash is significantly cheaper than 2.5 Flash, and is/was better than 2.5-Flash-Lite before this latest update. It's a great workhorse model for basic text parsing/summary/image understanding etc. Though looks like 2.5-Flash-Lite will make it redundant.