3 ms·
Grok it's really expensive. I'm getting really amazing results using DeepSeek 4.1 Flash for fraction of the price.
by meerita 5d ago
Grok it's really expensive. I'm getting really amazing results using DeepSeek 4.1 Flash for fraction of the price.
- parineum 5d agoBrought to you by...
- meerita 5d agoBy no one. For the price of 1M token you can get more and with better results with other models.
- includenotfound 5d agoSure, if you're doing easy work. But Grok is a lot more intelligent and can handle harder tasks.
- testfrequency 5d agoWhat is the most secure way to use this model as someone who is lazy
- user43928 5d agoI understand DeepSeek 4.1 Flash is available on US providers with Zero Data Retention if that is what you are asking.
- sparkling 5d agoYes, but with subpar caching and higher cached token pricing, compared to directly using the DeepSeek platform.
- brcmthrowaway 5d agoLink for the lazy?
- microsoftedging 5d agoOne example is Opencode. https://opencode.ai/v2/docs/console/models/ https://opencode.ai/v2/docs/console/models/ "Privacy# All these models are hosted in the US. Providers follow a zero-retention policy and do not use your data for model training, with the following exceptions: Big Pickle: During its free period, collected data may be used to improve the model. DeepSeek V4 Flash Free: During its free period, collected data may be used to improve the model. MiMo-V2.5 Free: During its free period, collected data may be used to improve the model. Laguna S 2.1 Free: During its free period, collected data may be used to improve the model. Ling-3.0-tiny Free: During its free period, collected data may be used to improve the model. LongCat-2.0 Free: During its free period, collected data may be used to improve the model. North Mini Code Free: During its free period, collected data may be retained and used to improve the model. Do not submit personal or confidential data. See the provider’s Terms of Use and Privacy Policy. Nemotron 3 Ultra Free (NVIDIA free endpoints): Trial use only — do not submit personal or confidential data. Your use is logged for security purposes and to improve NVIDIA products and services. The logged session data for improvement purposes is not linked to your identity or any persistent identifier. For more information about data processing practices, see the Privacy Policy. By interacting with this endpoint, you consent to the collection, recording, and use of such information and the NVIDIA API Trial Terms of Service."
- thehamkercat 5d agoopenrouter, "together" provider is fastest (165 t/s at the time of writing) and has ZDR and all https://openrouter.ai/deepseek/deepseek-v4.1-flash?endpoint=57e1cb3e-9762-4ee1-a67c-259bd3af24b7#providers https://openrouter.ai/deepseek/deepseek-v4.1-flash?endpoint=...
- nicce 5d agoSadly, there is no way to tell if this is running with real weights or being heavily quantized.
- simlevesque 5d agoI like devcontainers
- drewnick 5d agoI use it on fireworks which is US/ZDR and pretty reliable. We run a few hundred million tokens/day through it for dollars. Many are cached, which is super duper cheap.
- _s_a_m_ 5d agoDeepSeek 4.1 Flash is garbage, it almost only produced trash code. if you do extremely dumb things it is maybe sometimes fine to use.
- yipinwong 5d agoNot only that all DeepSeek is all garbage. GLM or Kimi are better for my own personal projects. DS? uhm. it just keeps doing dumb crap
- vorticalbox 5d agoCompared to the deep seek, gml sure but compared to OpenAI and Anthropic it’s actually very cheap. In cursor I have switch over to grok for planning a composer for coding.
- brianwawok 5d agoMaybe mid priced is a better term for it lol.
- vorticalbox 5d agoSure I can accept that lol
- thefourthchime 5d agoIt’s a great value if you get Cursor Ultra. I basically have infinite tokens