6 ms·
Currently burning money quickly on official deepseek api. They are also increasing pricing starting today. V4 Flash 0731 still feels like the most outstanding m
by Gecko4072 2mo ago
Currently burning money quickly on official deepseek api. They are also increasing pricing starting today. V4 Flash 0731 still feels like the most outstanding model of the past few months and probably to come.
- Jsttan 2mo agoWhat is the new price through?
- Gecko4072 2mo agohttps://api-docs.deepseek.com/quick_start/pricing/ https://api-docs.deepseek.com/quick_start/pricing/ edit: there are banner announcements saying v4 flash pricing will increase first then overall by an undetermined amount
- minraws 2mo agoisn't it the same old pricing? did they increase V4 Pro pricing already?
- nchmy 2mo agoi dont see any price increase there... what am i missing?
- alecsm 2mo agoRight below the pricing it is stated that they plan to increase the prices in the near future.
- vdfs 2mo agoIt's a big confusion, some[0] say an email was sent about significant price increase, personal I haven't seen anything official [0] https://finance.yahoo.com/technology/ai/articles/deepseek-plans-significant-price-increase-025333567.html https://finance.yahoo.com/technology/ai/articles/deepseek-pl...
- surgical_fire 2mo agoThe email is real, I received it from DeepSeek itself. I probably received it because I buy tokens directly from them. No actual price increase however.
- lionkor 2mo agoAlso got the email. It warned of a future large price increase, and to carefully watch usage. I read it as a "hey we will make stuff more expensive, don't miss it"
- GrinningFool 2mo agoThe banner on account settings; and a blurb on the pricing page: "We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice."
- deleted 2mo ago[deleted]
- anigbrowl 2mo agoIt has been updated now, and will take effect next week: https://api-docs.deepseek.com/quick_start/pricing/ https://api-docs.deepseek.com/quick_start/pricing/ Briefly Pro is $2/1m output in off-peak periods, $4 in peak. Flash is $0.66/$1.32. Input tokens are still much cheaper. I don't mind these prices but I find the need to check against two different time brackets of unequal length an annoying distraction. I guess I need to make some little background app or plugin.
- nolist_policy 2mo agoDeepSeek V4 Flash is the "too cheap to meter" of AI. And you can run the full unquantized model locally for $8000 (2x DGX Spark) at full 1M context and decent speeds: https://github.com/elsung/dgx-spark-deepseek-v4-flash#-long-context--what-to-expect https://github.com/elsung/dgx-spark-deepseek-v4-flash#-long-...
- Eueudhsbsj32 2mo agoWhat's the new pricing? The prices on OpenRouter still look the same.
- notatoad 2mo agonobody is saying. just "more". but openrouter says they don't expect the price to change other than through the deepseek api, other people hosting the same model will keep charging the same price.
- Eueudhsbsj32 2mo agoUnfortunately cache reads with third party providers are all 10-50x more expensive than with DeepSeek, so they're not even close to as cost efficient for multi-round agent use.
- monster_truck 2mo agoIt's just a flat 1.5x during peak hours, they emailed this to everyone 2 months ago. So still effectively limitless.
- anon373839 2mo agoYeah, Dax from OpenCode said that it appears to just be traffic shaping, nothing to do with the inference economics. He also said that OC have already replicated the inference cost in internal experiments.
- monster_truck 2mo agoYou can derive a pretty solid yardstick of how things are going for China by what you can find on the aftermarket, currently there is a glut of nvidia 4080s that have had their memory doubled up to 32GB. I'd have to assume they got a good deal buying up piles of H100s or whatever else was eating rack space, or potentially took a loss because they have hit the constraints of the # of cards they can rack. On the OEM side of things 9070/XTs are also shooting back up in price now that we have <$100 USB 4 egpu docks. People like to complain about how expensive things have gotten but I think it's pretty neat that there's so much pressure for throughput that it's even viable to buy 4 docks and 4 $850 GPUs and still save money over a single 48GB card.
- igravious 2mo agoyup :) i'm doing opencode <-> openrouter <-> official deepseek api (i don't get the opencode hate, i like it) how are you doing it? am also using Kimi K3 via kimi-code and also GLM 5.2 via ZCode happy with all three, they're trailing frontier but i figure if i'm running GNU/Linux then i ought to favour open weights models with my €s -- reduced my usage of claude/gpt to the ~$20 tier just to keep abreast of claude_code/codex developments
- literallyroy 2mo ago> i don't get the opencode hate, i like it When the company I work for was evaluating it, there were multiple rough points. Their terms and conditions allowed training on prompts, the default behavior was to route prompts to their servers for conversation summary/labeling. One of their lead maintainers is also super toxic on many issues. Sorry this is all baseless with no links, I’m on my phone and locating those issues again isn’t something I have time for. It’s a good tool I just don’t like the privacy policies nor maintainers attitudes.
- HDBaseT 2mo ago1. The privacy policy was a bit misleading, but it has since been updated to reflect the exact state of things. [1]. For example, DeepSeek models have ZDR, although their ZDR contract is renewed monthly. It COULD change. You need to toggle a Setting in your account to use DS. 2. At one point (apparently) summary and title generations were handled by Grok. This has changed, by default it uses your 'small_model' configured in your config. By default, it will use a cheap model provided by your provider. E.g. if you have ChatGPT API connected, it will use the cheapest ChatGPT model. OpenRouter users MAY see it routed to a free model however. [2] [3] [1] - https://opencode.ai/docs/go/#privacy https://opencode.ai/docs/go/#privacy [2] - https://github.com/anomalyco/opencode/blob/9b805e1cc4ba4a98419ca13d9d487c4550af8ddf/packages/opencode/src/provider/provider.ts#L1385 https://github.com/anomalyco/opencode/blob/9b805e1cc4ba4a984... [3] - https://opencode.ai/docs/config/ https://opencode.ai/docs/config/
- eli 2mo agoThe Deepseek official API is good with excellent caching. But their privacy policy is unusually bad - they can train off your prompts and completions.
- trollbridge 2mo agoUse another provider from OpenRouter. I really don’t care if they train off my prompts.
- stanac 2mo agoV4 Pro 0813 isn't offered by other providers. I can't find this model on hugging face. It's probably not open, or not open yet.
- sschueller 2mo agoDeepseek seems to have gotten too cheap. I have been using it for a long time and it's at a point now where my credits balance barely moves even at max setting.
- killingtime74 2mo agoJust use opencode go, you get more bang for your buck. Same api