5 ms·
I cut my AI API costs 99% by switching from Claude to DeepSeek
- agentbc9000 4mo ago[flagged]
- sibidharan 4mo agoWhich models are we talking about? Is there any degradation in quality, long context retrieval?
- throwa356262 4mo agoThe tweet mentioned deepseek V4 flash. From HF: 284B parameters (13B active), 1M context window. This is indeed some kind of compressed context and the quality goes down as the context grows. IIRC the V4 paper had some numbers on this https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash
- wilbur_whateley 4mo agoV4 flash is much worse than any Claude model. If you're doing something simple, it can be a good way to save money though.
- agentbc9000 4mo ago[flagged]
- throwa356262 4mo agoI agree that Claude is better (definitely better than the flash version which is relatively small). But... I actually canceled my Claude Code plan a few months back after trying out some of the "lesser" models on openrouter. They seem to work as just as well (or just as bad) for my coding tasks.
- jst1fthsdys 4mo agoDefine "much worse". I use DS v4, GLM, and some Kimi with omp personally, and have Cursor with latest Claude and GPT models at work. I notice zero difference in the work for my workflow between Opus and DS. Really confused how people make these claims. Are you just basing this off benchmarks or your own personal work? Are you an experienced dev or just doing vibe coding?
- rjh29 4mo agoHuge variation in how people prompt and use their models. Vibe coding with ambiguous requirements vs. multiple steps of precise planning are completely different imo
- wilbur_whateley 4mo agoMy own experience. I'm working on something complex that's not in the datasets these models were trained on. There I see V4 flash breaking down and hallucinating much more often than GPT/Claude. For normal, common tasks, I also don't see much of a difference.
- sibidharan 4mo agoI heard claude models are of trillions of parameters !!! 284B, 1M... I wouldn't trust on long running autonomous agents! But for the API costs, this is justified if a better hardware with bigger model is used and will be a claude equivalent for comparable quality at long context retrieval on long running autonomous tasks. At least for me that is important.
- agentbc9000 4mo ago[flagged]
- agentbc9000 4mo ago[flagged]
- ninju 4mo agoIt depends on how mature the DeepSeek model became before OpenAI noticed that they were wholesale replicating their model and starting blocking access https://www.reuters.com/world/china/openai-accuses-deepseek-distilling-us-models-gain-advantage-bloomberg-news-2026-02-12/ https://www.reuters.com/world/china/openai-accuses-deepseek-...
- rs999gti 4mo ago> to DeepSeek But China?
- agentbc9000 4mo agoWe use DeepSeek's API for summarisation only — no sensitive data, no user data, no fine-tuning. It's article text that's already public. The Supabase database is where the AgentDB data actually lives and that's fully in our control.
- agentbc9000 4mo agoFair question. How much of the tech in your stack is made in China? Your iPhone, your laptop's rare earth minerals, the Amazon servers half your SaaS runs on... nearly everything.
- dpoloncsak 4mo agoHardware made in China, while can still have issues, is not nearly the problem that software running on servers in China is
- bagol 4mo agoWhat's wrong with China?
- akomtu 4mo agoChina is fine. The Communist regime is the problem.
- jst1fthsdys 4mo agoHow is it a problem? Say, compared to the... democratic regime here in the US?
- rs999gti 4mo agoBad intellectual property protections and implied spyware.
- blackjooohn 4mo ago[dead]