3 ms·
Making Luna, which was already very cheap and extremely capable, 5x cheaper is crazy. I use Sol at work but Luna at home, and while there's definitely a differe
by pavpanchekha 2mo ago
Making Luna, which was already very cheap and extremely capable, 5x cheaper is crazy. I use Sol at work but Luna at home, and while there's definitely a difference, it doesn't feel like night-and-day. After a year of ever-increasing prices it suddenly feels (between this, Kimi K3, GLM 5.2) that prices are falling again.
- deleted 2mo ago[deleted]
- jedberg 2mo ago> Sol vs Luna > it doesn't feel like night-and-day. I see what you did there. :)
- deklesen 2mo agoGood observation! Kudos
- deleted 2mo ago[deleted]
- dominotw 2mo agodepends on what you are doing. if you are doing verifiable tasks like fixing bugs then any model would do as long as you write the right verification.
- maxdo 2mo agois kimi that cheap? it's a very expensive model
- pixelesque 2mo agoIt's cheaper currently on many of the inference providers. Personally, I'm having surprisingly good results with DeepSeek 4 Pro at home, which is very good value for money: it's not as good as Claude / GPT 5.6 (I have Co-pilot license at work), but it's still really useful for code reviews, validating thoughts, and especially designing / writing unit tests for new (and old before refactoring) functionality. And it's very cheap per task. (Flash is even cheaper, but I've had issues with that on more complex tasks where it starts forgetting things and arguing with itself "but wait, let me read the function again").
- subarctic 2mo agoI tried out deepseek v4 pro via a couple providers from openrouter, and it's always getting 429s. Are you running it on your own hardware?
- Mashimo 2mo agoWorks fine for me via opencode go.
- pixelesque 2mo agoI wish!! No, I'm using it via OpenRouter in pi.dev - I just used it 30 mins ago... Providers (automatically selected): StreamLake and Baidu Qianfan.
- mark_l_watson 2mo agoI toggle back and forth between deepseek v 4 flash/pro on FireWorks.ai using OpenCode. Easy to toggle, I default to flash.
- fy20 2mo agoDeepSeek V4 Pro is ridiculously priced, especially when you take into account caching. According to the DeepSeek usage panel, 50M tokens have cost me $1.38. It's not the smartest and does like to overthink, but if you have well defined problems it's good for coding. Well... except all your data going to China. I just use it for personal projects.
- forsalebypwner 2mo agoYup, last month I did ~150mil tokens on DeepSeek v4 Pro for just under $3
- pixelesque 2mo agoOut of interest, are you using the DeepSeek plan? (I've been using it via OpenRouter and it's much more than that, but still cheap).
- jug 2mo agoKimi K3 is fairly cheap per token but thinks like a madman with poor self esteem.
- pioneer37 2mo agoIts just a matter of time at this point.These companies are working day and night to capture the market.
- oh_no 2mo agoI pretty strongly disagree about comparing this to Kimi and GLM, 5.2 was a big price hike for Chinese models, and Kimi K3 was a big price hike to that. K3 was within spitting distance of OpenAI pricing (more expensive than short context Terra, less than long context). And that's after months of OpenAI/Anthropic prices going up. Now we have an American lab drastically cutting a price, feels like this is the opposite of that trend.