3 ms·
I predict that no one will use this and everyone will use Kimi K3.
by nullbio 3mo ago
I predict that no one will use this and everyone will use Kimi K3.
- embedding-shape 3mo agoI've been playing around with K3 a bunch, but the verbosity of the reasoning makes complete e2e agent work basically cost the same as other smaller models, and I'm not seeing a huge difference in quality, just a way longer e2e completion time.
- sunaookami 3mo agoSame problem with every chinese model currently, they overthink way too much and take too much tokens and time.
- EgregiousCube 3mo agoA consequence of aggressive distillation?
- embedding-shape 3mo agoMore or less, yeah. I've found mild success with deepseek-v4-flash though, and also Qwen3.5-122B-A10B-NVFP4 running locally, especially in terms of "doesn't overthink every single prompt" and somewhat reasonable quality. Really wishing for a 3.8 update of the 122B variant, that'd be really competitive (for local usage) :)
- szundi 3mo ago[dead]
- jadbox 3mo agoWhat's the price difference?
- rubslopes 3mo agoWhy? Price? If the reason is performance, I've been using non-frontier models for cheap, and they run great for my needs (GLM 5.2, DeepSeek v4 Pro).
- rurban 3mo agoWe'll probably use it, but for images. Qwen is still the best for images