3 ms·
wow - $0.95 input/$4 output. If its anywhere near opus 4.6 that's incredible.
by pt9567 6mo ago
wow - $0.95 input/$4 output. If its anywhere near opus 4.6 that's incredible.
- corlinp 6mo agoThis should erase any doubt that AI Labs are making $$$ on API inference. Kimi 2.5 (which this is based on) is served at $0.44 input / $2 output by a ton of different providers on OpenRouter, 2.6 will certainly be similar. That's about 11X less than Opus for similar smarts.
- Lalabadie 6mo agoFamously, OpenAI and Anthropic are devoted to increasing efficiency before scaling up resource usage.
- amazingamazing 6mo agoHow does it erase any doubt? You’re implying Chinese things can’t be actually cheaper to produce than American which is laughable
- corlinp 6mo agoMost of those inference providers are American, and China is actually at a disadvantage here because of export restrictions - US companies are using newer and more efficient chips.
- amazingamazing 6mo agoIf it’s newer and efficient then why is the api more expensive?
- veber-alex 6mo agoPrice is set based on what people are willing to pay not based on actual costs.
- amazingamazing 6mo agoI’d believe that if they didn’t lower limits
- gessha 6mo agoIt’s worth noting that the US is very behind on energy infra and that might affect the cost calculations since data centers are electricity guzzlers. Also, not sure if CN has completely switched off Nvidia or still using them for training.