6 ms·
This should erase any doubt that AI Labs are making $$$ on API inference. Kimi 2.5 (which this is based on) is served at $0.44 input / $2 output by a ton of di
by corlinp 5mo ago
This should erase any doubt that AI Labs are making $$$ on API inference.
Kimi 2.5 (which this is based on) is served at $0.44 input / $2 output by a ton of different providers on OpenRouter, 2.6 will certainly be similar.
That's about 11X less than Opus for similar smarts.
- Lalabadie 5mo agoFamously, OpenAI and Anthropic are devoted to increasing efficiency before scaling up resource usage.
- amazingamazing 5mo agoHow does it erase any doubt? You’re implying Chinese things can’t be actually cheaper to produce than American which is laughable
- corlinp 5mo agoMost of those inference providers are American, and China is actually at a disadvantage here because of export restrictions - US companies are using newer and more efficient chips.
- amazingamazing 5mo agoIf it’s newer and efficient then why is the api more expensive?
- veber-alex 5mo agoPrice is set based on what people are willing to pay not based on actual costs.
- amazingamazing 5mo agoI’d believe that if they didn’t lower limits
- gessha 5mo agoIt’s worth noting that the US is very behind on energy infra and that might affect the cost calculations since data centers are electricity guzzlers. Also, not sure if CN has completely switched off Nvidia or still using them for training.