I think any new model not demonstrably maybe 20-30% over Deepseek v4 capabilities priced over the price per token of Deepseek is almost automatically deprecated as low use model (maybe for Planning).
Is Deepseek just eating cost or are people able to host their open models for comparable costs?
Kimi K2.7-Code: open-source coding model with better token efficiency
21–30 of 254 posts
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#22Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#23How is 2.7 a thing _now_ ? it's not even mentioned on moonshot's webpage..
https://platform.kimi.ai/docs/guide/kimi-k2-7-code-quickstar...
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#24I would really love to know if anyone has any experience with something like opencode + Kimi K2.6/2.7 now compared to Claude Code. What is better, what is worse, what is the cost comparison. I am currently paying $100 for the 5x Max plan, but Fable is running through the usage limits quite drastically and I cannot really say it's night and day compared to Opus. Also, I use this mostly for my side projects, so the $10…
I can only talk about GLM 5.1 which is roughly at sonnet 4 levels imo. It's good, does most tasks well that I throw at it, but will fail at anything congitive/complex. It gets stuck often. It costs ~6$ a month though
I use the oh-my-openagent planning system and haven’t used vanilla OpenCode enough to know how much that is contributing.
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#25I am still very new to the open-weight/source models. If anyone is using them full-time, I’d really love to hear about the setup and how they perform, as I am considering moving my org off Anthropic products.
These models have open weights, but at the moment most flagship models are practically accessible only through third-party model providers. The main exception is models in the ~30B parameter range, which can still be run on consumer-grade GPUs. That said, even consumer GPUs have become increasingly expensive and difficult to justify in recent years.
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#26I would really love to know if anyone has any experience with something like opencode + Kimi K2.6/2.7 now compared to Claude Code. What is better, what is worse, what is the cost comparison. I am currently paying $100 for the 5x Max plan, but Fable is running through the usage limits quite drastically and I cannot really say it's night and day compared to Opus. Also, I use this mostly for my side projects, so the $10…
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#27I was wondering how does Anthropic and likes keep competitive when Opus is ($5 / $25) 5x times more expensive compared to Kimi K2.6 ($0.7 / $3.4) or other Chinese models, while being only marginally better. My theory is that US enterprise just can't send data to Chinese and that's understandable, but is that "the moat"?
I say this as a relatively frequent user of Kimi models and generally a big fan. But on not-yet-gamed benchmarks like DeepSWE, Kimi K2.6 is beaten soundly by Claude Sonnet 4.6 ($3 / $15) and even slightly by GPT 5.4 Mini ($0.75 / $4.50).
There's no question Kimi models are very good for a lot of code tasks. They're the best quality open weight model. But to get similar overall outcomes as on Sonnet/Opus, on average you'll spend many more tokens and will have to do more managing of the model. You shouldn't look at price per token, you should look at how much you pay for the entire process.
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#28I am still very new to the open-weight/source models. If anyone is using them full-time, I’d really love to hear about the setup and how they perform, as I am considering moving my org off Anthropic products.
I use glm5.1 plus pi with a few customized skills and am very happy with it. I hadn’t touched my Claude 5x plan for a couple of weeks but opened it back up in Claude code when fable was released and did a few tasks and still was happy to return to glm/pi.
When I tried glm found it way way slower (omlx as runtime)
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#29- GPT-5.5: 62.7%
- Opus 4.8: 62.2%
- Kimi K2.7 Code: 56.3%
- Kimi K2.6: 48.2%
Re: Kimi K2.7-Code: open-source coding model with better token efficiency
#30I would really love to know if anyone has any experience with something like opencode + Kimi K2.6/2.7 now compared to Claude Code. What is better, what is worse, what is the cost comparison. I am currently paying $100 for the 5x Max plan, but Fable is running through the usage limits quite drastically and I cannot really say it's night and day compared to Opus. Also, I use this mostly for my side projects, so the $10…
The Kimi problem is it doesn’t follow instructions and goes off track often. Other than that it’s pretty decent (for the price).