Earlier quoted context omitted.
But they aren't keeping up They are lauded for the ability to cost ratio, or their ability to parameter ratio, but virtually everyone using LLMs for productive work are using ChatGPT/Gemini/Claude. They are kind of like Huffy bicycles. Good value, work well, but if you go to any serious event, no one will be riding one.
they are keeping up. i have been using just chinese models for the last 2 years. chatgpt/gemini/claude have marketing. there's nothing that you can do with those models that can't be done with deepseek, glm or kimi. if there is, do let us know.
This aligns with the benchmarks as well; they benchmark great for what they are, but still bottom of the barrel when competing for "state of the art."
And yes, it's great you daily Chinese models, but the vast majority of people try them, say "impressive", then go back to the most performant models.