Kimi K3-256k
31–40 of 172 posts
Re: Kimi K3-256k
#32A bit of topic. But how likely is it that the US will restrict Chinese open weight models and also force Euro countries to do the same? I think it will be effective within 6 months. The US is having a hard time staying competitive.
Re: Kimi K3-256k
#33Since Claude is the first time for me really, really out (TIL against my wished about https://status.claude.com/ ), I am now interested enough to see what else works. But ... when I click pricing, I see "Join a waitlist". Wtf? Are they really that good, so were totally surprised and overwhelmed by the requests, is this a marketing stunt, or do they just don't have the hardware being in china?
> Are they really that good, They did exceed the expecations of pretty much everyone! I've also blogged about using the model, it's a bit on the slow side but pretty good! > so were totally surprised and overwhelmed by the requests ... or do they just don't have the hardware being in china? Yes, this is mostly the case: https://x.com/Kimi_Moonshot/status/2078855608565207130 As a user, I much prefer that to service di…
I am now rather pretty pissed towards antrophic for stopping my flow and forcing me to search for alternatives.
Re: Kimi K3-256k
#34Re: Kimi K3-256k
#35A bit of topic. But how likely is it that the US will restrict Chinese open weight models and also force Euro countries to do the same? I think it will be effective within 6 months. The US is having a hard time staying competitive.
Re: Kimi K3-256k
#36What is the purpose of this? Just a hard cutoff below the actual context window? You could set that in your harness anyway.
> k3 (1M) consumes about twice as much quota as k3-256k Cheaper?
Re: Kimi K3-256k
#37Why are Anthropic and OpenAI even allowing their coding harness apps to be plugged into different model providers…? I’m surprised they haven’t figured out a way to clamp down on that by now.
Re: Kimi K3-256k
#38Why are Anthropic and OpenAI even allowing their coding harness apps to be plugged into different model providers…? I’m surprised they haven’t figured out a way to clamp down on that by now.
Re: Kimi K3-256k
#39Having a lot of active context increases the per-token cost (flops issued and bytes read per token out) so it makes sense to pass that cost on to users. I'm actually surprised it's implemented as a hard cutoff instead of a smooth gradient.
Re: Kimi K3-256k
#40This was posted 38 minutes ago, and as of 20 minutes ago, several Anthropic services are now designated as having a "major outage". Doubt these are related, but it made me laugh a little.
Anthropic services have outages on all days ending in y.