Kimi K3-256k
kimi.com
Kimi K3-256k
1–10 of 172 posts
Re: Kimi K3-256k
#2[flagged]
Re: Kimi K3-256k
#3[flagged]
Re: Kimi K3-256k
#4[flagged]
For me the sweet spot is somewhere under 500k depending on how extensive I want to get. You can build up a sizable effort project in half a million tokens with Claude, with Claude having all the context from ground 0 to wherever you're off at.
Re: Kimi K3-256k
#5This is just an API level change right? The model itself should be the same I think.
Re: Kimi K3-256k
#6[flagged]
For me the sweet spot is somewhere under 500k depending on how extensive I want to get. You can build up a sizable effort project in half a million tokens with Claude, with Claude having all the context from ground 0 to wherever you're off at.
I'm always curious what you guys are working on; every git repo I've run a local model on and stick below <100k to increase speed seems effective enough to scope patches and changes.
Re: Kimi K3-256k
#7> k3-256k is now available. Within 256k context, it delivers the same results. k3 (1M) consumes about twice as much quota as k3-256k.
Re: Kimi K3-256k
#8This isn't quantized, right? Just a smaller context?
Re: Kimi K3-256k
#9omg! new model!!
Re: Kimi K3-256k
#10This isn't quantized, right? Just a smaller context?
Its 256k context window. Quantization is orthogonal. We cant really tell directly so it could be quantized.