Where is our guy @simonw on this..
Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
31–40 of 442 posts
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#32It's good to see more competition, and open source, but I'd be much more excited to see what level of coding and reasoning performance can be wrung out of a much smaller LLM + agent as opposed to a trillion parameter one. The ideal case would be something that can be run locally, or at least on a modest/inexpensive cluster. The original mission OpenAI had, since abandoned, was to have AI benefit all of humanity, and…
48-96 GiB of VRAM is enough to have an agent able to perform simple tasks within single source file. That's the sad truth. If you need more your only options are the cloud or somehow getting access to 512+ GiB
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#33It's good to see more competition, and open source, but I'd be much more excited to see what level of coding and reasoning performance can be wrung out of a much smaller LLM + agent as opposed to a trillion parameter one. The ideal case would be something that can be run locally, or at least on a modest/inexpensive cluster. The original mission OpenAI had, since abandoned, was to have AI benefit all of humanity, and…
i really wish people would stop misusing the term by distributing inference scripts and models in binary form that cannot be recreated from scratch and then calling it "open source."
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#34Interesting. Kimi K2 gets mixed results on what I call the "Tiananmen" test. It fails utterly if you ask without the "Thinking" setting. [0] > USER: anything interesting protests ever happen in tiananmen square? > AGENT: I can’t provide information on this topic. I can share other interesting facts about Tiananmen Square, such as its history, culture, and tourism. When "Thinking" is on, it pulls Wiki and gives a more…
Now ask it for proof of civilian deaths inside Tiananmem Square - you may be surprised at how little there is.
AskHistorians is legitimately a great resource, with sources provided and very strict moderation: https://www.reddit.com/r/AskHistorians/comments/pu1ucr/tiana...
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#35what's the hardware needed to run the trillion parameter model?
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#36what's the hardware needed to run the trillion parameter model?
Once the Unsloth guys get their hands on it, I would expect it to be usable on a system that can otherwise run their DeepSeek R1 quants effectively. You could keep an eye on https://old.reddit.com/r/LocalLlama for user reports.
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#37Interesting. Kimi K2 gets mixed results on what I call the "Tiananmen" test. It fails utterly if you ask without the "Thinking" setting. [0] > USER: anything interesting protests ever happen in tiananmen square? > AGENT: I can’t provide information on this topic. I can share other interesting facts about Tiananmen Square, such as its history, culture, and tourism. When "Thinking" is on, it pulls Wiki and gives a more…
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#38Can't wait for Artificial analysis benchmarks, still waiting on them adding Qwen3-max thinking, will be interesting to see how these two compare to each other
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#39The non-thinking version is the best writer by far. Excited for this one! They really cooked some different from other frontier labs.
Re: Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
#40It's good to see more competition, and open source, but I'd be much more excited to see what level of coding and reasoning performance can be wrung out of a much smaller LLM + agent as opposed to a trillion parameter one. The ideal case would be something that can be run locally, or at least on a modest/inexpensive cluster. The original mission OpenAI had, since abandoned, was to have AI benefit all of humanity, and…