Earlier quoted context omitted.
Not to mention - it is American. This is the first competitive non-Chinese open weights model since what, Llama 3?
Gemma 4
Inkling: Our Open-Weights Model
151–160 of 324 posts
Re: Inkling: Our Open-Weights Model
#152Very nice, multi modal, largest open weight model that supports audio. Would be interesting to see how good the audio capability is. If you want to run locally, checkout https://github.com/danielhanchen/llama.cpp/tree/add-inkling https://unsloth.ai/docs/models/inkling https://huggingface.co/unsloth/inkling-GGUF https://huggingface.co/unsloth/inkling-NVFP4 This supposedly is better than KimiK2.7, as much hype as GLM5.…
Not to mention - it is American. This is the first competitive non-Chinese open weights model since what, Llama 3?
Depends whether America the continent or just the United States counts of course.
Re: Inkling: Our Open-Weights Model
#153Earlier quoted context omitted.
To compete against America. If your country has something like DeepSeek you really can't afford to let it fall as it's your best leverage if the US government decides to ban companies in your country from accessing American LLMs. And this is why there will never be a "DeepSeek of the US."
Considering how volatile things can get depending on who's president, I'd say even American companies need to "compete against America" if they don't want to get their rug pulled from under them (which, apparently, the legal system allows to easily happen in the US).
Re: Inkling: Our Open-Weights Model
#154Earlier quoted context omitted.
I’m trying to be charitable but your comment reads as “China bad” propaganda to me. Who cares that DeepSeek and Z.ai are Chinese companies?
I think practically every government will want to put restrictions on private companies building models. Frankly the EU and the US will practically be less involved and have more pushback from the public in this than China. I think that’s less “China bad” than recognizing that China is a more authoritarian state and has far more proclivity to interfere than western states. Maybe I’m wrong? What does deep seek say abo…
This is what Deepseek replied when I asked it with a burner account. Claims it doesn't have it in its training data... sure.
Re: Inkling: Our Open-Weights Model
#155Something on that level but multi-modal would be quite nice!
Re: Inkling: Our Open-Weights Model
#156Earlier quoted context omitted.
Its not as good as GLM 5.2 for agentic workflows while also being bigger. Competition is going to be ruthless because the super low cost to switching. There is also AllenAi in the US, but they have yet to produce a model at this scale. Thankfully, new contenders can come out of nowhere and do well, as long as they can produce a competitive model.
> Its not as good as GLM 5.2 for agentic workflows while also being bigger GLM 5.2 underwent extensive post-training and iteration since its original release to reach its current state. This seems like an extremely strong model for a first release, with a lot of potential for improvement, just like DS4. Sometimes I wish Meta had stuck with Llama 4 a bit longer to see how much further it could be pushed.
They overspent on llama 3 anyway so money ran dry, LeCun is good at running research, but budgets didn't stretch. Meta isn't investing in frontier big models anymore.
Re: Inkling: Our Open-Weights Model
#157Very preliminary testing so far, but there is something here, far beyond what the benchmarks suggest. Only ever saw such outperformance of public evals vs my private ones with Anthropic models and while it is far to early to make any judgement at this stage, this model will take up a lot of mine time in the coming weeks by the look of things. Only ever viewed Moonshot AIs models as something I'd be able to live with…
Being ahead of Fable 5 in any task, that is not included in public benchmarks and thus could be overfitted for, is impressive to say the least. Last time a model exceeded the expectations I had based on the release notes to such an extent was Haiku 4.5, which I still wish we got a solid replacement for.
Re: Inkling: Our Open-Weights Model
#158Earlier quoted context omitted.
The story of Reflection AI is supposedly that the company was faffing and failing at winning in the coding agent space, but was introduced to Jenson, who suggested they build an open-weight model and said he would fund it. That turned into a $2 billion financing with NVIDIA doing roughly $500 million and was a complete pivot. I think the bet would have to be that a US Open Weight company either: 1. Gets a lot of mone…
Jensen Huang is just trying to commoditize the complements to his GPUs.
Re: Inkling: Our Open-Weights Model
#159Re: Inkling: Our Open-Weights Model
#160Earlier quoted context omitted.
It could be but there are a host of companies going after open weights models: Arcee, Reflection, Llama (TBD on Meta's focus on closed-source versus open-source), etc. That said, the fine-tuning API + open weight model at least is a semblance of a viable business that could work so I will be curious about it. I'm not sure the synergy is fully there (why is someone with an open weights model privelaged to fine-tune it…
I don’t really get the business plan for open weights model companies, is the idea companies would pay them for serving?
[0] https://thinkingmachines.ai/tinker/
[1] https://www.joelonsoftware.com/2002/06/12/strategy-letter-v/