Live data from Hacker News

Inkling: Our Open-Weights Model

thinkingmachines.ai

151–160 of 324 posts

Re: Inkling: Our Open-Weights Model

#152
post #119

Very nice, multi modal, largest open weight model that supports audio. Would be interesting to see how good the audio capability is. If you want to run locally, checkout https://github.com/danielhanchen/llama.cpp/tree/add-inkling https://unsloth.ai/docs/models/inkling https://huggingface.co/unsloth/inkling-GGUF https://huggingface.co/unsloth/inkling-NVFP4 This supposedly is better than KimiK2.7, as much hype as GLM5.…

Not to mention - it is American. This is the first competitive non-Chinese open weights model since what, Llama 3?

North Mini Code by Cohere (HQd in Toronto) has honestly been very competitive in my personal assessment with many of the models coming out of the PRC. I'd position it below Moonshot AIs and Z.ais recent releases, but above the varieties of Qwen, Deepseek, MiMo, etc.

Depends whether America the continent or just the United States counts of course.

Re: Inkling: Our Open-Weights Model

#153

Earlier quoted context omitted.

To compete against America. If your country has something like DeepSeek you really can't afford to let it fall as it's your best leverage if the US government decides to ban companies in your country from accessing American LLMs. And this is why there will never be a "DeepSeek of the US."

Considering how volatile things can get depending on who's president, I'd say even American companies need to "compete against America" if they don't want to get their rug pulled from under them (which, apparently, the legal system allows to easily happen in the US).

Not to mention the entire world outside the US and China. China seems to have the edge in stability.

Re: Inkling: Our Open-Weights Model

#154

Earlier quoted context omitted.

I’m trying to be charitable but your comment reads as “China bad” propaganda to me. Who cares that DeepSeek and Z.ai are Chinese companies?

I think practically every government will want to put restrictions on private companies building models. Frankly the EU and the US will practically be less involved and have more pushback from the public in this than China. I think that’s less “China bad” than recognizing that China is a more authoritarian state and has far more proclivity to interfere than western states. Maybe I’m wrong? What does deep seek say abo…

https://chat.deepseek.com/share/uf8jih4q95lbq3s8t2

This is what Deepseek replied when I asked it with a burner account. Claims it doesn't have it in its training data... sure.

Re: Inkling: Our Open-Weights Model

#156
post #6

Earlier quoted context omitted.

Its not as good as GLM 5.2 for agentic workflows while also being bigger. Competition is going to be ruthless because the super low cost to switching. There is also AllenAi in the US, but they have yet to produce a model at this scale. Thankfully, new contenders can come out of nowhere and do well, as long as they can produce a competitive model.

> Its not as good as GLM 5.2 for agentic workflows while also being bigger GLM 5.2 underwent extensive post-training and iteration since its original release to reach its current state. This seems like an extremely strong model for a first release, with a lot of potential for improvement, just like DS4. Sometimes I wish Meta had stuck with Llama 4 a bit longer to see how much further it could be pushed.

Llama 4 wasn't deemed a success, and Meta pivoted away as its now former head of AI couldn't demonstrate, nor even showed interest in, business profit.

They overspent on llama 3 anyway so money ran dry, LeCun is good at running research, but budgets didn't stretch. Meta isn't investing in frontier big models anymore.

Re: Inkling: Our Open-Weights Model

#157
post #102

Very preliminary testing so far, but there is something here, far beyond what the benchmarks suggest. Only ever saw such outperformance of public evals vs my private ones with Anthropic models and while it is far to early to make any judgement at this stage, this model will take up a lot of mine time in the coming weeks by the look of things. Only ever viewed Moonshot AIs models as something I'd be able to live with…

Quick and still very early update, the model has (with web search disabled which was verified via the reasoning traces) accurately answered a number of questions focused on very niche details (engine specific maintenance in certain newtimers, very niche bag construction and material details) that I have only ever seen Gemini 3 and 3.1 Pro get correct. Neither Fable 5, nor GPT-5.6 Sol or any other model by any other lab has ever provided accurate information without web access for these specific questions for which an objectively correct answer absolutely exists and is general knowledge if one is versed in the specifics.

Being ahead of Fable 5 in any task, that is not included in public benchmarks and thus could be overfitted for, is impressive to say the least. Last time a model exceeded the expectations I had based on the release notes to such an extent was Haiku 4.5, which I still wish we got a solid replacement for.

Re: Inkling: Our Open-Weights Model

#158
post #62
post #52

Earlier quoted context omitted.

The story of Reflection AI is supposedly that the company was faffing and failing at winning in the coding agent space, but was introduced to Jenson, who suggested they build an open-weight model and said he would fund it. That turned into a $2 billion financing with NVIDIA doing roughly $500 million and was a complete pivot. I think the bet would have to be that a US Open Weight company either: 1. Gets a lot of mone…

Jensen Huang is just trying to commoditize the complements to his GPUs.

Cf Microsoft v Intel circa 1995

Re: Inkling: Our Open-Weights Model

#160
post #7

Earlier quoted context omitted.

It could be but there are a host of companies going after open weights models: Arcee, Reflection, Llama (TBD on Meta's focus on closed-source versus open-source), etc. That said, the fine-tuning API + open weight model at least is a semblance of a viable business that could work so I will be curious about it. I'm not sure the synergy is fully there (why is someone with an open weights model privelaged to fine-tune it…

I don’t really get the business plan for open weights model companies, is the idea companies would pay them for serving?

Thinky's main commercial product AFAIK is Tinker [0] - companies pay them to host their fine-tuning workloads and then the resulting fine-tuned models. I don't know if this is a good business plan, but I'm sure at least one person there has read Joel on Software [1].

[0] https://thinkingmachines.ai/tinker/

[1] https://www.joelonsoftware.com/2002/06/12/strategy-letter-v/

Post reply on HN