Live data from Hacker News

Inkling: Our Open-Weights Model

thinkingmachines.ai

71–80 of 324 posts

Re: Inkling: Our Open-Weights Model

#74
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

What is the business model for an open weight model?

Thinky has a potential answer in Tinker — give away the weights and charge for the SFT (and maybe RL down the line) to make the model more capable for specific tasks.

Re: Inkling: Our Open-Weights Model

#76
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

What is the business model for an open weight model?

The same business model that Deepseek is using.

Open-source models + services. This is more attractive because it doesn't lock in the vendors. If I grow larger, I can decide to deploy the open-source models.

Re: Inkling: Our Open-Weights Model

#77
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

What is the business model for an open weight model?

To compete against America. If your country has something like DeepSeek you really can't afford to let it fall as it's your best leverage if the US government decides to ban companies in your country from accessing American LLMs. And this is why there will never be a "DeepSeek of the US."

Re: Inkling: Our Open-Weights Model

#78
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

Hopefully they'll release some smaller models (<100B) that we can run on home hardware at faster than 10tok/s.

Re: Inkling: Our Open-Weights Model

#79

Earlier quoted context omitted.

> ...while KIMI and DeepSeek will release Fable-class models this week. What new model is DeepSeek releasing? Their current V4 Pro at Max reasoning is consistently worse than GLM 5.2 at Max reasoning, though the latter is close to Opus 4.8 at Extra/Max reasoning, albeit a little bit worse in my experience (though if they gave comparable amounts of tokens to Anthropic 5x Max subscription I could see myself moving over…

DeepSeekV4 was a preview model, read the papers. It's not the final model. They released it to demonstrate architectural capabilities. They are still training and the model release is planned within the next month.

If they somehow make it approach Fable in capability, I’ll be quite surprised!

Re: Inkling: Our Open-Weights Model

#80
post #6

Earlier quoted context omitted.

Its not as good as GLM 5.2 for agentic workflows while also being bigger. Competition is going to be ruthless because the super low cost to switching. There is also AllenAi in the US, but they have yet to produce a model at this scale. Thankfully, new contenders can come out of nowhere and do well, as long as they can produce a competitive model.

> Its not as good as GLM 5.2 for agentic workflows while also being bigger GLM 5.2 underwent extensive post-training and iteration since its original release to reach its current state. This seems like an extremely strong model for a first release, with a lot of potential for improvement, just like DS4. Sometimes I wish Meta had stuck with Llama 4 a bit longer to see how much further it could be pushed.

This is a great point
Post reply on HN