Live data from Hacker News

Inkling: Our Open-Weights Model

thinkingmachines.ai

131–140 of 324 posts

Re: Inkling: Our Open-Weights Model

#132
post #99
post #5

Seems like this is particularly good at instruction following, but not as strong at coding as others. It's always great to get more diversity of open weight models though! I'll need to test this out to see what its "personality" is like.

seems pretty dang snappy and I like it's tone/personality so far. > look at today's hackernews frontpage and generate me a daily briefing report (create an artifact) to read later for today's nerd news https://chat.home.jake.town/artifacts/019f679d-99e5-7000-b02...

This is the best voice/tone I've seen from any model so far. It's using filler words and phrases in places that normal people would put them, rather than sounding like a corporate customer support agent!

Re: Inkling: Our Open-Weights Model

#133

It's nice to see a strong long context open weights model that is multi-modal. There are many applications that will benefit from the strength in audio here and until z.ai and co work in visual this could be very strong for general agentic applications, though I see there's a bit of weakness in the benches for areas that might make that less true. Like all models need to slap it in your harness and do proper evals on…

MiniMax M3 and DeepSeek v4-Pro are highly capable long context open weight multi-modal models. But long-context is a trap, because performance still falls dramatically after 150k-200k context.

[dead]

Re: Inkling: Our Open-Weights Model

#134
post #128

Earlier quoted context omitted.

I’m trying to be charitable but your comment reads as “China bad” propaganda to me. Who cares that DeepSeek and Z.ai are Chinese companies?

It doesn't matter until it does. If the chinese government decides that open weight model releases are no longer allowed, that's a lot of companies that can't release new models. Same with the US government, etc. Having diversity is important.

It's a similar problem the human DNA solved by telling our teenage selves that our parents are dumb and we needed to move to a new tribe. Genetic diversity, but a digital equivalent.

Re: Inkling: Our Open-Weights Model

#135

Earlier quoted context omitted.

What is the moat? The time it takes for AI to rewrite an efficient inference stack for a new model? Considering most LLMs follow a similar architecture, adapting to a new model shouldn't take that much time.

I don't know why people keep saying there's no moat. There's no moat. Having a FUCK ton of money to train these gigantic fucking models and retain the brains to make it happen is a moat. You're not going to train one using a VPS from LowEndBox.

But if people can download it for free, that is a drawbridge across the moat competitors can use to get the same model.

Re: Inkling: Our Open-Weights Model

#136
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

I’m trying to be charitable but your comment reads as “China bad” propaganda to me. Who cares that DeepSeek and Z.ai are Chinese companies?

I think practically every government will want to put restrictions on private companies building models.

Frankly the EU and the US will practically be less involved and have more pushback from the public in this than China. I think that’s less “China bad” than recognizing that China is a more authoritarian state and has far more proclivity to interfere than western states.

Maybe I’m wrong? What does deep seek say about Tiananmen square in 1989?

Re: Inkling: Our Open-Weights Model

#137

Earlier quoted context omitted.

This is genuine, noob question: how is this different from AWS? I get that they're in very different businesses, but for both don't they have the issue that once a client gets big enough the client might decide to move the services in-house? Based on how much of the internet went down when that AWS data center crashed the answer is clearly "No" for AWS. Is that because of physical, real-world infrastructure? Are ther…

Data is heavy. I would say "it's risky and requires a lot of labor to migrate without corruption, loss of data" and also minimizing downtime. Sure anyone can run pg_backup, but can you do it across 90 databases? Can you do it live? Can you coordinate rollout of the process, cutover, and monitor for failure? What's the cost of egress for this? Is the team your A-team or the B-team? Can you trust this to the B-team? Is…

I don't think it's that difficult. Their servers are stateless too. S3 is easy to migrate.

Database is more difficult, but tons of people have done it successfully.... meanwhile people who host their own LLMs are relatively small in number in comparison.

Most companies don't do their own data centers mainly because it is more expensive and less reliable. It's something they can just pay for the problem to go away. The calculus for hosting your own LLM is probably similar.

Even Stripe who built their own coding agents and has tons of money/resources still decides not to host their own LLMs.

Still, many people will prefer open-weight models. It is similar to how we prefer linux but still use AWS/Render/and whatever. It doesn't lock us in, and we can move providers if we want to.

Re: Inkling: Our Open-Weights Model

#138

Very nice, multi modal, largest open weight model that supports audio. Would be interesting to see how good the audio capability is. If you want to run locally, checkout https://github.com/danielhanchen/llama.cpp/tree/add-inkling https://unsloth.ai/docs/models/inkling https://huggingface.co/unsloth/inkling-GGUF https://huggingface.co/unsloth/inkling-NVFP4 This supposedly is better than KimiK2.7, as much hype as GLM5.…

I also have been preferring kimik2.7 more than GLM5.2. I'm interested to give this a try.

Re: Inkling: Our Open-Weights Model

#139

Very nice, multi modal, largest open weight model that supports audio. Would be interesting to see how good the audio capability is. If you want to run locally, checkout https://github.com/danielhanchen/llama.cpp/tree/add-inkling https://unsloth.ai/docs/models/inkling https://huggingface.co/unsloth/inkling-GGUF https://huggingface.co/unsloth/inkling-NVFP4 This supposedly is better than KimiK2.7, as much hype as GLM5.…

Where are you getting supposedly. It does worse in most benchmarks

Re: Inkling: Our Open-Weights Model

#140
post #7
post #2

America needs its own DeepSeek or Z.ai, a lot of people (myself included) root for open chinese models to win because they have no other choice. Thinking Machines might be it.

It could be but there are a host of companies going after open weights models: Arcee, Reflection, Llama (TBD on Meta's focus on closed-source versus open-source), etc. That said, the fine-tuning API + open weight model at least is a semblance of a viable business that could work so I will be curious about it. I'm not sure the synergy is fully there (why is someone with an open weights model privelaged to fine-tune it…

I don’t really get the business plan for open weights model companies, is the idea companies would pay them for serving?
Post reply on HN