Live data from Hacker News

Granite 4.1: IBM's 8B Model Matching 32B MoE

firethering.com

1–10 of 223 posts

Re: Granite 4.1: IBM's 8B Model Matching 32B MoE

#4
I test drove it yesterday. It's pretty impressive at 8b. Runs on commodity hardware quickly.

Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent.

The 4b they released was not good for my needs but could probably handle tool calls or something

Re: Granite 4.1: IBM's 8B Model Matching 32B MoE

#5
Interesting to see a pivot away from MoE by both IBM and mistral while the larger classes of SOTA of models all seem to be sticking to it.

Quick vibe check of it- 8B @ Q6 - seems promising. Bit of a clinical tone, but can see that being useful for data processing and similar. You don't really want a LLM that spams you with emojis sometimes...

Re: Granite 4.1: IBM's 8B Model Matching 32B MoE

#6

I test drove it yesterday. It's pretty impressive at 8b. Runs on commodity hardware quickly. Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent. The 4b they released was not good for my needs…

Yea, No doubt Qwen 3.6 open weights are far more strong

Re: Granite 4.1: IBM's 8B Model Matching 32B MoE

#8

I test drove it yesterday. It's pretty impressive at 8b. Runs on commodity hardware quickly. Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent. The 4b they released was not good for my needs…

Yea, No doubt Qwen 3.6 open weights are far more strong

Why no doubt?
Post reply on HN