Granite 4.1: IBM's 8B Model Matching 32B MoE
1–10 of 223 posts
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#2Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#3edit: I just realised they do actually have a 30b release alongside this. Haven't tried it yet.
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#4Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent.
The 4b they released was not good for my needs but could probably handle tool calls or something
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#5Quick vibe check of it- 8B @ Q6 - seems promising. Bit of a clinical tone, but can see that being useful for data processing and similar. You don't really want a LLM that spams you with emojis sometimes...
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#6I test drove it yesterday. It's pretty impressive at 8b. Runs on commodity hardware quickly. Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent. The 4b they released was not good for my needs…
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#7Wish they also released an embedding model, in the line of their previous: compact (while good)...
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#8I test drove it yesterday. It's pretty impressive at 8b. Runs on commodity hardware quickly. Qwen3.6 35b a3b is still my local champion but I may use this for auto complete and small tasks. Granite has recent training data which is nice. If the other small models got fine tuned on recent data I don't know if I would use this at all, but that alone makes it pretty decent. The 4b they released was not good for my needs…
Yea, No doubt Qwen 3.6 open weights are far more strong
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#9Why people don't edit out obvious sloppification and expect to still have readers left
Re: Granite 4.1: IBM's 8B Model Matching 32B MoE
#10> Full stop. Why people don't edit out obvious sloppification and expect to still have readers left