Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
411–420 of 682 posts
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#412Earlier quoted context omitted.
You can partially tell by the tokeniser; which gives you some hint into the training corpus mix. is four Gemma4 tokens, but one Qwen3.6 token.
Where do you find this information for each model?
In that file you can see all tokens and their corresponding numeric codes.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#413Just asking, what is the recommended models for M3 MacBook with 18G memory? Seems modern local models are not available.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#414I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…
It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#415Earlier quoted context omitted.
I'm a 52 year old natural born US citizen whose ancestors have been here for generations and I'm currently anti-American. Why wouldn't I be? We've never been the shining beacon of light we would claim to be, but we're so fucking awful now.
When you say you are anti-American, what do you mean by that? Are you, for example, wishing for the demise of the United States? Do you want to tear down the 1st Amendment and the Statue of Liberty? Are you against democracy? Do you want our businesses and factories to shut down and go out of business? Are you willing or would you support foreign countries attacking our military at home and abroad? Are you cheering a…
I am very pro the "dream" of America, in terms of liberty, democracy, etc. According to every "democracy index" I'm aware of we're not doing so hot in that regard, generally rating as a flawed/deficient democracy and the trends are going in the wrong direction, fast.
> Do you want our businesses and factories to shut down and go out of business?
Generally, no, but this is way too open-ended of a question. I want good economic opportunity for everyone, including every US citizen. But relevant to the OP if Meta got snapped out of existence I think it would be a net positive for the world.
> Are you willing or would you support foreign countries attacking our military at home and abroad?
Nope. But across our entire history ask yourself how many foreign countries have attacked the US? Now ask yourself how many the US has attacked. With those numbers in mind, does the US seem like "the good guys"? really? ...really?
> Are you cheering against our athletes?
Nope, but I'm not cheering for them either just because they are American, I'm not a tribalist.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#416I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…
It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#417Earlier quoted context omitted.
While composing a reply to a comment throwing tons of shade on American AI, I took some time to check out the commenter’s HN profile. Their comment history was about 50% such comments. Their submission history started with an article about how Russia was unfairly blamed for some hacking campaign. It’s entirely possible that this is not a foreign influence campaign. Perhaps there’s a group here that is simply anti-Ame…
> On the other hand, one should not discount the value of HN as tastemaker and trendsetter. I would encourage dedicated readers here to aggressively and persistently discount the value of HN as a tastemaker and trendsetter. HN is actually a trailing indicator on tastes and trends, essentially by design. Things only make it to the front page if they get submitted and voted upward by a large number of people. That mean…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#418Earlier quoted context omitted.
I'd also argue this is the case for any company releasing open weights. They're not righteous, they're marketing. That's not necessarily a bad thing! They're releasing some great stuff for free and we benefit from that. Every company doing this has a motivation to not release these for free. Alibaba, Google, Moonshot, Thinking Machines, etc are not releasing their models for free because they love to. They want to gr…
This model doesn’t look solid at all. It comes months after the Qwen model, and in almost half the benchmarks, it performs worse than that. Plus, the next Qwen 3.8 is going to be announced this week. So, this model is DOA.
It also has a knowledge cutoff inside this year.
The main limitation is the smaller maximum recommended context.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#419The gguf is up and works, I don’t know if it’s them or unsloth that’s facilitated this but it’s nice because e.g. Inkling still doesn’t appear to have support in llama.cpp which makes it irrelevant to a class of user. Unfortunately I don’t have enough experience with Qwen 27B to immediately compare, but I do it’s Qwen 3.6 35B A3. It’s much slower obviously but it seems to be way more efficient with its thinking to th…
I don't really use the Qwen 3.6 27B though I do test the variants (Bonsai, ThinkingCap). I really like the 3.6 35B A3B for experiments, and it seems OK, but as you say, it spins round in thinking loops more than say the 26B Gemma 4 does. If Muse doesn't actually-wait itself as much it will be very interesting. I am just downloading it to run my small tests.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#420Earlier quoted context omitted.
I don't understand the desire to run own AI models for programming locally. No laptop is ever going to be as powerful and energy efficient to run anything close to OpenAI, Anthropic or Google models. A model you can run on a loptop is simply not going to work as well as it's needed for programming. Small models for linguistic work fine, but anything more sophisticated simply won't provide enough resources or power. O…
> A model you can run on a loptop is simply not going to work as well as it's needed for programming The models you can run on a high-spec laptop today are approximately where frontier models were 12-18mo ago (albeit at a lower tok/s rate). If you scan back through hn comments from that era, you’ll find plenty of people saying “this is powerful enough to massively increase my productivity”.