Live data from Hacker News

Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

research.meta.ai

411–420 of 682 posts

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#412
post #150

Earlier quoted context omitted.

You can partially tell by the tokeniser; which gives you some hint into the training corpus mix. is four Gemma4 tokens, but one Qwen3.6 token.

Where do you find this information for each model?

When you look on HuggingFace.co at the files of a model, for each model you will see a file "tokenizer.json".

In that file you can see all tokens and their corresponding numeric codes.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#413
post #331

Just asking, what is the recommended models for M3 MacBook with 18G memory? Seems modern local models are not available.

You can try this site, toggle your computer specs at the top for a refined list of models and tokens/sec.

https://www.canirun.ai

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#414
post #101

I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…

It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…

It makes me wonder why this dynamic exists here, and I do wonder at times how much our conversations here are influenced by China in a top-down fashion. I'd prefer to think that HN is pretty organic, but that is probably a naive thought.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#415

Earlier quoted context omitted.

I'm a 52 year old natural born US citizen whose ancestors have been here for generations and I'm currently anti-American. Why wouldn't I be? We've never been the shining beacon of light we would claim to be, but we're so fucking awful now.

When you say you are anti-American, what do you mean by that? Are you, for example, wishing for the demise of the United States? Do you want to tear down the 1st Amendment and the Statue of Liberty? Are you against democracy? Do you want our businesses and factories to shut down and go out of business? Are you willing or would you support foreign countries attacking our military at home and abroad? Are you cheering a…

> Are you, for example, wishing for the demise of the United States? Do you want to tear down the 1st Amendment and the Statue of Liberty? Are you against democracy? Do you want our businesses and factories to shut down and go out of business? Are you willing or would you support foreign countries attacking our military at home and abroad? Are you cheering against our athletes?

I am very pro the "dream" of America, in terms of liberty, democracy, etc. According to every "democracy index" I'm aware of we're not doing so hot in that regard, generally rating as a flawed/deficient democracy and the trends are going in the wrong direction, fast.

> Do you want our businesses and factories to shut down and go out of business?

Generally, no, but this is way too open-ended of a question. I want good economic opportunity for everyone, including every US citizen. But relevant to the OP if Meta got snapped out of existence I think it would be a net positive for the world.

> Are you willing or would you support foreign countries attacking our military at home and abroad?

Nope. But across our entire history ask yourself how many foreign countries have attacked the US? Now ask yourself how many the US has attacked. With those numbers in mind, does the US seem like "the good guys"? really? ...really?

> Are you cheering against our athletes?

Nope, but I'm not cheering for them either just because they are American, I'm not a tribalist.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#416
post #101

I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…

It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…

both countries are authoritarian. both are using this "free" tech to spy on people and control them.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#417

Earlier quoted context omitted.

While composing a reply to a comment throwing tons of shade on American AI, I took some time to check out the commenter’s HN profile. Their comment history was about 50% such comments. Their submission history started with an article about how Russia was unfairly blamed for some hacking campaign. It’s entirely possible that this is not a foreign influence campaign. Perhaps there’s a group here that is simply anti-Ame…

> On the other hand, one should not discount the value of HN as tastemaker and trendsetter. I would encourage dedicated readers here to aggressively and persistently discount the value of HN as a tastemaker and trendsetter. HN is actually a trailing indicator on tastes and trends, essentially by design. Things only make it to the front page if they get submitted and voted upward by a large number of people. That mean…

So - presumably you have another site in mind that is better. I'd be intrigued to know which one you would recommend.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#418
post #157

Earlier quoted context omitted.

I'd also argue this is the case for any company releasing open weights. They're not righteous, they're marketing. That's not necessarily a bad thing! They're releasing some great stuff for free and we benefit from that. Every company doing this has a motivation to not release these for free. Alibaba, Google, Moonshot, Thinking Machines, etc are not releasing their models for free because they love to. They want to gr…

This model doesn’t look solid at all. It comes months after the Qwen model, and in almost half the benchmarks, it performs worse than that. Plus, the next Qwen 3.8 is going to be announced this week. So, this model is DOA.

I don't really care that much about benchmarks, but having tested it on one of my puzzle prompts I can tell you that it solves it well, writes clearly, isn't noticeably slower than Qwen 3.6 27B and is much more terse in its reasoning (which will help with preserve-reasoning).

It also has a knowledge cutoff inside this year.

The main limitation is the smaller maximum recommended context.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#419
post #345
post #327

The gguf is up and works, I don’t know if it’s them or unsloth that’s facilitated this but it’s nice because e.g. Inkling still doesn’t appear to have support in llama.cpp which makes it irrelevant to a class of user. Unfortunately I don’t have enough experience with Qwen 27B to immediately compare, but I do it’s Qwen 3.6 35B A3. It’s much slower obviously but it seems to be way more efficient with its thinking to th…

I don't really use the Qwen 3.6 27B though I do test the variants (Bonsai, ThinkingCap). I really like the 3.6 35B A3B for experiments, and it seems OK, but as you say, it spins round in thinking loops more than say the 26B Gemma 4 does. If Muse doesn't actually-wait itself as much it will be very interesting. I am just downloading it to run my small tests.

the “Actually… But wait!” style responses are so annoying, even Claude opus struggles with this so I’d be interested if meta has done something to cut down on that while still giving good responses

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#420

Earlier quoted context omitted.

I don't understand the desire to run own AI models for programming locally. No laptop is ever going to be as powerful and energy efficient to run anything close to OpenAI, Anthropic or Google models. A model you can run on a loptop is simply not going to work as well as it's needed for programming. Small models for linguistic work fine, but anything more sophisticated simply won't provide enough resources or power. O…

> A model you can run on a loptop is simply not going to work as well as it's needed for programming The models you can run on a high-spec laptop today are approximately where frontier models were 12-18mo ago (albeit at a lower tok/s rate). If you scan back through hn comments from that era, you’ll find plenty of people saying “this is powerful enough to massively increase my productivity”.

Slower is meaningfully dumber when you’re time bounded and need all the inference time compute you can get.
Post reply on HN