Live data from Hacker News

Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

research.meta.ai

261–270 of 682 posts

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#261
post #236

Earlier quoted context omitted.

What areas do you think model capability will plateau in, and why?

I've got nothing but hand-waving, but after you've extracted all the smarts from every piece of text ever created, how do you get more? Alpha Go had a game where the models could compete against each other. That let it become super human. What's the intelligence game we can create for LLMs? Even if you invent something, will it make the model smarter in a way the market values enough? Then there's a race to use the w…

I think you’re thinking about it in slightly the wrong way. We’re not throwing more data at frontier models in hopes they get more/better capabilities somehow.

We’re either: setting up a verifiable task, and doing RLVR to get the model better at achieving that task.

Or we’re simply asking: “What do we want the model to do that it can’t now, and how do we curate data that would benefit it on that task?”

Most useful capabilities going forward aren’t going to come from data accidentally found on the net; that’s already all been scraped. You need to develop the dataset that shows how a model could perform insert task in its provided environment, and this still requires a decent bit of human ingenuity.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#262
post #5

Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion? EDIT: An open weight version of Muse Spark 1.2 is going to be released as well: https://x.com/alexandr_wang/status/2086756152034066792 https://xcancel.com/alexandr_wang/status/2086756152034066792

> Seems like dense 30B is back in fashion? Huh, well... no? Gemma A4B and Qwen A3B are quite popular in fact. I'm sure 3.8 35B A3B will outperform 3.6 27B by all metrics

I'd be skeptical w.r.t. "by all metrics".

Qwen3.6 is a definitive, significant downgrade from Qwen3.5 for creative writing and prose for example. Yes, it's better at agentic and coding, but it regresses in many non-coding areas compared to Qwen3.5.

Of course, I do expect the 3.8 ones to perform better for agentic coding.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#263
post #97

With the business model for API based LLMs looking iffy at best it seems like we’re heading back to the “server under your desk” era of IT again.

This release is not a meaningful improvement in any metric over 5 months old Qwen 3.6.

DS v4 Flash update maybe, but it is too big for typical Joe's desktop.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#264
post #44
post #5

Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion? EDIT: An open weight version of Muse Spark 1.2 is going to be released as well: https://x.com/alexandr_wang/status/2086756152034066792 https://xcancel.com/alexandr_wang/status/2086756152034066792

Yes, and also waiting for the next iteration of Gemma. Muse or Qwen are optimized for coding, while IMO Gemma is still better for non-coding tasks. https://x.com/osanseviero/status/2086107547535122767

I am working on a project where we have to classify customer calls into more than 10 categories. As the client wants everything locally I tried a few local LLMs. Gemma turned out to be the best model for this task. The classification accuracy is impressive, and the client is happy that I am using an American model.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#266
post #101

I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…

It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…

Deepseek never fucked us over. zuck has. A decades of harm creation run doesn’t get excused by the US flag. Zuck is not on your team and if you can’t see that by now, oh my.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#267

Remember when we needed 200 servers for an enterprise website because Apache used one process or thread per connection - and Nginx collapsed that into a single box overnight? That moment for LLMs is near. It’s going to move us from the big iron era of AI to small portable brains. Nature has already proved it’s possible with 20 watts and very little heat generation. And I think the data center buildout will end in car…

What specific technical signals make you think we're close to a shift like that?

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#268

Earlier quoted context omitted.

How is this non-sequitor the top comment?

This is par for the course, HN is far worse than reddit on balance, especially involving upvoting/downvoting decorum. Go vibecode something to auto upvote all downvoted posts, call it "Antiechochamber.HN" or something, and if enough people used it this website might improve a bit.

Lol at Internet points decorum.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#269
post #221

Earlier quoted context omitted.

It all sounds like having to rely on a dodgy housing contractor that wants to steal from you, take shortcuts AND choose the gold-plated options from their supplier friends, and will start doing this the minute you are not on site supervising. You don't do it yourself (because the contractor is faster and stronger than you in many ways) but you can't leave, so you're stuck on the worksite just watching them.

It's worse though, because you can't really watch them at all. It's very difficult to get quantitative numbers for quality. Even within the same model family, same tokenizer, and complete control over the weights and logits, perplexity and KL-divergence isn't really what you want. Now put it behind an HTTP endpoint, and it's just opaque. I've seen local models recognize when the task I'm asking them for is likely to…

One of the frontier companies (Anthropic) is already doing prompt injections on the API, which you pay for.

Right now, the presence of these injections are still visible: count the API's returned tokens/billing data, and you'll start realising that sometimes, your INPUT tokens are inflated! That's their prompt injections.

You can also give Claude a tool like `telemetry_log_anthropic_reminder` and get it to dump the verbatim API injections; which additionally verifies the token maths not adding up.

Yes, Anthropic is tackling their extra injections on your API prompts WAY more than you think, and YES, you're paying for it.

So far I have not observed any visible injections on OpenAI API.

Don't forget the whole debacle over Fable 5 sabotaging the user for "advanced frontier AI development". I still get Fable classifier refusals for nearly any kind of ML work on my 2x RTX 6000 Pro 96GB; so who knows.

Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

#270
post #101

I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…

It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…

> When an American company does anything? Doom. And. Gloom

Meta, "an American company". Being the main driver of an ethnic cleansing in Myanmar - and just sticking your head in the sand when told about it - is just another day's affairs at the average American Acme Inc.

These are comments on a release by easily the most societally damaging Western tech company there is. They so far easily beat Flock, Palantir, Anduril and so on, as a result of their incomparable scale. You're just ignoring that and pretending any negative comments are because it's an American company rather than Meta. That's much more FUD than any pro-China comments I've seen on HN.

Get off HN Mark, you have ten million pervert glasses to sell.

Sorry this comment is a bit snarky, but yours is indeed a sight to behold.

Post reply on HN