Earlier quoted context omitted.
DSV4 Flash 0731 already runs on RTX 4090 24GB + 128GB system RAM at a usable tok/s and quantization.
You personally? Just curious. Context window is also a factor and ram isn’t really cheap. Sparks are assembled units which I like.
Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
361–370 of 682 posts
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#362Earlier quoted context omitted.
While composing a reply to a comment throwing tons of shade on American AI, I took some time to check out the commenter’s HN profile. Their comment history was about 50% such comments. Their submission history started with an article about how Russia was unfairly blamed for some hacking campaign. It’s entirely possible that this is not a foreign influence campaign. Perhaps there’s a group here that is simply anti-Ame…
Would be nice if there were a hn feature, userscript, or plugin to just filter comments from new accounts. Bonus if there was some sentiment analysis or llm-based analysis to filter out unsubstantiated inflammatory comments too.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#363Earlier quoted context omitted.
> I don't think there's a strong consensus on Hacker News. Even something like the time of day an article is posted might get different engagement depending on who is active in which time zones. Based on my own experience and reading, I do think there's a general consensus on this site but I could certainly be wrong about that. I'm less concerned about hypocrisy per se, it's more that the arguments that are used, eve…
> It's common, if not inevitable, for people who feel strongly about $topic to conclude that the system (or the community, or the mods, etc.) are biased against their side. One is far more likely to notice whatever data points that one dislikes because they go against one's view and overweight those relative to others. This is probably the single most reliable phenomenon on this site. Keep in mind that the people wit…
It's not even about sides, if for the last few hundred days you read a few AI related threads a day, then you notice that almost all arguments are rehashed, literally it's the same thing repeated using different words for 80%+ of comments on almost every AI thread. I started skipping most of it because there is genuinely nothing new or interesting added to these discussions.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#364Earlier quoted context omitted.
Of the two competing models Meta compare Glimmer to in the post, one is Google's Gemma 4. At this size open weight model, a Western company was already state of the art, Meta is joining that competition. And my memory is that Gemma 4 got little criticism or doom/gloom. And no, it isn't Chinese.
Likewise the Inkling open weights announcement, Thinking Machines model, was also not criticised. The comment about Meta is because of particular dislike of Meta, because of their business model, and how harmful they've ultimately turned out to be for the world - disproportionately so relative to their benefits to the world, compared to other big tech companies.
This is certainly what many people around here appear to believe, but there are lots and lots of people who get much more value out of Meta's products than those of any other tech company. Whatsapp alone is probably the most useful tech product for many, many people.
That being said, FB/Meta have done a bunch of awful stuff, but to say that they're worse than Google/Amazon/Microsoft is not necessarily obvious.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#365Meta is rocking AI. As of last week I have been using their excellent muse coding harness with their model Muse Spark 1.2. Starting this morning I am running their new local 30B model muse-glimmer on my old MacMini 32G using Ollama (remember to increase the context size!) and pi coding harness. I am getting good results with muse-glimmer running locally, with the caveat that everything runs slowly (e.g., give it a ta…
seems to underperform on Terminal Bench compared with qwen3.6-27b: 51.7 vs 60.7
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#366Pelican, rendered by Muse Glimmer on my Mac running LM Studio (with this model release: https://lmstudio.ai/models/muse-glimmer ): https://tools.simonwillison.net/markdown-svg-renderer#url=ht... It has all of the components of a pelican riding a bicycle, though not exactly arranged in the right order! (For comparison, here are the pelicans I got from Muse Spark 1, 1.1, and 1.2: https://bsky.app/profile/simonwillison.…
Maybe a sign that they didn't have SVG pelicans in the dataset
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#367Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#368Earlier quoted context omitted.
While composing a reply to a comment throwing tons of shade on American AI, I took some time to check out the commenter’s HN profile. Their comment history was about 50% such comments. Their submission history started with an article about how Russia was unfairly blamed for some hacking campaign. It’s entirely possible that this is not a foreign influence campaign. Perhaps there’s a group here that is simply anti-Ame…
Would be nice if there were a hn feature, userscript, or plugin to just filter comments from new accounts. Bonus if there was some sentiment analysis or llm-based analysis to filter out unsubstantiated inflammatory comments too.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#369Earlier quoted context omitted.
It’s really interesting timing, Qwen over thinking is what kills it for me. I’m just glad we have more options in this size class now.
I've been using Qwen3.6 35B A3B, and with reasoning turned on, I'd say 2/3 (give or take) of the tokens for a response are thinking tokens. Which at 70+ tps locally, that isn't that awful. I run an 80k context across 4-10 "agents" for my solo TTRPG, where Qwen is the GM, each NPC at a location, the director, and the narrator. Each turn is about 45-60 seconds to generate all of the various responses. The GM and direct…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#370Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion? EDIT: An open weight version of Muse Spark 1.2 is going to be released as well: https://x.com/alexandr_wang/status/2086756152034066792 https://xcancel.com/alexandr_wang/status/2086756152034066792
It’s really interesting timing, Qwen over thinking is what kills it for me. I’m just glad we have more options in this size class now.