With the business model for API based LLMs looking iffy at best it seems like we’re heading back to the “server under your desk” era of IT again.
Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
431–440 of 682 posts
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#432Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion? EDIT: An open weight version of Muse Spark 1.2 is going to be released as well: https://x.com/alexandr_wang/status/2086756152034066792 https://xcancel.com/alexandr_wang/status/2086756152034066792
Based on the benchmarks, it seems that Muse Glimmer barely edges out against Qwen3.6 27B, except for tool-calling skills (MCP, etc.). I wouldn't be surprised if they released it now because they are afraid they wouldn't beat Qwen3.8 27B.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#433Earlier quoted context omitted.
> On the other hand, one should not discount the value of HN as tastemaker and trendsetter. I would encourage dedicated readers here to aggressively and persistently discount the value of HN as a tastemaker and trendsetter. HN is actually a trailing indicator on tastes and trends, essentially by design. Things only make it to the front page if they get submitted and voted upward by a large number of people. That mean…
So - presumably you have another site in mind that is better. I'd be intrigued to know which one you would recommend.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#434Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#435Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion? EDIT: An open weight version of Muse Spark 1.2 is going to be released as well: https://x.com/alexandr_wang/status/2086756152034066792 https://xcancel.com/alexandr_wang/status/2086756152034066792
Makes me feel hopeful. Things felt more positive around the llama 3 era. Now it’s like a dark, dreadful race.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#436Remember when we needed 200 servers for an enterprise website because Apache used one process or thread per connection - and Nginx collapsed that into a single box overnight? That moment for LLMs is near. It’s going to move us from the big iron era of AI to small portable brains. Nature has already proved it’s possible with 20 watts and very little heat generation. And I think the data center buildout will end in car…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#437Earlier quoted context omitted.
I'm a 52 year old natural born US citizen whose ancestors have been here for generations and I'm currently anti-American. Why wouldn't I be? We've never been the shining beacon of light we would claim to be, but we're so fucking awful now.
When you say you are anti-American, what do you mean by that? Are you, for example, wishing for the demise of the United States? Do you want to tear down the 1st Amendment and the Statue of Liberty? Are you against democracy? Do you want our businesses and factories to shut down and go out of business? Are you willing or would you support foreign countries attacking our military at home and abroad? Are you cheering a…
I believe that many Americans that were previously dismissive or ambivalent regarding critiques of US activity at home and abroad (either due to patriotism, realpolitik apologia, or general naïveté) are now re-evaluating some of the beliefs they hold about their country in light of the chronic political dysfunction and an absolutely breathtaking extent of corruption being perpetrated in broad daylight today (as well as indications that extensive corruption has long festered among our elite class, surfaced via the Epstein revelations).
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#438Earlier quoted context omitted.
It’s really interesting timing, Qwen over thinking is what kills it for me. I’m just glad we have more options in this size class now.
I've been using Qwen3.6 35B A3B, and with reasoning turned on, I'd say 2/3 (give or take) of the tokens for a response are thinking tokens. Which at 70+ tps locally, that isn't that awful. I run an 80k context across 4-10 "agents" for my solo TTRPG, where Qwen is the GM, each NPC at a location, the director, and the narrator. Each turn is about 45-60 seconds to generate all of the various responses. The GM and direct…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#439Earlier quoted context omitted.
While composing a reply to a comment throwing tons of shade on American AI, I took some time to check out the commenter’s HN profile. Their comment history was about 50% such comments. Their submission history started with an article about how Russia was unfairly blamed for some hacking campaign. It’s entirely possible that this is not a foreign influence campaign. Perhaps there’s a group here that is simply anti-Ame…
> On the other hand, one should not discount the value of HN as tastemaker and trendsetter. I would encourage dedicated readers here to aggressively and persistently discount the value of HN as a tastemaker and trendsetter. HN is actually a trailing indicator on tastes and trends, essentially by design. Things only make it to the front page if they get submitted and voted upward by a large number of people. That mean…
> AI X/Twitter — a few hundred accounts effectively set the narrative in the first 24 hours; vibe checks here outrun benchmarks.
> r/LocalLLaMA — the open-weights kingmaker; a model that fails here doesn't get quantized, and unquantized means unadopted.
> Hacker News, and increasingly YouTube/Discord for the practitioner layer.
Source: https://pellmell.ai/s/aefaa217b57ed50be9e2a4b8c9f3173e
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#440Earlier quoted context omitted.
Meta doesn't need to be "redeemed". They have two of the most popular social media apps in the world. And theyll prob survive without ever having you work there
I read that as, "Meta is a piece of shit but they're rich and don't care." (Which I guess I agree with.)
Tech equivalent of "I wouldn't date Sydney Sweeney, I'm not into blondes". Cool story bro