Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
301–310 of 682 posts
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#302I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…
This is Apache 2.0, which is quite permissive. Just accept the gift. These kind of responses are hilarious. Someone gives something for free (and indeed this is entirely free) and the top comment is pure complaint.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#303I lament the comments saying this in any way redeems Meta (the company). The researchers releasing this stuff have almost nothing to do with Meta other than being bankrolled by the slaughterhouse. You aren't the customer, you are the pawn in big tech's game of thrones. Your good will is a commodity to be traded, almost literally. It will be used against you the moment it's convenient. This is open weights because Met…
How is this non-sequitor the top comment?
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#304Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#305Meta is rocking AI. As of last week I have been using their excellent muse coding harness with their model Muse Spark 1.2. Starting this morning I am running their new local 30B model muse-glimmer on my old MacMini 32G using Ollama (remember to increase the context size!) and pi coding harness. I am getting good results with muse-glimmer running locally, with the caveat that everything runs slowly (e.g., give it a ta…
Newb question but I’m curious what would help it to run faster? Would it need more vRAM or just system memory?
I have two GPU rigs both with 2x RTX Pro 6000, can get ~250 tk/s decode with deepseek-v4-flash in native mixed precision. For context, in antirez's dwarfstar project he only gets ~20-40 tk/s on the same model @ 2bpw on M5 Max.
The latter is for sure usable if it's your only option, but it's really hard for me to personally go back to speeds like that when I've experienced the former.
(Also worth noting dwarfstar only has experimental support for dspark spec dec, when that lands it will definitely give a big boost at higher acceptance rates)
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#306Meta is rocking AI. As of last week I have been using their excellent muse coding harness with their model Muse Spark 1.2. Starting this morning I am running their new local 30B model muse-glimmer on my old MacMini 32G using Ollama (remember to increase the context size!) and pi coding harness. I am getting good results with muse-glimmer running locally, with the caveat that everything runs slowly (e.g., give it a ta…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#307Earlier quoted context omitted.
> It's common, if not inevitable, for people who feel strongly about $topic to conclude that the system (or the community, or the mods, etc.) are biased against their side. One is far more likely to notice whatever data points that one dislikes because they go against one's view and overweight those relative to others. This is probably the single most reliable phenomenon on this site. Keep in mind that the people wit…
I'm certainly pro-America and anti-communist/fascist as a bias, but I don't really care about whether AI tools are open-weight or not. I just use the products that best fit my needs. I just think the arguments put forth regarding China and open-weight models and strategy are not very good. It just so happens that China is the only real competitor in the AI space and so they are who get talked about the most in compar…
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#308Earlier quoted context omitted.
Do AI companies make release plans based on upcoming other models like this? I would think all the processes that go into the repository and weight infrastructure pre-training, checkpointing, knowledge distillation, model compression, post training pipeline, ecosystem integrations, inference API, benchmarking, human eval/safety/alignment, docs, etc... all that dictates the release schedule.
There has been a long history of AI model releases made shortly before or after a major planned release by another company. Almost always to upstage or steal thunder. Just recently, Minimax H3 released as open weights on the eve of Seedance 2.5 global availability. It's not as good, but it's good enough and it's completely open. Flux 3, which is nowhere near as good as either, suddenly announced their release once ne…
Minimax H3 can run exceptionally fast (10 minutes for a 15 second 0.5mp video and that's stock cuda 13), works on 16 GB VRAM GPUs, etc. If Flux3 is anything like Flux2, it’s going to require an absolute monster truck of a machine and still run significantly slower. Even if it’s a better model, that won’t matter as much if nobody releases any LoRAs or fine-tunes for it.
Not to mention BFL licensing often feels deceptively confusing and restrictive.
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#309Earlier quoted context omitted.
It’s rather amusing to me to read comments like this, and then simultaneously whenever a Chinese company or team releases open-weight models or whatever there is a giant round of applause, America is so behind, and there’s nothing but positive things to say about the intelligent, creative, and well-intentioned Chinese engineers (which is true, America certainly doesn’t have a monopoly on great people). Don’t you know…
There is a big astroturfing going on social media platforms by the chinese. Did you notice 'day in a life of unmarried 30 yr old lady in china' videos flooding usa social media. Regular ppl in the west now hold mildly positive views of the ccp and how 'advanced' china is than usa. Then there are europeans who now are looking for china to give them the technology handout now that relationship with usa has soured.
See:
https://collaborate.princeton.edu/en/publications/decentrali...
https://yiqingxu.org/replications/lpxx2025.html
https://cira.exovera.com/conferences/engineers-of-the-human-...
Re: Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
#310Earlier quoted context omitted.
I was about a week away from buying a very tricked out MacBook Pro with 128 GB RAM, but was on vacation and worried about it arriving while I was away, and then the price hikes went into effect. Grumble. Oh, well. Serves me right.
Lol, I still think about buying that now, even after the price hike. FOMO.