Live data from Hacker News

Muse Spark 1.3

developer.meta.com

151–160 of 475 posts

Re: Muse Spark 1.3

#151
post #70

Earlier quoted context omitted.

Gotta be honest that I’m tired of the “I hate Zuck and Meta so much” comments every time Meta does anything. Ditto Elon/X. Fine, I get it. I don’t like Zuck either. But the post is about Muse Spark 1.3. What do you think about that? If you don’t like it because Meta made it, then maybe just don’t use it and stay silent.

What I'm tired of is the top story (or five) on HN every day announcing Spark Opus Fable Grok Gemini v4.1i3-F. Like, who actually cares? Are people excited for the new benchmarks? Is it interesting to read the model cards? And look, part of my job is to use these things and part of my job is to pick EC2 servers, too. The front page of HN is increasingly resembling one of those endless AWS pricing lists. And yeah, I d…

I think a lot of people are curious where the "knee" is on gains and productivity, particularly in the agentic space, which is where the real value is. A lot of us are being forced to shoe-horn this stuff into existing products, and knowing how much of the task the model can do now, vs having to build a complex custom harness, is valuable information to have. A year and a half ago it took our dev maybe six weeks of struggling with LangChain to approximate what Claude + MCP server can do today. The MCP server took us perhaps 2 days to build and 3 more to get it production ready. Today that MCP server gets 2-3 commits per month. I absolutely want to know when new models come out.

As for smaller models, we run a pretty wide variety of agentic workload doing data enrichment and, increasingly, a bunch of evaluation jobs to alert a human to review certain scenarios etc. These all run on the smaller 27B and 35B class models, and tooling behavior has improved DRAMATICALLY since april. The latest qwen 3.8 model has a 95% success tool call rate during internal testing and about 94% real world. That's about 3% better than the 35B-A3B model we're using today, but the 35B MoE is so much faster then 3% is worth the trade-off.

Re: Muse Spark 1.3

#153
post #141

Earlier quoted context omitted.

Dario's idea of an ethical focus seems to be keeping powerful models out of the hand of anyone unethical, which coincidentally is everyone except him.

Yeah this would be a great point if it were true and they didn’t give Mythos access to companies to fix bugs, which they did and have. It’s genuinely a difficult question. Not black and white. The models are really good at finding bugs, as demonstrated by people using Fable to reverse engineer. People make it sound like he’s just making it up.

This would be more convincing if mythos was something uniquely special and not something merely a couple months ahead of everyone else. It was great marketing though.

Re: Muse Spark 1.3

#154
post #150
post #47

Meta is one of those companies where, if there is anything remotely comparable, I'm happy to pay more to not use them. They've had a profoundly negative impact on society and Zuckerberg is not who I want controlling the future at the top of AI. I feel the same about Grok w/ Elon. I will pay extra to use someone else. I'm not an Amodei stan, but of all of these people he seems to have the most ethical focus. Again, no…

Strong disagree with the Anthropic being good at all part. This is not defending anyone else, but… Anthropic leadership repeatedly presents themselves as uniquely morally qualified to steward agi and decide how humanity should get access to it. Yet they have repeatedly failed basic morality tests. Pirating books for financial gain. The newer Sony/Warner music case shows this is pattern behavior. Aggressively scraping…

It makes me sad that people don’t see right through Anthropic’s gambit.

They want to position AI as an insurmountable threat in order to regulate away any future competitors. They’re trying to speedrun regulatory capture.

Re: Muse Spark 1.3

#155
post #141

Earlier quoted context omitted.

Dario's idea of an ethical focus seems to be keeping powerful models out of the hand of anyone unethical, which coincidentally is everyone except him.

Yeah this would be a great point if it were true and they didn’t give Mythos access to companies to fix bugs, which they did and have. It’s genuinely a difficult question. Not black and white. The models are really good at finding bugs, as demonstrated by people using Fable to reverse engineer. People make it sound like he’s just making it up.

They gave access, but considering that they wouldn't even sign the "don't ban open weights" letter, it's clear they would prefer to have tight control over who they bless with that access.

Re: Muse Spark 1.3

#156
post #3

llm -m meta-ai/muse-spark-1.3 "Generate an SVG of a pelican riding a bicycle" https://tools.simonwillison.net/markdown-svg-renderer?url=ht... 4.2266 cents, 38 seconds. For comparison here's Muse Spark 1.2, which animated it without me asking it to: https://tools.simonwillison.net/markdown-svg-renderer?url=ht... The 1.3 one is definitely better - better bicycle frame, better wing, better pelican hat. UPDATE: Here's an…

Is there any point anymore regarding this svg test? I would not be surprised if in the training they're fine tuned for this task too

You win this thread's prize:

https://news.ycombinator.com/item?id=49538333

Re: Muse Spark 1.3

#157
post #5
post #3

llm -m meta-ai/muse-spark-1.3 "Generate an SVG of a pelican riding a bicycle" https://tools.simonwillison.net/markdown-svg-renderer?url=ht... 4.2266 cents, 38 seconds. For comparison here's Muse Spark 1.2, which animated it without me asking it to: https://tools.simonwillison.net/markdown-svg-renderer?url=ht... The 1.3 one is definitely better - better bicycle frame, better wing, better pelican hat. UPDATE: Here's an…

Is there a reason these pelicans always have roughly the same composition (side-view, 2d, biking right, flat ground beneath, etc)? I don't see any of that detailed in the prompt, yet they all seem to generate roughly the same image of differing quality.

Not when rendered via POV-Ray:

https://blog.nawaz.org/posts/2025/Oct/pelican-on-a-bike-rayt...

I plan to update it with more pelicans from all the models released since.

(Spoiler alert: They haven't improved much since then).

Re: Muse Spark 1.3

#158
As a product, would developers switch to a meta model/harness? I don’t think so.

Only way I see is if it becomes the new SOTA / frontier, does anyone think Meta will surpass Anthropic or OpenAI?

I still can’t get my head around why language models are an existential threat to Meta - they own the platforms people watch adds on?

Re: Muse Spark 1.3

#159
post #47

Meta is one of those companies where, if there is anything remotely comparable, I'm happy to pay more to not use them. They've had a profoundly negative impact on society and Zuckerberg is not who I want controlling the future at the top of AI. I feel the same about Grok w/ Elon. I will pay extra to use someone else. I'm not an Amodei stan, but of all of these people he seems to have the most ethical focus. Again, no…

> he seems to have the most ethical focus

He wants to build a tech-god kept in chains whose power he parcels out to the unwashed masses he deems worthy like some sort of high priest of intelligence.

And that is being charitable and going by the interpretation that he actually believes what he says.

Re: Muse Spark 1.3

#160
post #150

Earlier quoted context omitted.

Strong disagree with the Anthropic being good at all part. This is not defending anyone else, but… Anthropic leadership repeatedly presents themselves as uniquely morally qualified to steward agi and decide how humanity should get access to it. Yet they have repeatedly failed basic morality tests. Pirating books for financial gain. The newer Sony/Warner music case shows this is pattern behavior. Aggressively scraping…

It makes me sad that people don’t see right through Anthropic’s gambit. They want to position AI as an insurmountable threat in order to regulate away any future competitors. They’re trying to speedrun regulatory capture.

It is so obvious yet most people don't want to see it.
Post reply on HN