Live data from Hacker News

Muse Spark: Scaling towards personal superintelligence

ai.meta.com

41–50 of 392 posts

Re: Muse Spark: Scaling towards personal superintelligence

#41
post #22

This would have been an amazing release 6 months ago. But the industry moves so fast, this is a trite release. Maybe it’s best for Meta to sell their superintelligence division. I don’t think Zuck’s vision is particularly compelling.

I never understood why meta decided to join the race. They don’t sell compute like Google or Microsoft. Why not let others do the hard work and integrate their LLMs in your systems if needed? I assume it’s because they have Instagram, Facebook, WhatsApp, Thread data and feel they should be the ones using them for training, but it’s really not obvious how having a frontier AI lab benefits their business

Pumps up the stock price.

Re: Muse Spark: Scaling towards personal superintelligence

#42

How is that Meta spent so much money for talent and hardware, but the model barely matches Opus 4.6? Especially, looking at these numbers after Claude Mythos, feels like either Anthropic has some secret sauce, or everyone else is dumber compared to the talent Anthropic has

> has some secret sauce

Yup, it's called test-time compute. Mythos is described as plenty slower than Opus, enough to seriously annoy users trying to use it for quick-feedback-loop agentic work. It is most properly compared with GPT Pro, Gemini DeepThink or this latest model's "Contemplating" mode. Otherwise you're just not comparing like for like.

Re: Muse Spark: Scaling towards personal superintelligence

#43
This really reinforces the idea that the AI race and the Railroad Mania of the 19th century are very similar.

So many different companies are going to have similarly powerful ai that there will be no moat around it and it will be cheap. They will never earn their investment back.

Re: Muse Spark: Scaling towards personal superintelligence

#44

How is that Meta spent so much money for talent and hardware, but the model barely matches Opus 4.6? Especially, looking at these numbers after Claude Mythos, feels like either Anthropic has some secret sauce, or everyone else is dumber compared to the talent Anthropic has

Meta did a bunch of mistakes, and look like Zuckerberg spent a lot of money on talent and made big swings to change it (that happened about a year ago)

I think it’s unrealistic to expect them to come back from that pit to the top in one year, but I wouldn’t rule them out getting there with more time. That’s a possible future. They have the money and Zuckerberg’s drive at the helm. It can go a long way.

Re: Muse Spark: Scaling towards personal superintelligence

#45

How is that Meta spent so much money for talent and hardware, but the model barely matches Opus 4.6? Especially, looking at these numbers after Claude Mythos, feels like either Anthropic has some secret sauce, or everyone else is dumber compared to the talent Anthropic has

It's benchmaxxed.

If they actually matched Opus 4.6 on such a short timeline, it would have been mighty impressive. (Keep in mind this is a new lab and they are prohibited from doing distills.)

Re: Muse Spark: Scaling towards personal superintelligence

#48

Question: since they've rebooted their approach to AI... have they given up on open models? There's no mention of open source or open weights or access to the models beyond their hosted services.

This may be too large to run locally anyway. Maybe they will distill down some smaller open versions later.

Re: Muse Spark: Scaling towards personal superintelligence

#49

How is that Meta spent so much money for talent and hardware, but the model barely matches Opus 4.6? Especially, looking at these numbers after Claude Mythos, feels like either Anthropic has some secret sauce, or everyone else is dumber compared to the talent Anthropic has

It's not even on par with Sonnet. It's on par with open source models and it not even open source and sit behind a private preview API.

Might as well not release anything.

Re: Muse Spark: Scaling towards personal superintelligence

#50

How is that Meta spent so much money for talent and hardware, but the model barely matches Opus 4.6? Especially, looking at these numbers after Claude Mythos, feels like either Anthropic has some secret sauce, or everyone else is dumber compared to the talent Anthropic has

It's benchmaxxed. If they actually matched Opus 4.6 on such a short timeline, it would have been mighty impressive. (Keep in mind this is a new lab and they are prohibited from doing distills.)

how do you know it's benchmaxxed?
Post reply on HN