Live data from Hacker News

Muse Code and Muse Spark 1.2

research.meta.ai

131–140 of 265 posts

Re: Muse Code and Muse Spark 1.2

#132
post #76

Earlier quoted context omitted.

DeepSeek is really crazy cheap, though, and they don't have a giant pool of other invasive personal data to correlate it with.

I hate to say this and this is because I fucking despise meta. But between DeepSeek and Meta, and trust they handle the training data correctly, I trust meta.

Right about DeepSeek, but with their history of handling data, I would never trust Meta either.

Re: Muse Code and Muse Spark 1.2

#133
post #116

Earlier quoted context omitted.

when has Meta ever not broken their contractual obligations (I am being serious here)? are we seriously discussing/expecting any sort of privacy related to Meta? you can pay whatever they want, they will train and use your data, I figured this is not something that should be discussed but obviously I have been mistaken...

This is one of those situations where we are both completely baffled by the position held by the other person. The fact that Facebook has so much experience taking advantage of people's private data is one of the reasons I believe them when they say they won't be doing it when you pay them for that service.

How does money change that trust? They certainly have breached their word on this in the past (for non-paying users of Facebook).

Re: Muse Code and Muse Spark 1.2

#136
post #117

Unfortunately I find this too high risk, I entered my credit card, but can not set a limit. The best I can do is get an email alert. I feel like I am one oopsie away from getting a 100 dollar bill.

Yep - exactly my thoughts. No way am I trying this out without a billing limit - seems crazy given the pricing strategy to not have a top-up and pay as you go.

Re: Muse Code and Muse Spark 1.2

#137
post #126

Earlier quoted context omitted.

1.2 is now on openrouter. You could try it there, they have hard limits per key.

absolutely and fair... you do lose out though on the mega discounted endpoint they have there.

Virtual card with a limit?

Re: Muse Code and Muse Spark 1.2

#138

They chose to compare against Open AI’s mid tier model Terra instead of Sol and still lost some benchmark against it. They left Opus in and got beat in all but one benchmark. Nothing wrong with trying to improve, but why the marketing games? Instead of trying to say in the post you’re “closer” to frontier, first set a clear goal to beat the Chinese labs on price or performance and demonstrate it convincingly. Then wh…

It’s pretty clear they’re really not attempting to compete that way. They’re using a profitable ad business to be able to undercut and buy some business to stay relevant.

Re: Muse Code and Muse Spark 1.2

#139
post #94
post #56

Earlier quoted context omitted.

My conclusion is the opposite. If benchmarks were meaningless, surely Meta would be able to find some benchmark that shows they are better than Sol and Fable. The fact that they can't do that tells me that benchmarks still do mean something.

Muse 1.1 performed relatively well according to benchmarks, putting it within spitting distance of the premier models. However, based on the results I got from it and the review videos I watched, it wasn’t even close. Opus 5 is incredible at making games. Almost like a generation better than other models from my experience. You won't see that if you just look at the popular benchmarks.. You have to test each model on…

> Opus 5 is incredible at making games.

This is a bit vague. What sort of games with what technology?

Re: Muse Code and Muse Spark 1.2

#140

I am cynical about this, but I would be careful about giving Meta access even to my open source code by myself.

You think OpenAI, Anthropic... are good guys??

We’re all trying to keep track of who is least bad in various ways at any moment in time.

There’s nothing wrong with that, given the landscape we’re all living in.

Post reply on HN