Earlier quoted context omitted.
Yup, they do quite poorly on random non-coding tasks: https://aibenchy.com/compare/minimax-minimax-m2-7-medium/moo...
It’s worth also comparing Qwen 3.5, it’s a very strong model. Different benchmarks give different results, but in general Qwen 3.5, GLM 5, and Kimi K2.5 are all excellent models, and not too far from current SOTA models in capability/intelligence. In my own non-coding tests, they were better than Gemini 3.1 flash. They’re comparable to the best American models from 6 months ago.
$500 GPU outperforms Claude Sonnet on coding benchmarks
51–60 of 311 posts
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#52Earlier quoted context omitted.
Some of those local model enthusiasts can actually afford solar panels.
You are still incurring a cost if you use the electricity instead of selling it back to the grid
Our peak import rate is 3x higher than our solar export rate. In other words, we’d need to sell 3 kWh hours of energy to offset the cost of using 1 kWh at peak.
We’re currently in the process of accepting a quote for home batteries. The rates here highly incentivise maximising self-use.
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#53Will open source or local llms kill the big AI providers eventually? If so when? I can see maybe basic chat, not sure about coding and images yet
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#54I’d encourage devs to use MiniMax, Kimi, etc for real world tasks that require intelligence. The down sides emerge pretty fast: much higher reasoning token use, slower outputs, and degradation that is palpable. Sadly, you do get what you pay for right now. However that doesn’t prevent you from saving tons through smart model routing, being smart about reasoning budgets, and using max output tokens wisely. And optimiz…
Yup, they do quite poorly on random non-coding tasks: https://aibenchy.com/compare/minimax-minimax-m2-7-medium/moo...
Needless to say, benchmarks are limited and impressions vary widely by problem domain, harness, written language, and personal preference (simplicity vs detail, tone, etc.). If personal experience is the only true measure, as with wine, solving this discovery gap is an interesting challenge (LLM sommelier!), even if model evolution eventually makes the choice trivial. (I prefer Gemini 3 for its wide knowledge, Sonnet 4.6 for balance, and GLM-5 for simplicity.)
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#55Earlier quoted context omitted.
Some of those local model enthusiasts can actually afford solar panels.
You are still incurring a cost if you use the electricity instead of selling it back to the grid
And this is with no income tax or VAT on sold electricity.
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#56Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#57Will open source or local llms kill the big AI providers eventually? If so when? I can see maybe basic chat, not sure about coding and images yet
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#58Will open source or local llms kill the big AI providers eventually? If so when? I can see maybe basic chat, not sure about coding and images yet
Financial gravity will kill them when returns don't match stratospheric expectations.
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#59I’d encourage devs to use MiniMax, Kimi, etc for real world tasks that require intelligence. The down sides emerge pretty fast: much higher reasoning token use, slower outputs, and degradation that is palpable. Sadly, you do get what you pay for right now. However that doesn’t prevent you from saving tons through smart model routing, being smart about reasoning budgets, and using max output tokens wisely. And optimiz…
Re: $500 GPU outperforms Claude Sonnet on coding benchmarks
#60Earlier quoted context omitted.
Financial gravity will kill them when returns don't match stratospheric expectations.
I hope so too, but I think it's wishful thinking. Be prepared for the mother of all financial bailouts from the world governments to make sure that doesn't happen
I hope you are not going to say, "to avoid a global recession or depression caused by the popping of the AI bubble". That would be unnecessary and harmful (in its second-order effects), and governments do have advisors who are competent enough in economics to advise against such a move.