Live data from Hacker News

$500 GPU outperforms Claude Sonnet on coding benchmarks

github.com

141–150 of 311 posts

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#141
post #50

Will open source or local llms kill the big AI providers eventually? If so when? I can see maybe basic chat, not sure about coding and images yet

Some open source models will cross the chasm, some big ai providers will too, and in both case they will have their specific use cases.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#142

Earlier quoted context omitted.

Why is that? The $200 per month subscription comes with a ton of usage. Opus 4.6 is available on the $20 plan too

> The $200 per month subscription comes with a ton of usage. $200 dollars + VAT is half of my rent. I know HN is not a good place to rant on this subject, but I'm often flabbergasted about the number of people here that lives in a bubble with regard to the price of tech. Or just prices in general. I remember someone who said a few years ago (I'm paraphrasing): "You could just use one of the empty room in your house!"…

You think I don't understand that? I'm friends with people who make little more than that amount per month.

But it's not all that relevant to this conversation. It's not like this is the first time economic inequality is a thing.

It's just as relevant to me factoring in your salary the next time I go buy a car.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#143

Earlier quoted context omitted.

"Opus 4.6 is available on the $20 plan too"

Anthropic’s $20 plan gives you such a pittance of tokens that it’s borderline unusable for anything more than a few scripts or a toy app. If $20 is all you have you’d do _much_ better going with chatgpt

That's simply not true at all.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#144

I’d encourage devs to use MiniMax, Kimi, etc for real world tasks that require intelligence. The down sides emerge pretty fast: much higher reasoning token use, slower outputs, and degradation that is palpable. Sadly, you do get what you pay for right now. However that doesn’t prevent you from saving tons through smart model routing, being smart about reasoning budgets, and using max output tokens wisely. And optimiz…

> I’d encourage devs to use MiniMax, Kimi, etc for real world tasks that require intelligence. I use MiniMax daily, mostly for coding tasks, using pi-coding-agent mostly. > The down sides emerge pretty fast: much higher reasoning token use, slower outputs, and degradation that is palpable. I don't care about token use, I pay per request in my cheap coding plan. I didn't notice slower outputs, it's even faster than An…

What is this 10€ per month subscription that you are talking about?

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#145
post #50

Will open source or local llms kill the big AI providers eventually? If so when? I can see maybe basic chat, not sure about coding and images yet

It'd be nice if they do, but I don't really see how. Training these open-weight local LLMs is still insanely expensive and hard to do, even if it's cheaper and faster than what the big corps are doing.

I don't get the financial motive for someone to keep funding these open-weight model training programs other than just purposefully trying to kill the big AI providers.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#146

Earlier quoted context omitted.

> I’d encourage devs to use MiniMax, Kimi, etc for real world tasks that require intelligence. I use MiniMax daily, mostly for coding tasks, using pi-coding-agent mostly. > The down sides emerge pretty fast: much higher reasoning token use, slower outputs, and degradation that is palpable. I don't care about token use, I pay per request in my cheap coding plan. I didn't notice slower outputs, it's even faster than An…

What is this 10€ per month subscription that you are talking about?

MiniMax token plan

https://platform.minimax.io/docs/guides/pricing-token-plan

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#147

Earlier quoted context omitted.

>I'm often flabbergasted about the number of people here that lives in a bubble with regard to the price of tech Sorry, no. You live in the bubble, the people you think are living in a bubble are actually doing the very opposite and taking advantage of the lack of bubbles in our globally connected world. Today, basically anyone can sell any bullshit to billions of people around the world. We’ve never lived in less of…

I guess all those people who live in not-SF just can't be bothered to succeed!

To be fair if you think only people in SF can afford that you do kind of live in a bubble.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#148

Earlier quoted context omitted.

"Opus 4.6 is available on the $20 plan too"

Anthropic’s $20 plan gives you such a pittance of tokens that it’s borderline unusable for anything more than a few scripts or a toy app. If $20 is all you have you’d do _much_ better going with chatgpt

My usage is in the $60 tier, but that doesn't exist so I have to cough up $100. And then get all shaky if I don't use up my weekly quota.

Re: $500 GPU outperforms Claude Sonnet on coding benchmarks

#149
post #148

Earlier quoted context omitted.

Anthropic’s $20 plan gives you such a pittance of tokens that it’s borderline unusable for anything more than a few scripts or a toy app. If $20 is all you have you’d do _much_ better going with chatgpt

My usage is in the $60 tier, but that doesn't exist so I have to cough up $100. And then get all shaky if I don't use up my weekly quota.

Do you mostly just hit the session limits? If so I know it's not ideal but you could wait an hour or two for that to reset. Not sure if that would work for you but just a suggestion
Post reply on HN