Live data from Hacker News

Claude Opus 4.1

anthropic.com

241–250 of 344 posts

Re: Claude Opus 4.1

#241

Opus 4(.1) is so expensive[1]. Even Sonnet[2] costs me $5 per hour (basically) using OpenRouter + Codename Goose[3]. The crazy thing is Sonnet 3.5 costs the same thing [4] right now. Gemini Flash is more reasonable[5], but always seems to make the wrong decisions in the end, spinning in circles. OpenAI is better, but still falls short of Claude's performance. Claude also gives back 400's from its API if you CTRL-C in…

Large models are for querying the model

Small models are for querying the context

Opus is cheap if you use it for its niche

Re: Claude Opus 4.1

#242

Earlier quoted context omitted.

In this case, they tried something and were told they were doing it wrong, and they know there's more than one way to do it wrong - wrong model, wrong tool using the model, wrong prompting, wrong task that you're trying to use it for. And of course you could be doing it right but the people saying it works great could themselves be wrong about how good it is. On top of that it costs both money and time/effort investm…

> I think it's pretty different from buying shoes. Shoe shopping is pretty complex, more so than trialing an AI model in my opinion. Are you a construction worker, a banker, a cashier or a driver? Are you walking 5 miles everyday or mostly sedentary? Do you require steel toed shoes? How long are you expecting them to last and what are you willing to pay? Are you going to wear them on long runs or take them river kaya…

Ya know, in the over half a century I've been on this planet, choosing a new pair of shoes is so low on my 'life's little annoyances' list that it doesn't even rise above the noise of all the stupid random things which actually do annoy me.

Maybe the problem is I don't take shoes seriously enough? Something to work on...

Re: Claude Opus 4.1

#243
post #179

Earlier quoted context omitted.

> I think it's pretty different from buying shoes. Shoe shopping is pretty complex, more so than trialing an AI model in my opinion. Are you a construction worker, a banker, a cashier or a driver? Are you walking 5 miles everyday or mostly sedentary? Do you require steel toed shoes? How long are you expecting them to last and what are you willing to pay? Are you going to wear them on long runs or take them river kaya…

> Shoe shopping is pretty complex, more so than trialing an AI model in my opinion. Oh c'mon, now you're just being disingenuous, trying to make an argument for argument's sake. No, shoe shopping is not more complicated than trialing a LLM. For all of those questions about shoes you are posing, either a) a purchaser won't care and won't need to ask them, or b) they already know they have specific requirements and wil…

Just play with the 'free tier' on whatever website does the AI thing and figure it out.

Maybe there's a need to try ten different ones but I just stuck with one and can now convince it to do what I want it to do pretty successfully.

Re: Claude Opus 4.1

#244
post #5

All three major labs released something within hours of each other. This anime arc is insane.

This is why you have PR departments. Being on top of the HN front page, news sites, etc matters a lot. Even if you can't be the first, it's important to dilute the attention as much as possible to reduce the limelight your competitors get.

How do they know when it's time? Corporate espionage? Or do they just have Next Thing queued up months in advance and ready to go.

Re: Claude Opus 4.1

#245
post #79

o3 and o3-pro are just so good. Sonnet goes off the deep end too often and Opus, in my experience, is not as strong at reasoning compared to OpenAI, despite the higher costs. Rarely do we see a worse, more expensive product win - but competition is good and I’m rooting for Anthropic nonetheless!

Off the deep end?

Probably referring to it's tendency to over-complicate things to the point you have to step in and be like "WTF are you even talking about... Wouldn't it be a lot simpler to just use the original, well planned out design?"

Which it does a lot...

Re: Claude Opus 4.1

#246

Earlier quoted context omitted.

Every time that Sonnet is acting like it has brain damage (which is once or twice a day), I switch to Opus and it seems to sort things out pretty fast. This is unscientific anicdata though, and it could just be that switching models (any model) would have worked.

This is a great use case for sub-agents IMO. By default, sub-agents use sonnet. You can have opus orchestrate the various agents and get (close to) the best of both worlds.

Great, now even computers need to leave the IC track if they want continued career progression.

Re: Claude Opus 4.1

#247
post #79

o3 and o3-pro are just so good. Sonnet goes off the deep end too often and Opus, in my experience, is not as strong at reasoning compared to OpenAI, despite the higher costs. Rarely do we see a worse, more expensive product win - but competition is good and I’m rooting for Anthropic nonetheless!

OpenAI also has Flex processing[1] for o3. I've spent most of my time with Gemini 2.5, but lately been trying out a ton of o3 as it seems to work quite well and I get really cheap tokens (~95% of my agentic tokens are cached which is 75% discount and flex mode adds 50% for $0.25 / million input tokens)

[1] https://platform.openai.com/docs/guides/flex-processing?api-...

Re: Claude Opus 4.1

#248
Claude lost me after I used it for a day. Their pricing model is bonkers. There is no way any developer in their right mind would go with Claude.
Post reply on HN