This is gonna be game-changing for the next 2-4 weeks before they nerf the model. Then for the next 2-3 months people complaining about the degradation will be labeled “skill issue”. Then a sacrificial Anthropic engineer will “discover” a couple obscure bugs that “in some cases” might have lead to less than optimal performance. Still largely a user skill issue though. Then a couple months later they’ll release Opus 4…
I’m disappointed that this type of discourse has now entered HN. I expected a more evidence-based less “nerf cycle” discussion over here.
Claude Opus 4.5
481–490 of 525 posts
Re: Claude Opus 4.5
#482Re: Claude Opus 4.5
#483With less token usage, cheaper pricing, and enhanced usage limits for Opus, Anthropic are taking the fight to Gemini and OpenAI Codex. Coding agent performance leads to better general work and personal task performance, so if Anthropic continue to execute well on ergonomics they have a chance to overcome their distribution disadvantages versus the other top players.
Re: Claude Opus 4.5
#484Earlier quoted context omitted.
Literally "cancelled" my Anthropic subscription this morning (meaning disabled renewal), annoyed hitting Opus limits again. Going to enable billing again. The neat thing is that Anthropic might be able to do this as they massively moving their models to Google TPUs (Google just opened up third party usage of v7 Ironwood, and Anthropic planned on using a million TPUs), dramatically reducing their nvidia-tax spend. Whi…
I was one of you two, too. After a frustrating month on GPT Pro and a half a month letting Gemini CLI run a mock in my file system I’ve come back to Max x20. I’ve been far more conscious of the context window. A lot less reliant on Opus. Using it mostly to plan or deeply understand a problem. And I only do so when context low. With Opus planning I’ve been able to get Haiku to do all kinds of crazy things I didn’t thi…
Re: Claude Opus 4.5
#485Re: Claude Opus 4.5
#486This is gonna be game-changing for the next 2-4 weeks before they nerf the model. Then for the next 2-3 months people complaining about the degradation will be labeled “skill issue”. Then a sacrificial Anthropic engineer will “discover” a couple obscure bugs that “in some cases” might have lead to less than optimal performance. Still largely a user skill issue though. Then a couple months later they’ll release Opus 4…
Re: Claude Opus 4.5
#487Earlier quoted context omitted.
I’m disappointed that this type of discourse has now entered HN. I expected a more evidence-based less “nerf cycle” discussion over here.
This is nothing new and it's been discussed numerous times. Would you also say we need more evidence that Meta is tracking people?
Re: Claude Opus 4.5
#488Earlier quoted context omitted.
https://www.youtube.com/watch?v=DtePicx_kFY "There's something still not quite right with the current technology. I think the phrase that's becoming popular is 'jagged intelligence'. The fact that you can ask an LLM something and they can solve literally a PhD level problem, and then in the next sentence they can say something so clearly, obviously wrong that it's jarring. And I think this is probably a reflection of…
There is something not right with expecting that artificial intelligence will have the same characteristics as human intelligence. (I am answering to the quote)
Re: Claude Opus 4.5
#489All the users in the comments here complaining about API limits and usage limits have missed the boat. You're not the target audience. This AI is not for you. It's not for consumers and end users. This AI is for the multi-billion and trillion-dollar businesses who are signing massive contracts to get these models enabled for their entire company. I've been using Sonnet 4.5 for months and never had a usage limit ever.…
Re: Claude Opus 4.5
#490Earlier quoted context omitted.
I never claimed that it was being done in secrecy. Here is another example: https://groq.com/blog/inside-the-lpu-deconstructing-groq-spe... . I have seen multiple people mention openrouter multiple times here on HN: https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que... Again, I'm not claiming malicious intent. But model performance depends on a number of factors and the end-user just sees benchmarks for a s…
All those are completely irrelevant. Quantization is just a cost optimization. People are claiming that Anthropic et all changes the quality of the model after the initial release, which is entirely different and the industry as a whole has denied. When a model is released under a certain version, the model doesn’t change. The only people who believe this are in the vibe coding community, believing that there’s some…
For example, in diffusion, there are some models where a Q8 quant dramatically changes what you can achieve compared to fp16. (I'm thinking of the Wan video models.) The point I'm trying to make is that it's a noticeable model change, and can be make-or-break.