Live data from Hacker News

Claude Opus 4.6

anthropic.com

621–630 of 1001 posts

Re: Claude Opus 4.6

#621
post #562

Earlier quoted context omitted.

It didn't use web search. But for sure it has some internal knowledge already. It's not a perfect needle in the hay stack problem but gemini flash was much worse when I tested it last time.

I think the OP was implying that it's probably already baked into its training data. No need to search the web for that.

[deleted]

Re: Claude Opus 4.6

#622
post #194
post #71

Earlier quoted context omitted.

It’s extremely successful, not sure what it explains other than your biases

Anthropic has perhaps the most embarrassing status page history I have ever seen. They are famous for downtime. https://status.claude.com/

Shades of Fail Whale

Re: Claude Opus 4.6

#623

Earlier quoted context omitted.

after the models get good enough to replace coders they will be able to start increasing the subscriptions back up

At $100k/yr the joke that AI means "actual Indians" starts to make a lot more sense... it is cheaper than the typical US SWE, but more than a lot of global SWEs.

No - because the AI will be super human. No human even at $1mm a year would be competitive with a $100k/yr corresponding AI subscription.

See people get confused. They think you can charge __less__ for software because it's automation. The truth is you can charge MORE, because it's high quality and consistent, once the output is good. Software is worth MORE than a corresponding human, not less.

Re: Claude Opus 4.6

#624

I just tested both codex 5.3 and opus 4.6 and both returned pretty good output, but opus 4.6's limits are way too strict. I am probably going to cancel my Claude subscription for that reason: What do you want to do? 1. Stop and wait for limit to reset 2. Switch to extra usage 3. Upgrade your plan Enter to confirm · Esc to cancel How come they don't have "Cancel your subscription and uninstall Claude Code"? Codex last…

How else are they going to supplement their own development expenses? The more Claude Anthropic needs the less Claude the customer will get. By their own admission that is how the Anthropic model works. Their end value is in using vibe coders and engineers alike to create a persistent synthetic developer that replaces their own employees and most of their customers.

Scalable Intelligence is just a wrapper for centralized power. All Ai companies are headed that way.

Re: Claude Opus 4.6

#625

I'm still not sure I understand Anthropic's general strategy right now. They are doing these broad marketing programs trying to take on ChatGPT for "normies". And yet their bread and butter is still clearly coding. Meanwhile, Claude's general use cases are... fine. For generic research topics, I find that ChatGPT and Gemini run circles around it: in the depth of research, the type of tasks it can handle, and the qual…

Why would I even use Claude for asking something on their web, considering that chips away my claude code usage limit?

Their limit system is so bad.

Re: Claude Opus 4.6

#626
I feel like I can't even try this on the Pro plan because Anthropic has conditioned me to understand that even chatting lightly with the Opus model blows up usage and locks me out. So if I would normally use Sonnet 4.5 for a day's worth of work but I wake up and ask Opus a couple of questions, I might as well just forget about doing anything with Claude for the rest of the day lol. But so far I haven't had this issue with ChatGPT. Their 5.2 model (haven't tried 5.3) worked on something for 2 FREAKING HOURS and I still haven't run into any limits. So yeah, Opus is out for me now unfortunately. Hopefully they make the Sonnet model better though!

Re: Claude Opus 4.6

#628

Earlier quoted context omitted.

Claude sucks at non English languages. Gemini and ChatGPT are much better. Grok is the worst. I am a native Czech speaker and Claude makes up words and Grok sometimes respond in Russian. So while I love it for coding, it’s unusable for general purpose for me.

Claude code (opus) is very good in Polish. I sometimes vibe code in polish and it's as good as with English for me. It speaks a natural, native level Polish. I used opus to translate thousands of strings in my app into polish, Korean, and two Chinese dialects. Polish one is great, and the other are also good according to my customers.

Your game is amazing!

I wish there was a "Reset" button to go back to the original position.

Where are you in Poland?

Re: Claude Opus 4.6

#629

Earlier quoted context omitted.

I kinda agree. Their model just doesn't feel "daily" enough. I would use it for any "agentic" tasks and for using tools, but definitely not for day to day questions.

Why? I use it for all and love it. That doesn't mean you have to, but I'm curious why you think it's behind in the personal assistant game.

My 2 cents:

All the labs seem to do very different post training. OpenAI focuses on search. If it's set to thinking, it will search 30 websites before giving you an answer. Claude regularly doesn't search at all even for questions it obviously should. It's postraining seems more focused on "reasoning" or planning - things that would be useful in programming where the bottleneck is: just writing code without thinking how you'll integrate it later and search is mostly useless. But for non coding - day to day "what's the news with x" "How to improve my bread" "cheap tasty pizza" or even medical questions, you really just want a distillation of the internet plus some thought

Re: Claude Opus 4.6

#630

I wonder if I’ve been in A/B test with this. Claude figured out zig’s ArrayList and io changes a couple weeks ago. It felt like it got better then very dumb again the last few days.

[dead]

What companies do you interact with that don’t A/B test?
Post reply on HN