Live data from Hacker News

Claude 4

anthropic.com

11–20 of 1001 posts

Re: Claude 4

#11
> Try Claude Sonnet 4 today with Claude Opus 4 on paid plans.

Wait, Sonnet 4? Opus 4? What?

Re: Claude 4

#16
“GitHub says Claude Sonnet 4 soars in agentic scenarios and will introduce it as the base model for the new coding agent in GitHub Copilot.”

Maybe this model will push the “Assign to CoPilot” closer to the dream of having package upgrades and other mostly-mechanical stuff handled automatically. This tech could lead to a huge revival of older projects as the maintenance burden falls.

Re: Claude 4

#17
> we’ve significantly reduced behavior where the models use shortcuts or loopholes to complete tasks. Both models are 65% less likely to engage in this behavior than Sonnet 3.7 on agentic tasks

Sounds like it’ll be better at writing meaningful tests

Re: Claude 4

#18
Sooo... it can play Pokemon. Feels like they had to throw that in after Google IO yesterday. But the real question is now can it beat the game including the Elite Four and the Champion. That was pretty impressive for the new Gemini model.

Re: Claude 4

#19
Allegedly Claude 4 Opus can run autonomously for 7 hours (basically automating an entire SWE workday).
Post reply on HN