Claude 4
51–60 of 1001 posts
Re: Claude 4
#52Allegedly Claude 4 Opus can run autonomously for 7 hours (basically automating an entire SWE workday).
Re: Claude 4
#53I'm happy that tool use during extended thinking is now a thing in Claude as well, from my experience with CoT models that was the one trick(tm) that massively improves on issues like hallucination/outdated libraries/useless thinking before tool use, e.g.
o3 with search actually returned solid results, browsing the web as like how i'd do it, and i was thoroughly impressed – will see how Claude goes.
Re: Claude 4
#54I want GenAI to become better at tasks that I don't want to do, to reduce the unwanted noise from my life. This is when I'll pay for it, not when they found a new way to cheat a bit more the benchmarks.
At work I own the development of a tool that is using GenAI, so of course a new better model will be beneficial, especially because we do use Claude models, but it's still not exciting or interesting in the slightest.
Re: Claude 4
#55Re: Claude 4
#56I've found myself having brand loyalty to Claude. I don't really trust any of the other models with coding, the only one I even let close to my work is Claude. And this is after trying most of them. Looking forward to trying 4.
Re: Claude 4
#57Re: Claude 4
#58Love to try the Claude Code VScode extension if the price is right and purchase-able from China.
Re: Claude 4
#59Re: Claude 4
#60I'll look at it when this shows up on https://aider.chat/docs/leaderboards/ I feel like keeping up with all the models is a full time job so I just use this instead and hopefully get 90% of the benefit I would by manually testing out every model.