Claude Opus 4.7
11–15 of 15 posts
Re: Claude Opus 4.7
#12For example, SWE-bench Pro improved ~11% compared with Opus 4.6. Should one interpret it as 4.7 is able to solve more difficult problems? or 11% less hallucinations?
Re: Claude Opus 4.7
#13Re: Claude Opus 4.7
#14Re: Claude Opus 4.7
#15Earlier quoted context omitted.
> Not related to this release, but is anyone aware of what's happening with Deepseek? Given that no-one is talking about DeepSeek, I assume it is coming this month. They are still releasing research papers and that is what really matters and not the .1 increment releases of AI models to massage benchmarks or create hype around.
There's been months of "DeepSeek v4 next week!" rumours and none have panned out. They're either stuck/dead or they're sitting on something really fantastic that they only want to release once they've perfected it. My realistic side thinks the former, my optimism on the latter. In the meantime, GLM 5.1 is actually really good.