Live data from Hacker News

SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

cognition.com

21–30 of 151 posts

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#21

Earlier quoted context omitted.

Funny, the cheerleading at HN for leading Chinese models, but a non Chinese lab (building on top of a Chinese model) gets dissed here.

all the open source models are a waste of time relative to the bleeding edge from openai/anthropic

Not true since a few months, genuinely try GLM 5.2 and Minimax M3, especially in adversarial/gating... as a general model, I can agree, but as a coding model, they are not bad, comparable to maybe Opus 4.5 in real usage which is quite impressive.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#22
post #8

Open source for the win! Imagine how far community might have pushed if 2 past versions of 'morally superior' Anthropic and 'completely Open AI' open sourced their models for the community to build on top of them

Is this open source? I can't find a link to download the weights.

It's based on an open weight model (Kimi 2.7) so shouldn't it also be open weight?

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#23
post #4

These models are never as good, the benchmarks dont tell the full story

Funny, the cheerleading at HN for leading Chinese models, but a non Chinese lab (building on top of a Chinese model) gets dissed here.

It's almost as if HN users aren't all the same.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#24
I'm looking forward to trying this out. I've been using SWE 1.6 quite a lot for grunt work alongside Opus for higher level planning and tricky stuff - a good combo.

As a (former) Windsurf user I'm pretty happy with the progress of the Cognition/Devin ecosystem after they took over Windsurf, now known as Devin Desktop.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#25

A company whose first demo was completely fraudulent announces that its model beats GPT-5.5, on its own benchmark? I’m gonna wait a little before I trust this. This whole company seems to optimize for raising money and impressing VCs. Lying about their products, ignoring consumer market to target enterprise, bragging about how they work their employees like slaves, and writing these posts full of intimidating technic…

Link for this?

https://news.ycombinator.com/item?id=40008109

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#27

We need more models that optimize for coding and that can be cheaper than frontier models, like what SWE 1.7 and composer 2.5 are trying to do. I don't think there's an effort to make something GLM-5.2 level but focused only on coding.

Qwen was doing something like this with their coder models. But alas, they seem not to be releasing those anymore. Last one was Qwen3-coder-next.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#28
post #25

Earlier quoted context omitted.

Link for this?

https://news.ycombinator.com/item?id=40008109

Is it just me or does all that* seem pretty tame by today’s standards? Not saying it’s right, but it barely raises eyebrows. Sounds like a pretty typical startup demo.

* Based on the first comment in the link that claims to summarize the video.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#29

Earlier quoted context omitted.

Funny, the cheerleading at HN for leading Chinese models, but a non Chinese lab (building on top of a Chinese model) gets dissed here.

all the open source models are a waste of time relative to the bleeding edge from openai/anthropic

At work I wouldn't want to use anything else. Compared to my salary a Claude subscription (or two) is cheap

For hobby projects I've completely switched to DeepSeek v4 pro. I spend less than on a $10 Claude plan and am not subjected to quota limits (when I have time and motivation, the last thing I want is a 5 hour quota running out). And the difference in model performance is fine for those smaller projects, most of which will end up abandoned or in a state of "good enough" anyways

And for utility tasks, those 30b models are also great. I'm a big fan of gemma4

Post reply on HN