Live data from Hacker News

SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

cognition.com

61–70 of 151 posts

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#61
post #17

Not finding anything about this while searching huggingface: https://huggingface.co/search/full-text?q=SWE-1.7 i assume this is another closed source model?

Yes, and not only that but you can't even access it via API, you can only use it in Devin (formerly Windsurf).

I'm an OpenCode user, but I'll fall back to Claude Code if I want to use Opus end to end for something, given my company has a subscription. But I'm not using yet another tool and subscription for a model that isn't even winning.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#63

Earlier quoted context omitted.

Qwen was doing something like this with their coder models. But alas, they seem not to be releasing those anymore. Last one was Qwen3-coder-next.

Its crazy that OpenAI and Anthropic themselves aren't doing that. No attempts at reducing inference cost for code as far as I know from them.

OpenAI do have codex models, which are half the price. I haven't used them enough to comment on the quality though.

I remember them saying a few years ago that, they didn't think it was worth specializing models for code, because their general purpose models kept beating them. I guess they changed their mind? Since they did start making codex models again.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#64

Earlier quoted context omitted.

yeah but why waste your time on these models, just use the one that gets the better results

I was going to respond until I saw your account name lol.

haha i outsource my thinking to the smartest model

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#65
post #60

While I am skeptical of the results here, I am very excited for this new trend of making models faster. Running capable models at 1k TPS is more valuable for me than running better models at 30 TPS. I can only imagine the trend continues to move from "let's only make models smarter" to just incremental intelligence gains but with step improvements in speed.

Why? I'm personally on the opposite end. Less babysitting/higher quality means more time goes back to me/the user. 1000tps of bad code means you have to keep validating the output and circling back.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#66
post #60

While I am skeptical of the results here, I am very excited for this new trend of making models faster. Running capable models at 1k TPS is more valuable for me than running better models at 30 TPS. I can only imagine the trend continues to move from "let's only make models smarter" to just incremental intelligence gains but with step improvements in speed.

Indeed. For me opus 4.8 is good enough. If only it would be 100 times faster. You could run it in self verification loops much much faster. It sometimes takes 15 minutes for me to complete a simple task. For example configuring AWS agentcore and deploying an agent on it. Takes forever with Claude with constant issues it tries to solve.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#67
post #63

Earlier quoted context omitted.

Its crazy that OpenAI and Anthropic themselves aren't doing that. No attempts at reducing inference cost for code as far as I know from them.

OpenAI do have codex models, which are half the price. I haven't used them enough to comment on the quality though. I remember them saying a few years ago that, they didn't think it was worth specializing models for code, because their general purpose models kept beating them. I guess they changed their mind? Since they did start making codex models again.

They also stopped making them again.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#68
https://devin.ai/pricing

Apparently 'free' on the $20/mo Devin plan (presumably within some quota still)

and that is "via Cerebras at 1000 TPS" according to the announcement

I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#69

Would have been worth a consideration if it could have been used beyond it's own harness. Unfortunately, doesn't seem to be the case. https://x.com/theodormarcu/status/2074896486047834380

Ugh, that changes everything. If I wanted an arranged marriage I could go back to Claude Code.
Post reply on HN