Live data from Hacker News

SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

cognition.com

141–150 of 151 posts

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#141

Earlier quoted context omitted.

I used SWE-1.5 and 1.6 when it was Windsurf (before Devin Desktop), it's not that bad (grunt work, tests, can actually plan and implement some medium level stuff) but you get a much much better value and better models (GPT-5.4^) going with a Codex subscription (plus you get resets). That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekl…

FWIW, Cognition has all the Sonnet/Opus/Fable models, and all the GPT ones, as well as GLM, Kimi, and Gemini.

But they don't appear to subsidize them to the same degree. I've only been using Devin for less than a month, but I've been hitting the limits of the $20/month plan way more quickly than I'd expect, and definitely more quickly than with Claude Code or Codex.

So far, Cursor provides the best value for their subscription, but I have to imagine they're basically lighting money on fire. There's no way their current pricing is sustainable.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#142

Earlier quoted context omitted.

> Today my "coding" sessions often enough begin with real life problems intuition is that your sessions consists of 10% of domain related reasoning, and 90% of code plumbing. Those 90% could be moved to cheap and efficient specialized and focused model.

But that 10% is the most important part! Getting the plumbing wrong means you might have bugs or your code is brittle. Getting the domain-specific business logic wrong means your product doesn't fundamentally solve the correct problem.

[dead]

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#144

https://devin.ai/pricing Apparently 'free' on the $20/mo Devin plan (presumably within some quota still) and that is "via Cerebras at 1000 TPS" according to the announcement I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?

The "Lightning" (Cerebras) variant isn't free, only the regular one, which runs closer to 50 TPS in my experience with SWE 1.6.

Oh, the free plan said "Slow" so I thought maybe the others had the fast version :)

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#145

Earlier quoted context omitted.

I used SWE-1.5 and 1.6 when it was Windsurf (before Devin Desktop), it's not that bad (grunt work, tests, can actually plan and implement some medium level stuff) but you get a much much better value and better models (GPT-5.4^) going with a Codex subscription (plus you get resets). That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekl…

FWIW, Cognition has all the Sonnet/Opus/Fable models, and all the GPT ones, as well as GLM, Kimi, and Gemini.

It's not very clear on the website, couldn't find a list of models or quota/pricing per model anywhere

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#146
it is obvious at this stage that most of the gains are in distillation post training and having good RL simulations. the moat of the private labs are just their capability of stealing open research and data and locking the few bit they innovate on the top of it.

It will work until they IPO.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#147
post #95

I’ve unfortunately had to temper my excitement with Cognition’s models/products given the amount of unwarranted hype they created with Devin on first release, but hopefully this is good.

I work with Devin daily (as well as Claude and a few others) and I can attest it's not a cheap product but it's a good one and it saves me a bunch of time.

Yea I cant convince myself to pay api pricing unless my employer is doing it

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#148
post #11

Unrelated: what's the point of "*equal contribution"? Why would someone specify this

Because papers are often referred to by the first author’s name, and often the first author is the primary researcher and therefore deserves the extra credit. When two or more primary authors are equally involved, they’ll often do a random ordering but annotate this so that no one thinks one did more than the others.

some journals and colleges actually have a policy to always use random order to help fight the "senior researcher gets all the credit" culture in academia.

theres a lot of cases where a prof forced their students to put them first even if they had an advisor role, or even credit someone for zero real work because they threatened to block submission and prevent the students from getting their degree.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#150
post #149

Heads up to anyone else curious, I installed the Devin CLI and SWE-1.7 is not currently available there.

It is now, at least for me. Latest Devin CLI version: v3000.1.27 I don't think SWE is geo-restricted.

Looks like now it shows up but it's not available to try on a free plan.
Post reply on HN