Live data from Hacker News

SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

cognition.com

131–140 of 151 posts

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#131

Earlier quoted context omitted.

200 bucks a month?

Is a lot of money. The majority of people here aren't willing to spend $200/mo for coding unless their little projects provide a comparable value back to them. For context, I'm paying under $30/year and get GLM-5.2. An extra $2300/year isn't going to get me much better outcomes.

it seems cheap for what is borderline AGI

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#132

Earlier quoted context omitted.

Defining what "coding" means now, and how quickly we fall off the capability cliff seems increasingly important. Today my "coding" sessions often enough begin with real life problems, where I discuss domain or inter-domain things, ranging from business, economics, psychology, etc. Being able to do all of that with one model is something I am willing to pay a premium for. Of course not having to pay the premium, becau…

> Today my "coding" sessions often enough begin with real life problems intuition is that your sessions consists of 10% of domain related reasoning, and 90% of code plumbing. Those 90% could be moved to cheap and efficient specialized and focused model.

But that 10% is the most important part! Getting the plumbing wrong means you might have bugs or your code is brittle. Getting the domain-specific business logic wrong means your product doesn't fundamentally solve the correct problem.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#133

https://devin.ai/pricing Apparently 'free' on the $20/mo Devin plan (presumably within some quota still) and that is "via Cerebras at 1000 TPS" according to the announcement I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?

I used SWE-1.5 and 1.6 when it was Windsurf (before Devin Desktop), it's not that bad (grunt work, tests, can actually plan and implement some medium level stuff) but you get a much much better value and better models (GPT-5.4^) going with a Codex subscription (plus you get resets).

That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekly deals reducing credits for premium models. Now it's barely any messages for paid models, completely watered down.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#134

https://devin.ai/pricing Apparently 'free' on the $20/mo Devin plan (presumably within some quota still) and that is "via Cerebras at 1000 TPS" according to the announcement I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?

I used SWE-1.5 and 1.6 when it was Windsurf (before Devin Desktop), it's not that bad (grunt work, tests, can actually plan and implement some medium level stuff) but you get a much much better value and better models (GPT-5.4^) going with a Codex subscription (plus you get resets). That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekl…

FWIW, Cognition has all the Sonnet/Opus/Fable models, and all the GPT ones, as well as GLM, Kimi, and Gemini.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#135

Earlier quoted context omitted.

Defining what "coding" means now, and how quickly we fall off the capability cliff seems increasingly important. Today my "coding" sessions often enough begin with real life problems, where I discuss domain or inter-domain things, ranging from business, economics, psychology, etc. Being able to do all of that with one model is something I am willing to pay a premium for. Of course not having to pay the premium, becau…

but that is not model problem most agentic coding app can use powerful model for planning/reasoning then use "budget" model to do ground work

> most agentic coding app can use powerful model for planning/reasoning then use "budget" model to do ground work

I've had terrible success using budget models to do ground work. The justifications that the budget models will use and document, polluting the rest of the session, are sometimes just insane. Like making code compatible with a bug that was implemented within the same session, not handling errors due to precedence in the code it just implemented, etc. I DO have success using the heavy models with lower effort, and using budget models on relatively changes post ground work. But major planning and initial ground work, I just get absolutely slop if I use a budget model.

If you're doing web stuffs, or GUI, then the budget models seem fine.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#136
post #85

Earlier quoted context omitted.

Dario is convinced that will create SkyNet, and so no, it will never happen. Only the blessed members of the True Church Of Effective Altruism can approach the Ark of the Covenant. The unwashed cannot be trusted.

Rationalism and EA is Scientology for the Bay Area.

It's just the 2020s version of Ayn Rand's "Objectivism." Distillation of exploitative personality traits covered over in sophistry and philosophical excuses. Lets people be dickheads and be smugly superior about it at the same time.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#137

Earlier quoted context omitted.

Is a lot of money. The majority of people here aren't willing to spend $200/mo for coding unless their little projects provide a comparable value back to them. For context, I'm paying under $30/year and get GLM-5.2. An extra $2300/year isn't going to get me much better outcomes.

it seems cheap for what is borderline AGI

The point is that the less capable models are also borderline AGI. You're paying 10-100x more to get a few percentage points improvement in performance.

Put another way, what I get for my under $3/mo is better than what you were getting 3-5 months ago paying $200/mo. So you're paying a lot just to be ahead by a few paltry months.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#138

Earlier quoted context omitted.

And to do that you’ll need development so until we’re all out of a job they’ll keep pushing. Once automating is automated it’s done.

Any day now... Just a bit more space in the context window, trust me bro.

We had decades of "just more and faster bits" in computing and that produced a lot of new capabilities.

Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence

#139
post #85

Earlier quoted context omitted.

Rationalism and EA is Scientology for the Bay Area.

It's just the 2020s version of Ayn Rand's "Objectivism." Distillation of exploitative personality traits covered over in sophistry and philosophical excuses. Lets people be dickheads and be smugly superior about it at the same time.

Creating elaborate ideologies so people can justify their actions with "I'm doing this for you and us all and/or the greater good", is as old as time.
Post reply on HN