Earlier quoted context omitted.
200 bucks a month?
Is a lot of money. The majority of people here aren't willing to spend $200/mo for coding unless their little projects provide a comparable value back to them. For context, I'm paying under $30/year and get GLM-5.2. An extra $2300/year isn't going to get me much better outcomes.
SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
131–140 of 151 posts
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#132Earlier quoted context omitted.
Defining what "coding" means now, and how quickly we fall off the capability cliff seems increasingly important. Today my "coding" sessions often enough begin with real life problems, where I discuss domain or inter-domain things, ranging from business, economics, psychology, etc. Being able to do all of that with one model is something I am willing to pay a premium for. Of course not having to pay the premium, becau…
> Today my "coding" sessions often enough begin with real life problems intuition is that your sessions consists of 10% of domain related reasoning, and 90% of code plumbing. Those 90% could be moved to cheap and efficient specialized and focused model.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#133https://devin.ai/pricing Apparently 'free' on the $20/mo Devin plan (presumably within some quota still) and that is "via Cerebras at 1000 TPS" according to the announcement I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?
That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekly deals reducing credits for premium models. Now it's barely any messages for paid models, completely watered down.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#134https://devin.ai/pricing Apparently 'free' on the $20/mo Devin plan (presumably within some quota still) and that is "via Cerebras at 1000 TPS" according to the announcement I live on Opus 4.8 High and their benchmark scores SWE-1.7 slightly higher ... if at all realistic that sounds like a great deal ... too good to be true?
I used SWE-1.5 and 1.6 when it was Windsurf (before Devin Desktop), it's not that bad (grunt work, tests, can actually plan and implement some medium level stuff) but you get a much much better value and better models (GPT-5.4^) going with a Codex subscription (plus you get resets). That company truly subsidized its user base to the extreme before, the $15/mo subscription was the best value on Earth paired with weekl…
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#135Earlier quoted context omitted.
Defining what "coding" means now, and how quickly we fall off the capability cliff seems increasingly important. Today my "coding" sessions often enough begin with real life problems, where I discuss domain or inter-domain things, ranging from business, economics, psychology, etc. Being able to do all of that with one model is something I am willing to pay a premium for. Of course not having to pay the premium, becau…
but that is not model problem most agentic coding app can use powerful model for planning/reasoning then use "budget" model to do ground work
I've had terrible success using budget models to do ground work. The justifications that the budget models will use and document, polluting the rest of the session, are sometimes just insane. Like making code compatible with a bug that was implemented within the same session, not handling errors due to precedence in the code it just implemented, etc. I DO have success using the heavy models with lower effort, and using budget models on relatively changes post ground work. But major planning and initial ground work, I just get absolutely slop if I use a budget model.
If you're doing web stuffs, or GUI, then the budget models seem fine.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#136Earlier quoted context omitted.
Dario is convinced that will create SkyNet, and so no, it will never happen. Only the blessed members of the True Church Of Effective Altruism can approach the Ark of the Covenant. The unwashed cannot be trusted.
Rationalism and EA is Scientology for the Bay Area.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#137Earlier quoted context omitted.
Is a lot of money. The majority of people here aren't willing to spend $200/mo for coding unless their little projects provide a comparable value back to them. For context, I'm paying under $30/year and get GLM-5.2. An extra $2300/year isn't going to get me much better outcomes.
it seems cheap for what is borderline AGI
Put another way, what I get for my under $3/mo is better than what you were getting 3-5 months ago paying $200/mo. So you're paying a lot just to be ahead by a few paltry months.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#138Earlier quoted context omitted.
And to do that you’ll need development so until we’re all out of a job they’ll keep pushing. Once automating is automated it’s done.
Any day now... Just a bit more space in the context window, trust me bro.
Re: SWE-1.7 Reach Near GPT 5.5 and Opus Intelligence
#139Earlier quoted context omitted.
Rationalism and EA is Scientology for the Bay Area.
It's just the 2020s version of Ayn Rand's "Objectivism." Distillation of exploitative personality traits covered over in sophistry and philosophical excuses. Lets people be dickheads and be smugly superior about it at the same time.