Like, stop toying around with token limits and just focus on more efficient models.
Once the local models are good enough we are so abandoning these elephants.
I'm gonna walk as soon as possible.
71–80 of 275 posts
Like, stop toying around with token limits and just focus on more efficient models.
Once the local models are good enough we are so abandoning these elephants.
I'm gonna walk as soon as possible.
Anthropic seems cooked right now in terms of compute
(nobody uses claude anymore, it's too over-subscribed)
Earlier quoted context omitted.
During this period however they’ve released models which lots of people report to be significantly more chatty. Framing it purely as a promotion ending feels like you’re giving them a bit too much credit. I doubt it’s a coincidence they started this promotion a week before the release of 4.8.
I think running a model unsupervised in anything above medium is a sucker move to burn more tokens. The high effort models can be great in limited context, but unsupervised they too often end up navel gazing. High doesn't always mean smarter, but it always burns more tokens. Medium and low seem to be decent for day to day tasks.
Once on a sandbox, realizes nothing works in its sandbox, then again outside the sandbox.
Fucking hell.
Looking like this will be the last month with Anthropic. Between the outages and just overall crap utility of Opus/Fable lately...
They'll all be getting around to this sooner or later, its just too expensive
It's hard to put a number on it, but even accounting for all the time in meetings, talking with stakeholders and developers etc I'm over 10% more productive overall. I earn substantially more than $2000/month, the ROI is there
It's only expensive compared to the currently very strong offering from OpenAI. Or other models - my hobby projects are all on DeepSeek
Anthropic seems cooked right now in terms of compute
Sideline LLMs have free compute and offer cheap prices to draw in crowd. Becomes flavor of the month LLM.
Back to step one.
It should be pretty clear by now that token prices are predominately a function of available compute.
I got a max account during this promotion and it's been very fun but I'm a little burned out and am kind of looking forward to going back and tinkering with game engines without the help of an LLM for code gen (will still use it for documentation questions but I can do that with the free tier)
For fun/side stuff, handwritten code makes total sense, if you’re burnt out. For anything that makes money, it feels like a huge step down in productivity, even with the downsides of agent-assisted code.
The max account is for home projects where I'm basically making some common tools but tailored to myself (diet tracker, note app etc).
It's been great but across work and home it's just too much.
Looking like this will be the last month with Anthropic. Between the outages and just overall crap utility of Opus/Fable lately...
also codex + sol is really good and quite fast.
I think the difference between the Anthropic token maximization approach (vibe code all the things!) and OpenAI's focus on efficiency, terseness and token reduction are going to be the defining features of who wins the long-term race. My money is on the more efficient solution. Even if Anthropic can win some benchmarks by using 3x tokens over 3x time, it is a terrible base to build toward the future. Users are no lon…
I think the difference between the Anthropic token maximization approach (vibe code all the things!) and OpenAI's focus on efficiency, terseness and token reduction are going to be the defining features of who wins the long-term race. My money is on the more efficient solution. Even if Anthropic can win some benchmarks by using 3x tokens over 3x time, it is a terrible base to build toward the future. Users are no lon…
Isn’t the race between Chinese open-weight models and the others more decisive for the future?