Anthropic is a bit nuts, I had $260 of credits on my max account for the extra usage the other night. It was expiring, so I figured I'll fire up an agentic swarm to deep dive and make some deep changes to some old cold bases.. literally 25 minutes or less, $260 burnt, it didn't get get into the implementation, just wrote a ton of useless plans for the most part. It really opened my eyes to what they expect to charge…
And at that cost they're still not profitable. It's going to be a bumpy road ahead...
Qwen3.8 Max now ranked as the best overall model by agentic index
111–120 of 364 posts
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#112Anthropic is a bit nuts, I had $260 of credits on my max account for the extra usage the other night. It was expiring, so I figured I'll fire up an agentic swarm to deep dive and make some deep changes to some old cold bases.. literally 25 minutes or less, $260 burnt, it didn't get get into the implementation, just wrote a ton of useless plans for the most part. It really opened my eyes to what they expect to charge…
And at that cost they're still not profitable. It's going to be a bumpy road ahead...
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#113Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#114Earlier quoted context omitted.
VCs are footing the bill for that $200 subscription.
They have something like 80% gross margins, are at a $100B/yr ARR, and are growing at 10x per year... If that keeps up, they're going to be doing more revenue than Google in a year ($400B ARR, 20% per year growth)
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#115Earlier quoted context omitted.
And at that cost they're still not profitable. It's going to be a bumpy road ahead...
What’s the blast radius of this bubble popping? It’s all private investment still right?
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#116Opus is still first in Intelligence Index followed by Fable, GPT 5.6, Kimi K3 then Qwen 3.8 max. https://artificialanalysis.ai/#intelligence Our leaderboard combines Arena ELO, AA Intelligence index, latency and speed and goes: #1 Opus 5 #2 Kimi K3 #3 Qwen3.8 Max #4 GPT 5.6 Sol Source: http://pellmell.ai/leaderboard . This jumps around a lot based on the top throughput and latency of whatever provider happens to be b…
Opus-5 is practically unusable (for complex tasks) in this sense - its updates are voluminous, and dense with cryptic language (there are numerous reddit threads complaining about this, so it's not just me). I often have to ask it to re-state concisely in plain terms.
For a fairly gnarly task, after fighting with with Claude-Code + Opus-5, I ported my session to Codex + GPT-5.6-sol, and it was like a breath of fresh air.
Arguably a key aspect of intelligence is concise, clear communication, and current benchmarks miss that, at least as far as I'm aware. I would think some arena-type benchmarks where humans rate responses would measure this, though I'm not sure which those are.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#117Anthropic is a bit nuts, I had $260 of credits on my max account for the extra usage the other night. It was expiring, so I figured I'll fire up an agentic swarm to deep dive and make some deep changes to some old cold bases.. literally 25 minutes or less, $260 burnt, it didn't get get into the implementation, just wrote a ton of useless plans for the most part. It really opened my eyes to what they expect to charge…
Maybe they consider that hiring a person to do it would have cost at least as much and taken much more time, so paying them is a bargain.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#118Earlier quoted context omitted.
I have Fable plan and Opus implement. I haven't had any major issues working this way; however, Opus does seem plain fucking stupid compared to what I experienced with Sonnet previously.
> however, Opus does seem plain fucking stupid Infuriatingly so, in a way I don't remember Opus 4.8 being, but maybe I've just been ruined by Fable 5.
I got so used to it, when they finally pulled access for me and I had to go back to Opus I felt like I was working with my hands tied.
I finally know what those women with AI boyfriends felt like when their app updated and it won't dirty talk with them anymore.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#119Any benchmark showing Opus 5 as the best just loses credibility for me. Anyone who's actually used Opus 5 daily knows what I'm talking about.
Re: Qwen3.8 Max now ranked as the best overall model by agentic index
#120Earlier quoted context omitted.
And at that cost they're still not profitable. It's going to be a bumpy road ahead...
People keep saying this but from what we’ve seen, Anthropic models are marginally profitable and earn back their costs over their lifetime. The company is burning money building the next versions and other ventures (e.g. verticals), but the models themselves have been profitable.