Is it on topic to complain about the various claude-isms in this article? I don't know any actual humans that write titles like "Two floors the rate card hides". I find my brain disengages once I suspect something of being written by an LLM. If the author didn't put much effort into writing it, should I expect them to have put much effort into fact-checking it? Edit: this specific title has been deleted from the arti…
The real prices of frontier models
41–50 of 91 posts
Re: The real prices of frontier models
#42My take from this is that Anthropic is screwing us again. I hope AMD shows them up again.
Could you elaborate?
Re: The real prices of frontier models
#43The real elephant in the room is pricing for KV cache writes and reads. That makes all the difference for tasks with large context.
Re: The real prices of frontier models
#44An individual token, and the level of energy it represents (electricity, or relative effectiveness per model) increasingly seems the space of obsfucation. This space can be increasingly avoided by becoming, and remaining, efficient and effective with prompts.
That is one of the interesting things about Neuralwatt cloud. Their pricing is based on energy rather than tokens (actually they have a token-based alternative, but claim the energy pricing results in 95% cheaper results). I've tried out their subscription offer and it does seem like you get a lot more usage even on the cheap plan. However since the energy metering is pretty much unique to them (at least that I've se…
How we drive AI will cause mileage variances, but over time the improved practices can measurably change.
Re: The real prices of frontier models
#45Re: The real prices of frontier models
#46Is it on topic to complain about the various claude-isms in this article? I don't know any actual humans that write titles like "Two floors the rate card hides". I find my brain disengages once I suspect something of being written by an LLM. If the author didn't put much effort into writing it, should I expect them to have put much effort into fact-checking it? Edit: this specific title has been deleted from the arti…
"Authoritative: it is the same count Anthropic bills against."
"This reframes a headline that looked like good news."
Re: The real prices of frontier models
#47Re: The real prices of frontier models
#48Providers change tokenizers all the time with model updates, and it's often not even possible to query/figure out how text is tokenized without actually just sending the LLM a request.
Just switch to charging for bytes of intelligence. Please. Claude Shannon figured this out decades ago.
Re: The real prices of frontier models
#49Earlier quoted context omitted.
If I need to put 2 and 2 together for you: https://en.wikipedia.org/wiki/Shrinkflation
Would you mind sharing your actual thoughts instead of this vague posting? What does AMD have to do with Anthropic?
Re: The real prices of frontier models
#50> You will see people claim Claude uses 2x to 4x the tokens of GPT. Our measurements do not support that, and overstating it would undercut the real point.
It's not because a single prompt represents only 1.7x the number of tokens that a model doesn't use 4x as many tokens as another, when running as an agent. This doesn't take at all the number of tokens of the output into account, and the number of tokens of the potential tool calls from this output, which directly feeds back to input tokens.
The article also has a very small test set (16 documents), all of very small length (15K tokens at most, when models go up to 1M in context and agents routinely exceed this and have to summarize).
Complete garbage article.