On top of that, saving 90% of input tokens != saving 90% "of tokens", output is wildly more expensive.
[1] especially if it's a really old model like Gemini 2.5!
101–110 of 156 posts
On top of that, saving 90% of input tokens != saving 90% "of tokens", output is wildly more expensive.
[1] especially if it's a really old model like Gemini 2.5!
It takes true corporate dedication to publish technical thought leadership on a page that actively fights your ability to read it.
not X, but Y... so this is very likely Opus 5 text, given timing and how it reads. has all the marks of it, one being very weird wording which makes it so hard to read the text.
What great productivity gains are Spotify achieving in making their product worse?
Eg. On opencode there's explorer subgagent that we can set to use lower level model. Many people even use haiku level model for this.
It takes true corporate dedication to publish technical thought leadership on a page that actively fights your ability to read it.
"The seat license isn't what hurts, it's the tokens." not X, but Y... so this is very likely Opus 5 text, given timing and how it reads. has all the marks of it, one being very weird wording which makes it so hard to read the text.
> modes are the load-bearing piece:
Okay, this write-up is filled with Claudisms.
What great productivity gains are Spotify achieving in making their product worse?
The spotify desktop app is one of the worst pieces of software by a major company I have ever used.
15 seconds.
Incredible stuff.