Earlier quoted context omitted.
I'd be more OK with it if prices weren't subsidized so much and people actually had to pay to ask opus how to pee, then maybe we would realize we don't need beefier models for everything
There isn't much evidence at all that inference is "subsidised" (and by whom?) Training is quite expensive and it does look likely that the American providers have been doing that at a loss. In any case, you can go buy a MacBook Pro M5 48GB or an AMD R9700 and run Qwen 3.6 35B-A3B (a very capable model) and the only "subsidy" is you plugging it in, and 140W is not exactly a huge amount of power (roughly 50¢ per day i…
Well... why else would the major providers now tighten the screws on per-token pricing?