I've only ever seen linear increases. When did Moore's law even _start_?
The End of Moore's Law for AI? Gemini Flash Offers a Warning
61–70 of 78 posts
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#62It can be just Google trying to capitalize Gemini's increasing popularity. Until 2.5 Gemini was a total underdog. Less so since 2.5.
Since Gemini CLI was recently released, many people on the "free" tier noticed that their sessions immediately devolved from Gemini 2.5 Pro to Flash "due to high utilization". I asked Gemini itself about this and it reported that the finite GPU/TPU resources in Google's cloud infrastructure can get oversubscribed for Pro usage. Google (no secret here) has a subscription option for higher-tier customers to request guaranteed provisioning for the Pro model. Once their capacity gets approached, they must throttle down the lower-tier (including free) sessions to the less resource-intensive models.
Price is one lever to move once capacity becomes constrained. Yet, as the top voted comment of this post explains, it's not honest to simply label this as a price increase. They raised Flash pricing on input tokens but lowered pricing on output tokens up to certain limits -- which gives creedence to the theory that they are trying to shape the demand in order for it to better match their capacity.
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#63I think the big thing that really surprised me. Llama 4 maverick is 16x 17b. So 67GB of size. The equivalency is 400billion. Llama 4 behemoth is 128x 17b. 245gb size. The equivalency is 2 trillion. I dont have the resources to be able to test these unfortunately; but they are claiming behemoth is superior to the best SAAS options via internal benchmarking. Comparatively Deepseek r1 671B is 404gb in size; with pretty…
"Sir, I'm delighted to report that the productivity and insights gained outclass anything available from four years ago. We are clearly winning."
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#64It could just as well have been Google reducing subsidisation. From the outside that would look exactly the same
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#65Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#66Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#67"In a move that at first went unnoticed, Google significantly increased the price of its popular Gemini 2.5 Flash model" It's not quite that simple. Gemini 2.5 Flash previously had two prices, depending on if you enabled "thinking" mode or not. The new 2.5 Flash has just a single price, which is a lot more if you were using the non-thinking mode and may be slightly less for thinking mode. Another way to think about t…
I really hate the thinking. I do my best to disable it but don't always remember. So often it just gets into a loop second guessing itself until it hits the token limit. It's rare it figures anything out while it's thinking too but maybe that's because I'm better at writing prompts.
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#68Personally, I'm rooting for RWKV / Mamba2 to pull through, somehow. There's been some work done to increase their reasoning depths, but transformers still beat them without much effort.
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#69Can anyone explain the economics of Anthropic's Max plan pricing to me? I have friends on the $100/month plan using well over $800 of tokens per month with Claude Code (according to ccusage). I certainly don't use Claude Code as much if I'm not on a flat rate plan, the cost spirals out of control very quickly. I understand that a subscription makes for more predictable revenue and that there will be people on the Max…
Re: The End of Moore's Law for AI? Gemini Flash Offers a Warning
#70"In a move that at first went unnoticed, Google significantly increased the price of its popular Gemini 2.5 Flash model" It's not quite that simple. Gemini 2.5 Flash previously had two prices, depending on if you enabled "thinking" mode or not. The new 2.5 Flash has just a single price, which is a lot more if you were using the non-thinking mode and may be slightly less for thinking mode. Another way to think about t…
I really hate the thinking. I do my best to disable it but don't always remember. So often it just gets into a loop second guessing itself until it hits the token limit. It's rare it figures anything out while it's thinking too but maybe that's because I'm better at writing prompts.