Engineers at google have publically stated that the models are too big and are far from their potencial. Glad they're being proven right with every release. They continue to focus on smaller models while openai and anthropic are increasing compute requirements for their SOTA models.
> Engineers at google have publically stated that the models are too big and are far from their potencial Can you link to a source?
Gemini 3.5 Flash
391–400 of 692 posts
Re: Gemini 3.5 Flash
#392Earlier quoted context omitted.
They were CPU killers but man those Flash websites were gorgeous (talking mostly about MU Online "private" servers)
It was probably the right call at the time with low bandwidth. Nowadays I bet flash would execute faster than most js heavy sites :D
Re: Gemini 3.5 Flash
#393The pelican is a lot : https://github.com/simonw/llm-gemini/issues/133#issuecomment... Not a great bicycle though, it forgot the bar between the pedals and the back wheel and weirdly tangled the other bars. Expensive too - that pelican cost 13 cents: https://www.llm-prices.com/#it=11&ot=14403&sel=gemini-3.5-fl...
Love your pelicans, as always. And that one is... Wow. I noticed the "Synthwave" aesthetic, which is enjoying quite some success since quite some time now, has found its way into AI models (even when it's not in the user's query). It's not the first time I see the sun at sunset with color bands etc. in AI-generated pictures. Don't know why it's now taking on in AI too. https://en.wikipedia.org/wiki/Synthwave Hence th…
So it's as relevant and baked-in to today as actual 80s synth-culture was in 2000.
Re: Gemini 3.5 Flash
#394The pelican is a lot : https://github.com/simonw/llm-gemini/issues/133#issuecomment... Not a great bicycle though, it forgot the bar between the pedals and the back wheel and weirdly tangled the other bars. Expensive too - that pelican cost 13 cents: https://www.llm-prices.com/#it=11&ot=14403&sel=gemini-3.5-fl...
That pelican looks like it's in Miami for a crypto conference.
Re: Gemini 3.5 Flash
#395Google shot it's shot with that alternative history artwork generation fiasco. Don't know why anyone would be too hot for them now. Dime a dozen at this point.
Re: Gemini 3.5 Flash
#396Earlier quoted context omitted.
This is not priced at inference cost. My guess: it's the price at which they make more money than if they rent the TPUs to other companies. The Gemini team has had trouble securing enough TPUs for their user's needs. They struggle with load and their rate limits are really bad. Maybe at a higher price, they have a better chance at getting more TPUs assigned?
The cost at such they could rent out the TPUs, i.e. the market rate, is the inference cost. Just because you are vertically integrated doesn't mean you get to discount the one business units products to the other. Doing so discounts the opportunity cost you pay and is just bad accounting.
Re: Gemini 3.5 Flash
#397Gemini 3.5 Flash's 2000 token clocks aren't bad. https://clocks.brianmoore.com/
Re: Gemini 3.5 Flash
#398Wow at the price hike. Still I think in the long run the Chinese will win if they're able to produce hardware comparable to Nvidia.
Re: Gemini 3.5 Flash
#399Re: Gemini 3.5 Flash
#400I have a tool to track these I've built Relatively speaking here's where it's at: score age size name 44.2 97 large GLM-5 (Reasoning) 44.7 187 - GPT-5.1 (high) 44.9 29 - Qwen3.6 Max Preview 45 0 - Gemini 3.5 Flash 45.5 27 large MiMo-V2.5-Pro 45.6 75 - GPT-5.4 (low) this is from artificial-analysis using https://github.com/day50-dev/aa-eval-email/blob/main/art-ana... I really don't know why people down vote me. What d…
We genuinely don't understand what your post is about. What is this tool? What are these numbers representative? Why are things sorted in that order?
You haven't communicated really anything at all. I am interested, I'd like to understand. Write a more complete post, please.