Don’t let the “flash” name fool you, this is an amazing model. I have been playing with it for the past few weeks, it’s genuinely my new favorite; it’s so fast and it has such a vast world knowledge that it’s more performant than Claude Opus 4.5 or GPT 5.2 extra high, for a fraction (basically order of magnitude less!!) of the inference time and price
Gemini 3 Flash: Frontier intelligence built for speed
131–140 of 609 posts
Re: Gemini 3 Flash: Frontier intelligence built for speed
#132At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#133At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#134I really wish these models were available via AWS or Azure. I understand strategically that this might not make sense for Google, but at a non-software-focused F500 company it would sure make it a lot easier to use Gemini.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#135Don’t let the “flash” name fool you, this is an amazing model. I have been playing with it for the past few weeks, it’s genuinely my new favorite; it’s so fast and it has such a vast world knowledge that it’s more performant than Claude Opus 4.5 or GPT 5.2 extra high, for a fraction (basically order of magnitude less!!) of the inference time and price
Just to point this out: many of these frontier models cost isn't that far away from two orders of magnitude more than what DeepSeek charges. It doesn't compare the same, no, but with coaxing I find it to be a pretty capable competent coding model & capable of answering a lot of general queries pretty satisfactorily (but if it's a short session, why economize?). $0.28/m in, $0.42/m out. Opus 4.5 is $5/$25 (17x/60x). I…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#136At this point in time I start to believe OAI is very much behind on the models race and it can't be reversed Image model they have released is much worse than nano banana pro, ghibli moment did not happen Their GPT 5.2 is obviously overfit on benchmarks as a consensus of many developers and friends I know. So Opus 4.5 is staying on top when it comes to coding The weight of the ads money from google and general direct…
This has been true for at least 4 months and yeah, based on how these things scale and also Google's capital + in-house hardware advantages, it's probably insurmountable.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#137Also, I hate that I cannot send the Google models in a "Thinking" mode like in ChatGPT. When I send GPT 5.1 Thinking on a legal task and tell it to check and cite all sources, it takes +10 minutes to answer, but it did check everything and cite all its sources in the text; whereas the Gemini models, even 3 Pro, always answer after a few seconds and never cite their sources, making it impossible to click to check the answer. It makes the whole model unusable for these tasks. (I have the $20 subscription for both)
Re: Gemini 3 Flash: Frontier intelligence built for speed
#138Re: Gemini 3 Flash: Frontier intelligence built for speed
#139This is awesome. No preview release either, which is great to production. They are pushing the prices higher with each release though: API pricing is up to $0.5/M for input and $3/M for output For comparison: Gemini 3.0 Flash: $0.50/M for input and $3.00/M for output Gemini 2.5 Flash: $0.30/M for input and $2.50/M for output Gemini 2.0 Flash: $0.15/M for input and $0.60/M for output Gemini 1.5 Flash: $0.075/M for inp…
Re: Gemini 3 Flash: Frontier intelligence built for speed
#140Does anyone else understand what the difference is between Gemini 3 'Thinking' and 'Pro'? Thinking "Solves complex problems" and Pro "Thinks longer for advanced math & code". I assume that these are just different reasoning levels for Gemini 3, but I can't even find mention of there being 2 versions anywhere, and the API doesn't even mention the Thinking-Pro dichotomy.
I think: Fast = Gemini 3 Flash without thinking (or very low thinking budget) Thinking = Gemini 3 flash with high thinking budget Pro = Gemini 3 Pro with thinking
>Fast = 3 Flash
>Thinking = 3 Flash (with thinking)
>Pro = 3 Pro (with thinking)