Earlier quoted context omitted.
Came and went in a flash
Because they are preparing Gemini 3.9 Flash
Gemini 3.8 Flash and 3.8 Flash Cyber
91–100 of 697 posts
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#92Currently top at https://deepswe.datacurve.ai - beating Opus 5! https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium! Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#93One place where I find the Flash models surprisingly bad is Google Search's "AI Mode". A recent example - I searched for how to unsubscribe from Pearson emails. Google Search "AI Mode" confidently gave me a sequence of steps along the lines of Settings > Profile > Email preferences > Unsubscribe. Of course, I looked for an unsubscribe link before asking Google. None of those options existed. The correct answer was th…
I think that's just a limitation on the size of the model. I'm pretty sure that they use a pretty small model in those summaries to save money, which naturally makes them a little less smart.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#94Earlier quoted context omitted.
The benchmark also doesn't include speed. You almost think something has gone wrong when using it because it returns full responses so incredibly fast.
Not just speed, also reliability. IME, Gemini's speed and quality doesn't degrade badly during weekday working hours compared to OAI, and especially Anthropic.
Not sure on consumer/product use though
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#95Currently top at https://deepswe.datacurve.ai - beating Opus 5! https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium! Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.
We'll see about that. I suspect benchmaxxing as all the labs do as I haven't found Gemini models to be nearly as good in agentic engineering compared to Claude or GPT models.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#96Currently top at https://deepswe.datacurve.ai - beating Opus 5! https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium! Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.
A fifth of the cost of Opus 5! Google is certainly pushing the completion with this.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#97Gemini Flash is also pretty cheap, so it's a great family for performing media analysis, like extracting structured data from images and video.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#98One place where I find the Flash models surprisingly bad is Google Search's "AI Mode". A recent example - I searched for how to unsubscribe from Pearson emails. Google Search "AI Mode" confidently gave me a sequence of steps along the lines of Settings > Profile > Email preferences > Unsubscribe. Of course, I looked for an unsubscribe link before asking Google. None of those options existed. The correct answer was th…
https://www.pearson.com/privacy-center/privacy-notices/full-... >We will not send marketing emails to a user who has opted out of receiving them. Any marketing communications we send will include an unsubscribe link at the end of the email. I don't think this is AI's fault. This is Pearson's publishing incorrect information and the only way to really know they are a bunch of lying assholes is to have an account and t…
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#99Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#100Earlier quoted context omitted.
Latest rumor is that 3.5 pro was struggling to be meaningfully better than flash, since iterations on flash were moving much faster than iterations on pro, likely due to model size (flash is estimated to be in the 200-400B range).
I found 3.5 pro to be much better than 3.5 flash, but 3.7 flash with high reasoning is comparable and way way faster.