Gemini 3.8 Flash and 3.8 Flash Cyber
121–130 of 697 posts
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#122I've been using Gemini 3.7 for my personal trip planning app. Across multiple benchmarks, it ranks higher on everything I tried: - Real world knowledge (when a thing opens and closes, the geographic region, historical facts). It's also the best at taking a cluster of places and working out a visiting order. - Photo ranking (which photo should be the hero). Gemini can tell whether a photo is of the thing or of the vie…
"Claude 3.7"?
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#123Earlier quoted context omitted.
This is in comparison to Fable: > https://tools.simonwillison.net/markdown-svg-renderer?url=ht... > Took just under 14 minutes to generate, and at 65927 output tokens cost me a hefty $3.30! So 50x cheaper - and how much faster?
The Gemini models have openly trained for SVG output, apparently with a specialism on animals in forms of transport! https://twitter.com/JeffDean/status/2024525132266688757
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#124Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#125Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#126Earlier quoted context omitted.
"Claude 3.7"?
I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#127Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#128After struggling with Gemini for months, I think the trick to getting the most out of the model is writing a really solid personal intelligence/instructions prompt. The results are night and day in terms of performance.
In all seriousness, gemini has the best interactive planning document/orchestration. Tell it to create a plan document and work through it with it and it will preform really well(in antigravity products). But this is the case with plan modes with every model, I just think the interactive document that antigravity uses is really well thought out.
Re: Gemini 3.8 Flash and 3.8 Flash Cyber
#129Currently top at https://deepswe.datacurve.ai - beating Opus 5! https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium! Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.
As of writing this comment, Claude Opus 5 has an intelligence score of 63, not 59 (it's not the same as Gemini 3.8 Flash). With a score of 59, Gemini 3.8 Flash is in eighth place, falling behind even Grok 4.6, Kimi k3, and GLM 5.3. https://imgur.com/a/BMOJBED