Live data from Hacker News

Gemini 3.8 Flash and 3.8 Flash Cyber

blog.google

121–130 of 699 posts

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#122
post #115
post #113

I've been using Gemini 3.7 for my personal trip planning app. Across multiple benchmarks, it ranks higher on everything I tried: - Real world knowledge (when a thing opens and closes, the geographic region, historical facts). It's also the best at taking a cluster of places and working out a visiting order. - Photo ranking (which photo should be the hero). Gemini can tell whether a photo is of the thing or of the vie…

"Claude 3.7"?

I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#123
post #85

Earlier quoted context omitted.

This is in comparison to Fable: > https://tools.simonwillison.net/markdown-svg-renderer?url=ht... > Took just under 14 minutes to generate, and at 65927 output tokens cost me a hefty $3.30! So 50x cheaper - and how much faster?

The Gemini models have openly trained for SVG output, apparently with a specialism on animals in forms of transport! https://twitter.com/JeffDean/status/2024525132266688757

I don’t know if you’re joking, but I don’t see anything in the linked tweet which suggests that is the case

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#126
post #122
post #115

Earlier quoted context omitted.

"Claude 3.7"?

I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.

incredible. further evidence supporting my personal stance to never ever let an LLM write or edit my writing intended for another human being to read. this is all me, baby

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#128

After struggling with Gemini for months, I think the trick to getting the most out of the model is writing a really solid personal intelligence/instructions prompt. The results are night and day in terms of performance.

slot machine addict thinks if he pushes buttons in a certain order the odds get better.

In all seriousness, gemini has the best interactive planning document/orchestration. Tell it to create a plan document and work through it with it and it will preform really well(in antigravity products). But this is the case with plan modes with every model, I just think the interactive document that antigravity uses is really well thought out.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#129

Currently top at https://deepswe.datacurve.ai - beating Opus 5! https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium! Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.

As of writing this comment, Claude Opus 5 has an intelligence score of 63, not 59 (it's not the same as Gemini 3.8 Flash). With a score of 59, Gemini 3.8 Flash is in eighth place, falling behind even Grok 4.6, Kimi k3, and GLM 5.3. https://imgur.com/a/BMOJBED

They are all much larger and more expensive models. Google does not have a frontier model right now, but for cheap ones, they are better than event the chinese models now.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#130
post #122
post #115

Earlier quoted context omitted.

"Claude 3.7"?

I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.

you didn’t even read your comment before you posted it?
Post reply on HN