Live data from Hacker News

Gemini 3.8 Flash and 3.8 Flash Cyber

blog.google

141–150 of 699 posts

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#142
post #85

Earlier quoted context omitted.

The Gemini models have openly trained for SVG output, apparently with a specialism on animals in forms of transport! https://twitter.com/JeffDean/status/2024525132266688757

I don’t know if you’re joking, but I don’t see anything in the linked tweet which suggests that is the case

Watch the video. It's from then-Gemini-lead Jeff Dean and the video shows off an animated pelican riding a bicycle, a frog on a penny-farthing, a giraffe driving a tiny car, an ostrich on roller skates, a turtle kickflipping a skateboard, and a dachshund driving a stretch limousine.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#143
post #113

I've been using Gemini 3.7 for my personal trip planning app. Across multiple benchmarks, it ranks higher on everything I tried: - Real world knowledge (when a thing opens and closes, the geographic region, historical facts). It's also the best at taking a cluster of places and working out a visiting order. - Photo ranking (which photo should be the hero). Gemini can tell whether a photo is of the thing or of the vie…

Can G3.7 use Google Maps for distance grounding?

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#144

Earlier quoted context omitted.

I found 3.5 pro to be much better than 3.5 flash, but 3.7 flash with high reasoning is comparable and way way faster.

There are no public release of 3.5 pro. Either its a typo, or you have some insider information

Check my profile?

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#145

Is the Gemini CLI still terrible compared to Claude Code and Codex? The harness the main thing holding back Google models as they could've been the best given all the advantages in compute capacity and training data they initially had, where now even the Google CEO said they're falling behind in agentic tasks, which is sort of a vicious cycle because RLHF relies on human usage.

it is antigravity now. It is ok

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#146
They've interestingly left out any mention of speed.

I have been testing 3.7 flash against 3.5 flash and it seems to lose every time in overall latency. Every benchmark I've seen seems to suggest the opposite[1] - that 3.7 flash is significantly (at times 2x) faster than 3.5 flash - but I have never been able to prove this out in real world use cases.

Has anyone found their latency numbers to actually be accurate? Is this why they've toned it down in this release? For context, I'm testing larger generation payloads that take 8-10 seconds in 3.5 flash and 15-25 seconds in 3.7 flash. Lowest reasoning settings in both cases.

1: https://artificialanalysis.ai/?speed=intelligence-vs-speed&m...

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#147
>"safety performance" - this starting to get long in the tooth. Gemini cut programming session 3 times for "safety reasons" yesterday for mentioning image generation (I need to generate bunch of those for infinite zoom virtual training app experience). After I got creative and managed to trick it to answer t was of course because "think of a children"

And in my other app I was debugging and using OpenAI to optimize some path it cut me off numerous times because it did not like JIT functionality (this is my commercial business rule evaluation engine that compiles rules to executable code inside the app to increase performance using asmjit library)

I am basically paying for them to waste my tokens and time on these 2 tasks

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#148
post #130
post #122

Earlier quoted context omitted.

I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.

you didn’t even read your comment before you posted it?

Eh that one is on me, if I think too much about my HN comment I end up deleting before posting it. I rely on the 1 min `delay` set in the profile page to fix before it goes live, but for some reason this time it was set to 0.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#149

Earlier quoted context omitted.

I found 3.5 pro to be much better than 3.5 flash, but 3.7 flash with high reasoning is comparable and way way faster.

There are no public release of 3.5 pro. Either its a typo, or you have some insider information

Googlers and some external workplaces have had 3.5 pro access for a few months now.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#150
post #64

Pelicans (thinking effort high, medium, low): https://tools.simonwillison.net/markdown-svg-renderer?url=ht... - high cost 8.9742 cents Here are the 3.7 pelicans for comparison: https://tools.simonwillison.net/markdown-svg-renderer.html?u... - high cost 8.4387 cents (I think thinking level low is a regression on 3.8 compared to 3.7.)

The fenders are a nice touch, but putting the fenders through the tires seems like a design flaw.
Post reply on HN