Live data from Hacker News

Gemini 3.8 Flash and 3.8 Flash Cyber

blog.google

331–340 of 699 posts

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#331
post #200

The speed combined with the fact that this thing is really good at HTML JavaScript is pretty exciting. Here's what I got for 1.8 cents and 13 seconds from the prompt "make me a cool thing in html": https://gisthost.github.io/?6a77bc41a81718c6aaa10d4ab243c59f Transcript here (it was part of a chat): https://gist.github.com/simonw/b6149a49d327164d67d62c3d12992...

Definitely cool. I noticed it felt a little janky on my PC despite being "60 FPS"...then I noticed the "60 FPS" is hard-coded into the HTML.

The new Bench-Maxxing!

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#332

"available to trusted defenders through our new Fairwind Program" Then why even bother announcing this? Ordinary people can use K3 and GLM 5.3 or whatever drops next and avoid all this hassle.

You're telling me for only 5x the cost and 1/10th the speed I can use a Chinese model which performs worse than Gemini 3.8 Cyber? And I get to do all the hosting and setup work myself instead of just using a model and framework which is already integrated with GCP? Dang!

I'm sorry, is this a bot that is optimized for sealioning? The point is that you don't have access to Cyber.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#333
I like Google's strategy here. These new Flash models of late (Flash 3.6, 3.7 and now 3.8) have obviously been distilled from a much larger unreleased model (Gemini 3.5 Pro, iirc from the rumors).

One aspect of model releases that don't get discussed as much are the cache invalidation (changes in underlying architecture, weights, or tokenizers); I assess Google seems to be squeezing the maximum out of the last 'Pro' version they released with 3.1 back in February.

Small models cataching up with their bigger siblings are fantastic news.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#334
post #97

The most interesting thing about the Gemini models is still their multi-modal support: they accept audio and video input, OpenAI and Anthropic's flagships are still image-only. Gemini Flash is also pretty cheap, so it's a great family for performing media analysis, like extracting structured data from images and video.

Interesting side note: although Opus is still image-only, you can still drag videos into Claude Code and it doesn't blink an eye; it just strips it down to a series of images to parse.

True multimodal support would be way better, but I have no issues pasting in full screen recordings while QA'ing games and having Claude identify and fix issues in the video.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#335
post #167

Earlier quoted context omitted.

It depends on how you're querying Gemini models. OpenRouter is the fastest by far. I'm guessing they bought the dedicated pipe from Google. Gemini via VertexAI and consumer API has pretty bad latency.

Yea I am testing through OpenRouter - have you noticed 3.7 flash being significantly faster?

I guess it might be relative, but switching from VertexAI endpoint to OpenRouter was like 2-3x faster for us.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#336

Earlier quoted context omitted.

Curious where did you hear this rumor?

All the talk on Reddit on Gemini 3.8 discussions: https://www.reddit.com/search/?q=gemini+3.9&cId=1650e403-bcf...

Accordit to reddit talk, Fable 5.1 is worse than Opus 4.6 and 8B models are smarter than Qwen 3.8 Max, I wouldn't take anything said there with any more reliability than an instagram short.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#337
post #175

Earlier quoted context omitted.

I wasn't trying to be precise originally, I just tried to fit activities into "morning / evening" buckets. I did the whole itinerary with Opus first, but when I gave it to Gemini 3.7 Flash to review, it started correcting it with "this place will close 5PM" or "this place is closed for good". It was right on every nit, so it was surprising how well the model knows these things. If I ever release this I'll probably ne…

When you called the Gemini API, did you opt in to using search grounding: tools=[{"type": "google_search"}] I'm curious whether in fact you were getting answers from the model weights (which is what I had assumed) or whether your API calls were resulting in web search tool calls.

Google AI person here:

Using grounding in Gemini is indeed backed by the same canonical data source for business information (like opening hours) as Google Maps. This stuff is available in its own API for a GCP fee, but we’ve built tooling to connect it to the Gemini agentic ecosystem as well.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#338
post #200

The speed combined with the fact that this thing is really good at HTML JavaScript is pretty exciting. Here's what I got for 1.8 cents and 13 seconds from the prompt "make me a cool thing in html": https://gisthost.github.io/?6a77bc41a81718c6aaa10d4ab243c59f Transcript here (it was part of a chat): https://gist.github.com/simonw/b6149a49d327164d67d62c3d12992...

OK, but what about a pelican in a bycicle.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#340

Earlier quoted context omitted.

> basically Google with a much better frontend and no ads/seo nonsense so far

Fair. But I think there's a good chance it stays that way on paid plans. YouTube Premium is still ad free. Also them having their own silicon means they don't have to pay the Nvidia tax and can keep costs a lot lower.

This thinking is why I am all in on GOOG shares. As a bonus, that means I’m getting part of Anthropic’s gains as well!
Post reply on HN