Live data from Hacker News

Gemini 3.8 Flash and 3.8 Flash Cyber

blog.google

211–220 of 697 posts

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#211
post #200

The speed combined with the fact that this thing is really good at HTML JavaScript is pretty exciting. Here's what I got for 1.8 cents and 13 seconds from the prompt "make me a cool thing in html": https://gisthost.github.io/?6a77bc41a81718c6aaa10d4ab243c59f Transcript here (it was part of a chat): https://gist.github.com/simonw/b6149a49d327164d67d62c3d12992...

Mission accomplished. That's both cool and fast.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#212

I don't use Gemini, but I thought `cool, let's give this new model a try`. Opened gemini.google.com, and I'm not even surprised. The drop down gives me the following options: - Flash-Lite - 3.6 Flash [new] - 3.1 Pro The above is why i don't use LLM products from Google. If the model is not available right this minute (heck, hours before the release!), then I'm not gonna bother getting back to it tomorrow, because tom…

[dead]

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#213

been absolutely loving 3.7 flash for coding. it feels very fast and quality is decent for implementing product features. usually use opus or sol for hardcore debugging.

Is there any good subscription and CLI harness to use Gemini models now?

I tested Gemini CLI while ago, and it was awful tbh.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#214

Who coined the phrase "cyber" for security related things lol. It's so 1999.

Cyber is more of an early 1990's thing, and I have no issue with it unlike most in the tech field. I feel like it dropped off in the late 90's and early 00's but made a comeback as hacking became a mainstream security issue.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#215

3.7 flash was by far the best model for image recognition tasks according to my benchmarks. 3.8 flash didn't regress any candidates and improved some specificity (positive ID of common name vs species name of exotic fruit, correct identification of cast/replica of artifact and statue) but is still relatively weaker (26/30) on esoteric public figures (Korean beatboxers). I'm going to have to make my benchmark harder.

I’m very curious about your esoteric public figures benchmark, do you ask it in English or Korean to identify the person? Does it change the result? I wonder if having data labeled in only a given language (or web sources in only a given language) change the output.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#216

I don't use Gemini, but I thought `cool, let's give this new model a try`. Opened gemini.google.com, and I'm not even surprised. The drop down gives me the following options: - Flash-Lite - 3.6 Flash [new] - 3.1 Pro The above is why i don't use LLM products from Google. If the model is not available right this minute (heck, hours before the release!), then I'm not gonna bother getting back to it tomorrow, because tom…

it's weird how the web ui doesn't show the latest flash options while the desktop/mobile apps update the same day as the release. I saw the model in the model selection (by coincidence) before seeing it show up on HN

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#217
post #179

I don't use Gemini, but I thought `cool, let's give this new model a try`. Opened gemini.google.com, and I'm not even surprised. The drop down gives me the following options: - Flash-Lite - 3.6 Flash [new] - 3.1 Pro The above is why i don't use LLM products from Google. If the model is not available right this minute (heck, hours before the release!), then I'm not gonna bother getting back to it tomorrow, because tom…

I'm a paid Gemini subscriber via Workspace Standard accounts and yet I also only have access to 3.6. So frustrating and confusing. Meanwhile Anthropic and OpenAI simply release a model everywhere (Fable on Pro only as a somewhat mild exception).

> I'm a paid Gemini subscriber via Workspace Standard accounts and yet I also only have access to 3.6.

Same and I have found it extremely annoying. I actually really like the Gemini models for question/answer stuff and reach for it before Claude (the other model family I have purchased) but it's getting long in the tooth at this point and I'm finding my Gemini usage shrinking to nearly 0.

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#218
I don't know if Google is having the worst marketing fumble or the most genius marketing one. Their "flash" models are very comparable to other companies' "pro" or "flagship" models. It seems to be a quite counterintuitive naming convention as it undersells the models.

Unless they have an even more powerful Gemini Pro in the oven...?

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#219
post #158

Earlier quoted context omitted.

Google has been doing staged roll outs on all their products since forever.

What stage of the roll out are we where I don’t even see 3.7-flash which was released 2-3 weeks ago?

https://aistudio.google.com/prompts/new_chat?model=gemini-3....

Re: Gemini 3.8 Flash and 3.8 Flash Cyber

#220
post #162

Earlier quoted context omitted.

They are all much larger and more expensive models. Google does not have a frontier model right now, but for cheap ones, they are better than event the chinese models now.

That's not being debated here. The initial reported numbers were false and this was simply pointed out. You're changing the subject.

Opus 5 medium has the same score as 3.8 flash on artificial analysis intelligence index.

Are you implying Google or Artificial Analysis are reporting false numbers? What's your source?

Post reply on HN