Live data from Hacker News

Gemini 3.5 Flash

blog.google

11–20 of 692 posts

Re: Gemini 3.5 Flash

#14
post #8

$1.5/m input tokens $9/m output tokens 6x the price of 3.1 flash lite

I don't think input/output pricing matters, 90% of the cost is cache. $0.15 is pretty good, but still very expensive.

10% of input pricing is standard especially compared to competition.

Re: Gemini 3.5 Flash

#16

Engineers at google have publically stated that the models are too big and are far from their potencial. Glad they're being proven right with every release. They continue to focus on smaller models while openai and anthropic are increasing compute requirements for their SOTA models.

Given the cost increase associated with this model, and previous model releases, I think the size is trending upwards, not down.

Re: Gemini 3.5 Flash

#19

Engineers at google have publically stated that the models are too big and are far from their potencial. Glad they're being proven right with every release. They continue to focus on smaller models while openai and anthropic are increasing compute requirements for their SOTA models.

Don’t let that fool yourself. Google will have SOTA models as big as or even bigger than their competitors.

They are just refining their current models while they finish training the next generation.

They will all come out at about the same time. Anthropic, OpenAi, Google, xAI

Post reply on HN