Earlier quoted context omitted.
I believethere’s also exponential dislike growing for Altman among most AI users, and that impacts how the brand/company is perceived.
Most AI users outside of HN does not have any idea of who Altman is. ChatGPT is in many circles synonymous to AI so their brand recognition is huge.
Gemini 3 Flash: Frontier intelligence built for speed
591–600 of 609 posts
Re: Gemini 3 Flash: Frontier intelligence built for speed
#592Earlier quoted context omitted.
I feel like that is part of their cloud strategy. If your company wants to pump a huge amount of data through one of these you will pay a premium in network costs. Their sales people will use that as a lever for why you should migrate some or all of your fleet to their cloud.
A few gigabytes of text is practically free to transfer even over the most exorbitant egress fee networks, but would cost “get finance approval” amounts of money to process even through a cheaper model.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#593Earlier quoted context omitted.
Why wouldn't you switch? The cost to switch is near zero for me. Some tools have built in model selectors. Direct CLI/IDE plug-ins practically the same UI.
Not OP, but I feel the same way. Cost is just one of the factor. I'm used to Claude Code UX, my CLAUDE.md works well with my workflow too. Unless there's any significant improvement, changing to new models every few months is going to hurt me more.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#594Earlier quoted context omitted.
Also https://artificialanalysis.ai/evaluations/omniscience Prepare to be amazed
I'm confused about the "Accuracy vs Cost" section. Why is Gemini 3 Pro so cheap? It's basically the cheapest model in the graph (sans Llama 4 and Mistral Large 3) by a wide margin, even compared to Gemini 3 Flash. Is that an error?
They have a similar chart that compares results across all their benchmarks vs. cost and 3 Flash is about half as expensive as 3 Pro there despite being four times cheaper per token.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#595Any word on if this using their diffusion architecture?
Re: Gemini 3 Flash: Frontier intelligence built for speed
#596Re: Gemini 3 Flash: Frontier intelligence built for speed
#597Earlier quoted context omitted.
A few gigabytes of text is practically free to transfer even over the most exorbitant egress fee networks, but would cost “get finance approval” amounts of money to process even through a cheaper model.
It sounds like you already know what sales peoples incentives are. They don't care about the tiny players who wanna use tiny slices. I was referring to people who are trying to push PB through these. GCPs policies make a lot of sense if they are trying to get major players to switch their compute/data host to reduce overall costs.
Re: Gemini 3 Flash: Frontier intelligence built for speed
#598Earlier quoted context omitted.
I love how every single LLM model release is accompanied by pre-release insiders proclaiming how it’s the best model yet…
Thats true though. All these announcements beat all the other models on most benchmarks and are then the best model yet. They can't see the future yet so they are not aware or care anyway that 2 weeks later someone says "hold my beer" and we get again better benchmark results from someone else. Exhausting and exciting
Re: Gemini 3 Flash: Frontier intelligence built for speed
#599Re: Gemini 3 Flash: Frontier intelligence built for speed
#600Earlier quoted context omitted.
Ok, but then your "post" isn't scientific by definition since it cannot be verified. "Post" is in quotes because I don't know what you're trying to but you're implying some sort of public discourse. For fun: https://chatgpt.com/s/t_694361c12cec819185e9850d0cf0c629
As ChatGPT said to you: > A secret benchmark is: Useful for internal model selection That's what I'm doing.
The root of this whole discussion was a post about how Gemini 3 outperformed other models on some presumably informal question benchmark (a"vibe test"?). When asked for the benchmark, the response from the op and and someone else was that secrecy was needed to protect the benchmark from contamination. I'm skeptical of the need in the op's cases and I'm skeptical of the effectiveness of the secrecy in general. In a case where secrecy has actual value, why even discuss the benchmark publicly at all?