Pelicans for 3.6 Flash and 3.5 Flash-Lite (Cyber isn't available to me through the API yet.) https://tools.simonwillison.net/markdown-svg-renderer#url=ht...
Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
591–600 of 616 posts
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#592Man I love Gemini models but these kind of pricing increase is just insanse. I have a little product and I have to keep increasing the price and reduce the limits because of this non-sense, and they did not even let us use the old models in near future, so I forced to update to the new model with basically no to little improvement because I don't even need that much. Google if you can read this, it okay to release ne…
What is your use case? Have you considered moving to open source / Chinese models? If gemini-2.5-flash-lite is good enough for your application, you will find even lower cost options with better performance outside of the Google ecosystem.
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#593Man I love Gemini models but these kind of pricing increase is just insanse. I have a little product and I have to keep increasing the price and reduce the limits because of this non-sense, and they did not even let us use the old models in near future, so I forced to update to the new model with basically no to little improvement because I don't even need that much. Google if you can read this, it okay to release ne…
Do not base products on models that are not open-weights. Doing it is like building a product on someone else's platform, you are entirely at their mercy, and even when they don't have any reason to hurt you, you are tiny enough that if any policy they want to enact hurts you as a side effect, no-one is going to care. You don't have to self-host the open-weights model, you just need to be able to source it from multi…
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#594Earlier quoted context omitted.
I'm in the same situation. But I was shocked discovering it goes both ways: many new Gemini functionalities are only accessible using a consumer account instead of a Workspace account. Also, Gemini is now the only major AI assistant with no support for MCP connectors. Instead of adding this to the core product, like ChatGPT and Claude did, somebody at Google decided that it was smarter to add this fundamental feature…
i think Google would see more success if they kept the CEO and everyone at the bottom (ie. doesn't manage anyone), and fired everyone else. Build a whole new management tree - the current people all do a terrible job.
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#595Earlier quoted context omitted.
Why was this flagged? This is absolutely true, Americans seem to not actually want to use Chinese models at least for coding, maybe for other inference use cases but I haven't seen it. No one I know uses anything but OpenAI and Anthropic even if Chinese models are better or cheaper in many use cases.
First 6 of the most used models on openrouter currently are Chinese; that's true for code generation and other coding-relevant tasks too when ranked by share of tokens. https://openrouter.ai/rankings#top-models Of course openrouter is not representative because most users directly go to the model provider but it still proves your claim is very far from "absolutely true".
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#596Pelicans for 3.6 Flash and 3.5 Flash-Lite (Cyber isn't available to me through the API yet.) https://tools.simonwillison.net/markdown-svg-renderer#url=ht...
I am growing tired of these pelicans posts every time a new model is published. Feels to me like low effort personal brand promotion. Just sharing my 2 cents.
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#597Earlier quoted context omitted.
> But why people choose to bring the mediocrity of Google into their lives is beyond me. Because life gets simpler and better if you are just using what is available on your phone and in your browser instead of constantly chasing current HN darling.
That’s one braindead strawman. There’s millions of apps to choose from on your phone. Alternatives to Google are decades old. But do enjoy getting your ass profiled in gmail and let me know if the ads in your mailbox are helpfully tailored to your needs and preferences!
And choosing one of them obviously makes life more complicated than not doing that.
> getting your ass profiled
Gmail knows nothing of my ass.
> let me know if the ads in your mailbox are helpfully tailored to your needs and preferences!
I couldn't, for I don't check them. I have enough life to live without degoogling crusade.
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#598Earlier quoted context omitted.
First 6 of the most used models on openrouter currently are Chinese; that's true for code generation and other coding-relevant tasks too when ranked by share of tokens. https://openrouter.ai/rankings#top-models Of course openrouter is not representative because most users directly go to the model provider but it still proves your claim is very far from "absolutely true".
A big slice of a small pie does not prove the majority. And I'm speaking as someone who is indeed a user of those OpenRouter Chinese models.
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#599Earlier quoted context omitted.
It says the original report was in the Information, which I can't see, but I'm skeptical that they includes the training cost? And how much that changes the figure?
That's the profit margin on inference, not overall. Each model does end up being profitable over its lifetime, but the money they're making is being immediately churned into buying more data centers & the training for the next giant model up, so they're not profitable overall at the moment. It's a bit like how Amazon kept churning their profits into more growth instead of taking the profit early. That said, Anthropic…
Re: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
#600Earlier quoted context omitted.
> China can keep up because it's cheaper to run a frontier lab there Not sure if this is what you meant, but their training runs are significantly cheaper. This was one of the big shockers from the Deepseek R1 paper. US foreign policy has helped to ensure that the Chinese are compute constrained, so they literally cannot buy the most expensive and powerful training rigs. This has led to a steady drumbeat of innovatio…
I’ve often thought it was the cost of electricity and crypto mining being banned a few years ago. China has a lot of “stranded electricity” which fits this usecase well.