Live data from Hacker News

Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

developers.googleblog.com

101–110 of 151 posts

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#101
post #35

This sounds interesting: "We will continue to offer a suite of safety filters that developers may apply to Google’s models. For the models released today, the filters will not be applied by default so that developers can determine the configuration best suited for their use case."

There's still basic filters even if you take all the ones that you can turn off from the UI all off. It's still not capable of summarizing some YA novels I tried to feed it because of those filters.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#102
post #32

Earlier quoted context omitted.

The Aider leaderboards seem like a good practical test of coding usefulness: https://aider.chat/docs/leaderboards/ . I haven't tried Cursor personally but I am finding Aider with Sonnet more useful that Github Copilot and its nice to be able to pick any model API. Eventually even a local model may be viable. This new Gemini model does not rank very high unfortunately.

Thanks for the link. That's unfortunate, though perhaps the benchmarks will be updated after this latest Gemini release. Cursor with Sonnet is great, I'll have to give Aider a try as well.

It's on the leaderboard, it's tied with qwen 2.5 72b and far below SOTA of o1, claude sonnet, and deepseek. (also below very old models like gpt-4-0314 lol)

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#103

Earlier quoted context omitted.

This is just incorrect. The OpenAI models hosted though Azure are HIPAA-compliant, and Antropic will also sign a BAA.

I'm open to being wrong. However for many industries your still running the risk of leaking data via a 3rd party service. You can run Llama3 on prem, which eliminates that risk. I try to reduce reliance on 3rd party services when possible. I still have PTSD from Saucelabs constantly going down and my manager berating me over it.

You are not technically wrong because a statement "there is a risk of leaking data" is not falsifiable. But your comment is performative cynicism to display your own high standards. For the very vast majority of people and companies, privacy standards-compliant services (like HIPAA-compliant) are private enough.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#104
post #93

I’ve used it. The API is incredibly buggy and flakey. A particular pain point is the “recitation error” fiasco. If you’re developing a real world app this basically makes the Gemini api unusable. It strikes me as a kind of “Potemkin” service. Google is aware of the issue and it has been open on google's bug tracker since March 2024: https://issuetracker.google.com/issues/331677495 There is also discussion on GitHub:…

[deleted]

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#105
post #87
post #76

Earlier quoted context omitted.

Probably not. Do they really believe they are going to knock OpenAI out of business, when the OpenAI models are better? Instead I think they are going after the "Android model". Recognize they might not be able to dethrone the leader who invented the space. Define yourself in the marketplace as the cheaper alternative. "Less good but almost as good." In the end, they hope to be one of a small number of surviving memb…

Cheapness has a quality all its own. Gemini is substantially cheaper to run (in consumer prices, and likely internally as well) than OpenAI's models. You might wonder, what's the value in this, if the model isn't leading? But cheaper inference could potentially be a killer edge when you can scale test-time compute for reasoning. Scaling test-time compute is, after all, what makes o1 so powerful. And this new Gemini d…

In Home Assistant you can use LLMs to control your Home with your voice. Gemini performs similar to the GPT models, and with the cost difference there is little reason to choose OpenAi

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#106
post #87
post #76

Earlier quoted context omitted.

Probably not. Do they really believe they are going to knock OpenAI out of business, when the OpenAI models are better? Instead I think they are going after the "Android model". Recognize they might not be able to dethrone the leader who invented the space. Define yourself in the marketplace as the cheaper alternative. "Less good but almost as good." In the end, they hope to be one of a small number of surviving memb…

Cheapness has a quality all its own. Gemini is substantially cheaper to run (in consumer prices, and likely internally as well) than OpenAI's models. You might wonder, what's the value in this, if the model isn't leading? But cheaper inference could potentially be a killer edge when you can scale test-time compute for reasoning. Scaling test-time compute is, after all, what makes o1 so powerful. And this new Gemini d…

Google still has an unbelievable training infrastructure advantage. The second they can figure out how to convert that directly to model performance without worrying about data (as the o1 blog post seemed to imply OAI had) they’ll be kings.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#107
post #41

Earlier quoted context omitted.

I can’t see that price on https://openai.com/api/pricing/ - it’s listing $5/m input and $15/m output for GPT-4o right now. No wait, correction: That’s confusing: it lists 4o first and then lists gpt-4o-2024-08-06 as $2.50/$10.

apologies: it's taken us a minute to switch the default `gpt-4o` pointer to the newest snapshot we're planning on doing that default change next week (October 2nd). And you can get the lower prices now (and the structured outputs feature) by manually specify `gpt-4o-2024-08-06`

> “You can”

No, “I” can’t.

Open AI has always trickled out model access, putting their customers into “tiers” of access. I’m not sufficiently blessed by the great Sam to have immediate access.

On, and Azure Open AI especially likes to drag their feet both consistently, and also on a per-region basis.

I live in a “no model for you” region.

Open AI says: “Wait your turn, peasant” while claiming to be about democratising access.

Google and everyone else just gives access, no gatekeeping.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#108
post #105
post #87

Earlier quoted context omitted.

Cheapness has a quality all its own. Gemini is substantially cheaper to run (in consumer prices, and likely internally as well) than OpenAI's models. You might wonder, what's the value in this, if the model isn't leading? But cheaper inference could potentially be a killer edge when you can scale test-time compute for reasoning. Scaling test-time compute is, after all, what makes o1 so powerful. And this new Gemini d…

In Home Assistant you can use LLMs to control your Home with your voice. Gemini performs similar to the GPT models, and with the cost difference there is little reason to choose OpenAi

Using either frontier model for basic edge device problems is wasteful. Use something cheap. We're asking "is there a profitable niche between the best & runner-up models?" I believe so.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#109
Cool, now all Google has to do is make it easier to onboard new GCP customers and more people will probably use it...its comical how hard it is to create a new GCP organization & billing account. Also I think more Workspace customers would probably try Gemini if it was a usage-based trial as opposed to clicking a "Try for 14 days" CTA to activate a new subscription.

Re: Two new Gemini models, reduced 1.5 Pro pricing, increased rate limits, and more

#110
post #94

Has anyone used Gemini Code Assist? I'm curious how it compares with Github Copilot and Cursor.

I have used Github Copilot extensively within VS Code for several months. The autocomplete - fast and often surprisingly accurate - is very useful. My only complaint is when writing comments, I find the completions distracting to my thought process. I tried Gemini Code Assist and it was so bad by comparison that I turned it off within literally minutes. Too slow and inaccurate. I also tried Codestral via the Continue…

Have you tried Gitlab Duo and if so, what are your thoughts on that?
Post reply on HN