Live data from Hacker News

Gemini AI

deepmind.google

291–300 of 1001 posts

Re: Gemini AI

#292

Is it live already at bard.google.com? Just tried it and still useless compared to GPT 3.5.

It depends on your region. In general these things take some time (hours) to go live globally to all enabled regions, and are done carefully. If you come back tomorrow or in a few days it's more likely to have reached you, assuming you're in an eligible region. It's probably best to wait until the UI actually tells you Bard has been updated to Gemini Pro. Previous Bard updates have had UI announcements so I'd guess (…

I don't understand how anyone can see a delayed EU launch as anything other than a red flag. It's basically screaming "we didn't care about privacy and data protection when designing this".

Re: Gemini AI

#293
To test whether bard.google.com is already updated in your region, this prompt seems to work:

    Which version of Bard am I using?
Here in Europe (Germany), I get:

    The current version is Bard 2.0.3. It is
    powered by the Google AI PaLM 2 model
Considering that you have to log in to use Bard while Bing offers GPT-4 publicly and that Bard will be powered by Gemini Pro, which is not the version that they say beats GPT-4, it seems Microsoft and OpenAI are still leading the race towards the main prize: Replacing search+results with questions+answers.

I'm really curious to see the next SimilarWeb update for Bing and Google. Does anybody here already have access to the November numbers? I would expect we can already see some migration from Google to Bing because of Bing's inclusion of GPT-4 and Dall-E.

Searches for Bing went throught the roof when they started to offer these tools for free:

https://trends.google.de/trends/explore?date=today+5-y&q=bin...

Re: Gemini AI

#294

One thing I noticed is I asked Bard "can you make a picture of a black cat?" It says no I can't make images yet. So I asked "can you find one in Google search?" It did not know what I meant by "one" (the subject cat from previous question). Chat GPT4 would have no issue with such context.

I reproduced your result, but then added "Didn't I just ask you for a picture of a black cat?" and it gave me some. Meh.

Re: Gemini AI

#295
post #186

For others that were confused by the Gemini versions: the main one being discussed is Gemini Ultra (which is claimed to beat GPT-4). The one available through Bard is Gemini Pro . For the differences, looking at the technical report [1] on selected benchmarks, rounded score in %: Dataset | Gemini Ultra | Gemini Pro | GPT-4 MMLU | 90 | 79 | 87 BIG-Bench-Hard | 84 | 75 | 83 HellaSwag | 88 | 85 | 95 Natural2Code | 75 |…

Thanks, I was looking for clarification on this. Using Bard now does not feel GPT-4 level yet, and this would explain why.

Re: Gemini AI

#296
post #286

This is hilarious for anyone who knows the area: "The best way to get from Lake of the Clouds Hut to Madison Springs Hut in the White Mountains is to hike along the Mt. Washington Auto Road. The distance is 3.7 miles and it should take about 16 minutes." What it looks like it's doing is actually giving you the driving directions from the nearest road point to one hut to the nearest road point to the other hut. An ear…

I mean even without knowing the area if you are hiking (which implies you are walking) 3.7 miles in 16 m then you are the apex predator of the world my friend. That's 20/25 km/h

Re: Gemini AI

#297
post #256

This announcement makes we wonder if we are approaching a plateau in these systems. They are essentially claiming close to parity with gpt-4, not a spectacular new breakthrough. If I had something significantly better in the works, I'd either release it or hold my fire until it was ready. I wouldn't let openai drive my decision making, which is what this looks like from my perspective. Their top line claim is they ar…

I don’t think we can declare a plateau just based on this. Actually, given that we have nothing but benchmarks and cherry picked examples, I would not be so quick to believe GPT-4V has been bested. PALM-2 was generally useless and plagued by hallucinations in my experience with Bard. It’ll be several months till Gemini Pro is even available. We also don’t know basic facts like the number of parameters or training set size.

I think the real story is that Google is badly lagging their competitors in this space and keeps issuing press releases claiming they are pulling ahead. In reality they are getting very little traction vs. OpenAI.

I’ll be very interested to see how LLMs continue to evolve over the next year. I suspect we are close to a model that will outperform 80% of human experts across 80% of cognitive tasks.

Re: Gemini AI

#298

I asked Bard, "Are you running Gemini Pro now?" And it told me, "Unfortunately, your question is ambiguous. "Gemini Pro" could refer to..." and listed a bunch of irrelevant stuff. Is Bard not using Gemini Pro at time of writing? The blog post says, "Starting today, Bard will use a fine-tuned version of Gemini Pro for more advanced reasoning, planning, understanding and more." (EDIT: it is... gave me a correct answer…

Your line of thinking also presupposes that Bard is self aware about that type of thing. You could also ask it what programming language it's written in, but that doesn't mean it knows and/or will answer you.

This is a common occurance I'm seeing lately. People treating these things as oracles and going straight to chatgpt/bard instead of thinking or researching for themselves

Re: Gemini AI

#299

I did some side-by-side comparisons of simple tasks (e.g. "Write a WCAG-compliant alternative text describing this image") with Bard vs GPT-4V. Bard's output was significantly worse. I did my testing with some internal images so I can't share, but will try to compile some side-by-side from public images.

Bard with pro is apparently text only:

> Important: For now, Bard with our specifically tuned version of Gemini Pro works for text-based prompts, with support for other content types coming soon.

https://support.google.com/bard/answer/14294096

I'm in the UK and it's not available here yet - I really wish they'd be clearer about what I'm using, it's not the first time this has happened.

Re: Gemini AI

#300

I did some side-by-side comparisons of simple tasks (e.g. "Write a WCAG-compliant alternative text describing this image") with Bard vs GPT-4V. Bard's output was significantly worse. I did my testing with some internal images so I can't share, but will try to compile some side-by-side from public images.

I'm researching using LLMs for alt-text suggestion for forum users, can you share your finding so far? Outside of GPT-4V I had good first results with https://github.com/THUDM/CogVLM

As a heads up, bard with gemini pro only works with text.
Post reply on HN