Live data from Hacker News

Gemini AI

deepmind.google

371–380 of 1001 posts

Re: Gemini AI

#371
post #200
post #60

One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…

To add to my comment above: Google DeepMind put out 16 videos about Gemini today, the total watch time at 1x speed is about 45 mins. I've now watched them all (at >1x speed). In my opinion, the best ones are: * https://www.youtube.com/watch?v=UIZAiXYceBI - variety of video/sight capabilities * https://www.youtube.com/watch?v=JPwU1FNhMOA - understanding direction of light and plants * https://www.youtube.com/watch?v=D…

Watching these videos made me remember this cool demo Google did years ago where their earpods would auto translate in realtime a conversation between two people talking different languages. Turned out to be demo vaporware. Will this be the same thing?

Re: Gemini AI

#372

I just tried out a vision reasoning task: https://g.co/bard/share/e8ed970d1cd7 and it hallucinated. Hello Deepmind, are you taking notes?

Is this something we really expect AI to get right with high accuracy with an image like that?

For one, there's a huge dark line that isn't even clear to me what it is and what that means for street crossings.

I am definitely not confident I could answer that question correctly.

Re: Gemini AI

#373

This is very cool and I am excited to try it out! But, according to the metrics, it barely edges out GPT-4 -- this mostly makes me _more_ impressed with GPT-4 which: - came out 9 months ago AND - had no direct competition to beat (you know Google wasn't going to release Gemini until it beat GPT-4) Looking forward to trying this out and then seeing OpenAI's answer

OpenAI had an almost five-year head-start with relevant data acquisition and sorting, which is the most important part of these models.

Google has the biggest proprietary moat of information of any company in the world I'm sure.

Re: Gemini AI

#376
post #324

So, better than GPT4 according to the benchmarks? Looks very interesting. Technical paper: https://goo.gle/GeminiPaper Some details: - 32k context length - efficient attention mechanisms (for e.g. multi-query attention (Shazeer, 2019)) - audio input via Universal Speech Model (USM) (Zhang et al., 2023) features - no audio output? (Figure 2) - visual encoding of Gemini models is inspired by our own foundational work o…

The table is *highly* misleading. It uses different methodologies all over the place. For MMLU, it highlights the CoT @ 32 result, where Ultra beats GPT4, but it loses to GPT4 with 5-shot, for example. For GSM8K it uses Maj1@32 for Ultra and 5-shot CoT for GPT4, etc. Then also, for some reason, it uses different metrics for Ultra and Pro, making them hard to compare. What a mess of a "paper".

It really feels like the reason this is being released now and not months ago is that that's how long it took them to figure out the convoluted combination of different evaluation procedures to beat GPT-4 on the various benchmarks.

Re: Gemini AI

#377
post #293

To test whether bard.google.com is already updated in your region, this prompt seems to work: Which version of Bard am I using? Here in Europe (Germany), I get: The current version is Bard 2.0.3. It is powered by the Google AI PaLM 2 model Considering that you have to log in to use Bard while Bing offers GPT-4 publicly and that Bard will be powered by Gemini Pro, which is not the version that they say beats GPT-4, it…

I'm getting little "PaLM2" badges on my Bard responses.

Re: Gemini AI

#378

Google is number 1 at launching also-rans and marketing sites with feature lists that show how their unused products are better than the competition. Someday maybe they’ll learn why nobody uses their shit.

Ah, yes, the company with by far the most users in the world - and no one uses their shit.

Re: Gemini AI

#379
While this must be an incredible technical achievement for the team, as a simple user I will only see value when Google ships a product that's better than OpenAI's, and that's yet to be seen.

Re: Gemini AI

#380
post #376
post #324

Earlier quoted context omitted.

The table is *highly* misleading. It uses different methodologies all over the place. For MMLU, it highlights the CoT @ 32 result, where Ultra beats GPT4, but it loses to GPT4 with 5-shot, for example. For GSM8K it uses Maj1@32 for Ultra and 5-shot CoT for GPT4, etc. Then also, for some reason, it uses different metrics for Ultra and Pro, making them hard to compare. What a mess of a "paper".

It really feels like the reason this is being released now and not months ago is that that's how long it took them to figure out the convoluted combination of different evaluation procedures to beat GPT-4 on the various benchmarks.

[deleted]
Post reply on HN