One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…
To add to my comment above: Google DeepMind put out 16 videos about Gemini today, the total watch time at 1x speed is about 45 mins. I've now watched them all (at >1x speed). In my opinion, the best ones are: * https://www.youtube.com/watch?v=UIZAiXYceBI - variety of video/sight capabilities * https://www.youtube.com/watch?v=JPwU1FNhMOA - understanding direction of light and plants * https://www.youtube.com/watch?v=D…
Gemini AI
371–380 of 1001 posts
Re: Gemini AI
#372I just tried out a vision reasoning task: https://g.co/bard/share/e8ed970d1cd7 and it hallucinated. Hello Deepmind, are you taking notes?
For one, there's a huge dark line that isn't even clear to me what it is and what that means for street crossings.
I am definitely not confident I could answer that question correctly.
Re: Gemini AI
#373This is very cool and I am excited to try it out! But, according to the metrics, it barely edges out GPT-4 -- this mostly makes me _more_ impressed with GPT-4 which: - came out 9 months ago AND - had no direct competition to beat (you know Google wasn't going to release Gemini until it beat GPT-4) Looking forward to trying this out and then seeing OpenAI's answer
OpenAI had an almost five-year head-start with relevant data acquisition and sorting, which is the most important part of these models.
Re: Gemini AI
#374Re: Gemini AI
#375Re: Gemini AI
#376So, better than GPT4 according to the benchmarks? Looks very interesting. Technical paper: https://goo.gle/GeminiPaper Some details: - 32k context length - efficient attention mechanisms (for e.g. multi-query attention (Shazeer, 2019)) - audio input via Universal Speech Model (USM) (Zhang et al., 2023) features - no audio output? (Figure 2) - visual encoding of Gemini models is inspired by our own foundational work o…
The table is *highly* misleading. It uses different methodologies all over the place. For MMLU, it highlights the CoT @ 32 result, where Ultra beats GPT4, but it loses to GPT4 with 5-shot, for example. For GSM8K it uses Maj1@32 for Ultra and 5-shot CoT for GPT4, etc. Then also, for some reason, it uses different metrics for Ultra and Pro, making them hard to compare. What a mess of a "paper".
Re: Gemini AI
#377To test whether bard.google.com is already updated in your region, this prompt seems to work: Which version of Bard am I using? Here in Europe (Germany), I get: The current version is Bard 2.0.3. It is powered by the Google AI PaLM 2 model Considering that you have to log in to use Bard while Bing offers GPT-4 publicly and that Bard will be powered by Gemini Pro, which is not the version that they say beats GPT-4, it…
Re: Gemini AI
#378Google is number 1 at launching also-rans and marketing sites with feature lists that show how their unused products are better than the competition. Someday maybe they’ll learn why nobody uses their shit.
Re: Gemini AI
#379Re: Gemini AI
#380Earlier quoted context omitted.
The table is *highly* misleading. It uses different methodologies all over the place. For MMLU, it highlights the CoT @ 32 result, where Ultra beats GPT4, but it loses to GPT4 with 5-shot, for example. For GSM8K it uses Maj1@32 for Ultra and 5-shot CoT for GPT4, etc. Then also, for some reason, it uses different metrics for Ultra and Pro, making them hard to compare. What a mess of a "paper".
It really feels like the reason this is being released now and not months ago is that that's how long it took them to figure out the convoluted combination of different evaluation procedures to beat GPT-4 on the various benchmarks.