So, better than GPT4 according to the benchmarks? Looks very interesting. Technical paper: https://goo.gle/GeminiPaper Some details: - 32k context length - efficient attention mechanisms (for e.g. multi-query attention (Shazeer, 2019)) - audio input via Universal Speech Model (USM) (Zhang et al., 2023) features - no audio output? (Figure 2) - visual encoding of Gemini models is inspired by our own foundational work o…
I miss when ML scientific papers had actual science in them. Now they all feel like ads.
Gemini AI
231–240 of 1001 posts
Re: Gemini AI
#232I asked Bard, "Are you running Gemini Pro now?" And it told me, "Unfortunately, your question is ambiguous. "Gemini Pro" could refer to..." and listed a bunch of irrelevant stuff. Is Bard not using Gemini Pro at time of writing? The blog post says, "Starting today, Bard will use a fine-tuned version of Gemini Pro for more advanced reasoning, planning, understanding and more." (EDIT: it is... gave me a correct answer…
Bard shows “PaLM2” in my answers, and it says “I can't create images yet so I'm not able to help you with that” when I ask it to do so, which Gemini ought to be able to since its transformer can output images. I don’t think Bard is using Gemini Pro, perhaps because the rollout will be slow, but it is a bit of a blunder on Google’s part to indicate that it now uses it, since many will believe that this is the quality…
https://i.imgur.com/VhnQYie.png
So powerful.
Re: Gemini AI
#233I asked Bard, "Are you running Gemini Pro now?" And it told me, "Unfortunately, your question is ambiguous. "Gemini Pro" could refer to..." and listed a bunch of irrelevant stuff. Is Bard not using Gemini Pro at time of writing? The blog post says, "Starting today, Bard will use a fine-tuned version of Gemini Pro for more advanced reasoning, planning, understanding and more." (EDIT: it is... gave me a correct answer…
```
do you use Lambda model?
I use a variety of large language models, including the Lambda model, to perform my tasks. However, for each specific task, the best model is chosen based on its strengths and weaknesses. The Lambda model is particularly good at tasks that require understanding and responding to complex and nuanced language, such as generating creative text formats of text content and answering open ended, challenging, or strange questions. However, it may not be the best choice for tasks that require high accuracy or low latency.
Here's a breakdown of the models I use:
LaMDA: This is a family of Transformer-based neural language models specialized for dialogue. It is particularly good at understanding and responding to complex and nuanced language, and can generate different creative text formats of text content.
PaLM: This is another large language model, similar to LaMDA, but with a focus on factual accuracy. It is best suited for tasks that require precise and reliable information retrieval.
Meena: This is an older, smaller language model that was the precursor to LaMDA. It is still used for some tasks that require a more lightweight model. Other models: I also have access to a variety of other specialized language models, such as those for summarization, translation, and question answering. The specific model that is used for each task is determined by a variety of factors, including the nature of the task, the available resources, and the desired level of performance.
```
Re: Gemini AI
#234I just tried out a vision reasoning task: https://g.co/bard/share/e8ed970d1cd7 and it hallucinated. Hello Deepmind, are you taking notes?
Re: Gemini AI
#235So just a bunch of marketing fluff? I can use GPT4 literally right now and it’s apparently within a few percentage points of what Gemini Ultra can do… which has no release date as far as I can tell. Would’ve loved something more substantive than a bunch of videos promising how revolutionary it is.
Re: Gemini AI
#236Re: Gemini AI
#237Not impressed with the Bard update so far. I just gave it a screenshot of yesterday's meals pulled from MyFitnessPal, told it to respond ONLY in JSON, and to calculate the macro nutrient profile of the screenshot. It flat out refused. It said, "I can't. I'm only an LLM" but the upload worked fine. I was expecting it to fail maybe on the JSON formatting, or maybe be slightly off on some of the macros, but outright ref…
> I just gave it a screenshot of yesterday's meals pulled from MyFitnessPal, told it to respond ONLY in JSON, and to calculate the macro nutrient profile of the screenshot > Not impressed This made me chuckle Just a bit ago this would have been science fiction
Thanks for the reminder. It is definitely cool to be living in the future.
Re: Gemini AI
#238Not impressed with the Bard update so far. I just gave it a screenshot of yesterday's meals pulled from MyFitnessPal, told it to respond ONLY in JSON, and to calculate the macro nutrient profile of the screenshot. It flat out refused. It said, "I can't. I'm only an LLM" but the upload worked fine. I was expecting it to fail maybe on the JSON formatting, or maybe be slightly off on some of the macros, but outright ref…
Sounded like the update is coming out next week- did you get early access?
Re: Gemini AI
#239The improvement over ChatGPT are counted in (very) few percents. Does it mean they have entered a diminishing returns phase or is it that each percent is much harder to get compared to the previous ones ?