Live data from Hacker News

Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

blog.google

911–920 of 1001 posts

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#911

I'm surprised they got rid of the Bard name. It struck me as a really smart choice since a Bard is someone who said things, and it's an old/archaic enough word to not already be in a zillion other names. Gemini, on the other hand, doesn't strike me as particularly relevant (except that perhaps it's a twin of ChatGPT?), and there are other companies with the same name. EDIT: I can see the advantage of picking a name t…

The Bard name gave me a warm fuzzy feeling immediately transporting me back to my youth playing (or at least trying to play) Bard's Tale. The name evoked adventure, excitement and a good dose of dread. And, the idea of it being "role playing" struck me as a master meta stroke. Gemini, from the mythological standpoint, seemed to make more sense to me from an overall business/marketing standpoint. "This AI thing right…

And similarly anyone playing modern tabletop RPGs will probably associated "Bard" with the smart, charismatic person who buffs the party and debuffs your enemies; perfect for an AI assistant

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#912
post #757
post #701

Earlier quoted context omitted.

Incredible. Gpt4 spots that the door is transparent and that changes things but has this great line > When you initially pick a door (in this case, door number 1 where you already see the car), you have a 1/3 chance of having picked the car (Asking it to explain this it correctly solves the problem but it's a wonderfully silly sentence) Edit - in a new chat it gets it right the first time

This is not convincing though that gpt4 actually understands the problem. Here's a slight variation I asked and it fails miserably. https://chat.openai.com/share/22a9027f-a2c1-428a-94a2-8fd918... I wonder what lends itself it answer correct in one situation but not the other? Was your question previously asked already and it recognized it whereas my question is different enough?

> Was your question previously asked already and it recognized it

Given that LLMs training data consists to a large extent of "stuff people have written on the internet", and The Monty Hall Problem is something that comes up as a topic for discussion on the internet not entirely infrequently - as well as having a wikipedia page - yes, I suspect that the words describing the monty hall problem being followed by words describing the correct solution appeared often in the training set, so LLMs are likely to reproduce that.

Words describing a problem similar to the monty hall problem are going to be less common, and probably have a lot of discussion about whether they accurately match the monty hall problem, and disagreement about what the right answer is. LLMs will confabulate something that looks like a plausible answer based on the language used in those discussions, because that's how they work. Whether they get a right answer is probably going to be much more up to chance.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#913

Earlier quoted context omitted.

> I take it for granted that all these services are going to be free. They are a goldmine for behavioral and persuasion engineers. They are also a goldmine for LLMs. Training on human text is necessary for AIs but it has one major flaw - it is so called "off-policy". That means it portrays human behavior and human errors. While human-AI chat logs portray AI errors, so they are better material to generate training dat…

> chatGPT is reportedly serving 10M customers and let's assume 10K tokens/month/user No way. Definitely too high once you remove their system prompts. > In one year they have 12T tokens, while their original training set for GPT-4 was rumored to be 13T tokens. This sounds great for understanding use, but the quality to train on seems terrible.

You might be right, a LLM alone doesn't improve by itself. But when it is part of a system like GPT's, then it can use web search, local RAG, code execution and also get human guidance and corrections. Clearly superior setup that improves over the LLM alone. I believe that is why OpenAI created GPT's, to lift a model at level N to level N+1.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#914

I'm surprised they got rid of the Bard name. It struck me as a really smart choice since a Bard is someone who said things, and it's an old/archaic enough word to not already be in a zillion other names. Gemini, on the other hand, doesn't strike me as particularly relevant (except that perhaps it's a twin of ChatGPT?), and there are other companies with the same name. EDIT: I can see the advantage of picking a name t…

As a Brit I’ll be glad to see the word association gradually return to that of our greatest playwright rather than something appropriated by big tech.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#915
post #757
post #701

Earlier quoted context omitted.

Incredible. Gpt4 spots that the door is transparent and that changes things but has this great line > When you initially pick a door (in this case, door number 1 where you already see the car), you have a 1/3 chance of having picked the car (Asking it to explain this it correctly solves the problem but it's a wonderfully silly sentence) Edit - in a new chat it gets it right the first time

This is not convincing though that gpt4 actually understands the problem. Here's a slight variation I asked and it fails miserably. https://chat.openai.com/share/22a9027f-a2c1-428a-94a2-8fd918... I wonder what lends itself it answer correct in one situation but not the other? Was your question previously asked already and it recognized it whereas my question is different enough?

Your link is not to GPT4, your link is to the free version of ChatGPT, aka gpt-3.5-turbo (you can tell because the icon is green, not purple).

GPT4 indeed understands your variant, as evidenced here: https://chat.openai.com/share/46916f21-c469-4e93-9bed-bbd18b...

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#916

Earlier quoted context omitted.

I'm not surprised -- I thought Bard was terrible branding. It's all associations with Shakespeare and poetry and medieval England, and as much as I might personally enjoy those, it's extremely backwards-looking, with archaic connotations. Also it sounds close to "beard" -- hairy stuff. Gemini sounds like the space program -- futuristic, a leap for mankind. It's got all the right emotional associations. It's a constel…

Interesting. I don’t like the name at all because it makes me think of people who take horoscopes seriously. You’re impression seems to be untainted by that which is nice

Funny, my mind goes to NASA and the Gemini program.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#917
post #787
post #756

Earlier quoted context omitted.

User: Which is heavier, two pounds of feathers or a pound of bricks? Gemini: Neither! They both weigh the same, which is two pounds. The key is in the wording of the question. You're comparing equal amounts of two different materials (two pounds of feathers vs. one pound of bricks). As long as the quantity is the same, the weight (force exerted by gravity) will always be the same regardless of the material. This is a…

This is surely in the training data.

[deleted]

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#918
post #756

I’ve been pretty excited to finally try Gemini advanced. So far pretty disappointed. Here’s my go-to test question - which even chat gpt 3.5 can get. Question: I walk up to a glass door. It has the word push on it in mirror writing. Should I push or pull the door, and why Gemini advanced: You should push the door. Here's why: * Mirror Writing: The word "PUSH" is written in mirror writing, meaning it would appear corr…

User: Which is heavier, two pounds of feathers or a pound of bricks? Gemini: Neither! They both weigh the same, which is two pounds. The key is in the wording of the question. You're comparing equal amounts of two different materials (two pounds of feathers vs. one pound of bricks). As long as the quantity is the same, the weight (force exerted by gravity) will always be the same regardless of the material. This is a…

Interesting. Based on this conversation[1], I think Gemini Ultra is massively overfit. Make it do unit conversions or use units it hasn't seen in the same framing before and it does well. But stay close enough to the original trick question (1 and 1) and it fails.

[1] https://g.co/gemini/share/94bfb8f9ebea

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#919

Earlier quoted context omitted.

That's a bummer. I just made this one up which GPT-4 failed and Gemini got right but definitely need to do more testing to see what's being taken from online and what is actual reasoning. https://i.imgur.com/3sNr3LW.png https://i.imgur.com/EIj0nZg.png Edit: When I did your prompt it got it right on the first try fwiwi https://i.imgur.com/E3zYEca.png

I tried through API with latest GPT-4-Turbo-Preview and this is what I got: ------ Let's break down the information given: The school is two blocks north of the pool. This means if you are at the pool and move two blocks north, you'll arrive at the school. The convenience store is one block south of the school. Therefore, if you start at the school and move one block south, you'll reach the convenience store. Based o…

You know you can share conversations right?
Post reply on HN