Live data from Hacker News

Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

blog.google

91–100 of 1001 posts

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#93

Gemini Ultra seems better on logic than GPT4. Still messing around testing but here's a prompt Ultra nailed but GPT4 completely botched: Tabitha likes cookies but not cake. She likes mutton but not lamb, and she likes okra but not squash. Following the same rule, will she like cherries or pears https://i.imgur.com/KW6gQbc.jpeg https://i.imgur.com/OSHSvLp.png

Note that Gemini pulled the answer off the Internet, while GPT-4 didn't. The answer can easily be found via Google search. Changing up the question a little, I reversed it and asked Ultra and it was unable to answer:

Jake likes coke but not pepsi. He likes corn but not popcorn, and he likes pens but not pencils. Will Jake like salmon or cheese?

https://i.imgur.com/lWU9HHS.png

edit: why was this downvoted? I don't understand Hacker News, and I've been here for over 12 years.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#94
post #85

Wow, I pay £1.50 currently, which is paid nicely via google rewards every month for a year, ticks along nicely. A £18.99/month bolt-on, nope, if couple quid, sure, but just priced out and what I would call top-end whale marketing price farming, which down the line will, I predict - half before years out at least.

I guess it's competing with ChatGPT+ at about £16 / month. And you get a bunch of extra storage and Workspace features from the Premium Google One too.

If it is actually as good as GPT-4, I can imagine lots of people swithing subscription to get all the other Google One stuff cheap/free. But you'd have to be very into Google - full benefit looks like it needs you to use various Workspace features from Google One for your whole family?

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#96
Looking forward to someone writing a review. So far Gemini has shown capabilities all over the place. Image generation quality feels a bit worse than what I've seen from DALL-E and Stable Diffusion. Does Ultra provide superior image generation capabilities?

Something I've noticed about Gemini is that usually it'll respond to my query correctly, but it's never the default draft. If I look through each draft one of the options will usually contain the correct answer though.

I'm pleased to find that capabilities have been improving. When Gemini was initially released, asking for something like "How many views have the last 5 mrbeast videos gotten?" wouldn't generate a useful reply. But now it lists the latest 5 videos and one of the drafts even includes the total added up.

Asking Gemini to generate video summaries seems to work really well on some videos, but for others it just gives an error... Are YouTube creators allowed to opt-out of Gemini interactions?

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#97

Gemini Ultra seems better on logic than GPT4. Still messing around testing but here's a prompt Ultra nailed but GPT4 completely botched: Tabitha likes cookies but not cake. She likes mutton but not lamb, and she likes okra but not squash. Following the same rule, will she like cherries or pears https://i.imgur.com/KW6gQbc.jpeg https://i.imgur.com/OSHSvLp.png

Proof of Gemini cheating: https://i.imgur.com/eYJDFjS.png

Answer about cherries falling from the sky...

(there is no question or context beforehand, this is the first question of the chat)

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#98

Gemini Ultra seems better on logic than GPT4. Still messing around testing but here's a prompt Ultra nailed but GPT4 completely botched: Tabitha likes cookies but not cake. She likes mutton but not lamb, and she likes okra but not squash. Following the same rule, will she like cherries or pears https://i.imgur.com/KW6gQbc.jpeg https://i.imgur.com/OSHSvLp.png

I would not be so quick to jump to conclusions. GPT-4 beats it easily in this simple logic puzzle: https://www.reddit.com/r/singularity/comments/1altttv/bard_a...

We need more data.

Re: Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced

#100

Gemini Ultra seems better on logic than GPT4. Still messing around testing but here's a prompt Ultra nailed but GPT4 completely botched: Tabitha likes cookies but not cake. She likes mutton but not lamb, and she likes okra but not squash. Following the same rule, will she like cherries or pears https://i.imgur.com/KW6gQbc.jpeg https://i.imgur.com/OSHSvLp.png

I would have never guessed the answer. With such little data available, one can invent any arbitrary rules to fit their favorite answer.

It would be more impressive to practical use cases, if a LLM simply said that it's impossible to guess without inventing their own reasoning or looking up the answer online.

Post reply on HN