Live data from Hacker News

Gemini AI

deepmind.google

711–720 of 1001 posts

Re: Gemini AI

#712

Gemini Ultra isn't released yet and is months away still. Bard w/ Gemini Pro isn't available in Europe and isn't multi-modal, https://support.google.com/bard/answer/14294096 No public stats on Gemini Pro. (I'm wrong. Pro stats not on website, but tucked in a paper - https://storage.googleapis.com/deepmind-media/gemini/gemini_... ) I feel this is overstated hype. There is no competitor to GPT-4 being released today. I…

I bet that it will land on Google's graveyard before it gets released worldwide.

Re: Gemini AI

#713

Earlier quoted context omitted.

He's not wrong. DeepMind spends time solving big scientific / large-scale problems such as those in genetics, material science or weather forecasting, and Google has untouchable resources such as all the books they've scanned (and already won court cases about) They do make OpenAI look like kids in that regard. There is far more to technology than public facing goods/products. It's probably in part due to the cultura…

While you are spot on, I cannot avoid thinking of 1996 or so. On one corner: IBM Deep Blue winning vs Kasparov. A world class giant with huge research experience. On the other corner, Google, a feisty newcomer, 2 years in their life, leveraging the tech to actually make something practical. Is Google the new IBM?

I think the analogy is kind of strained here - at the current stage, OpenAI doesn't have an overwhelming superiority in quality in the same way Google once did. And, if marketing claims are to be believed, Google's Gemini appears to be no publicity stunt. (not to mention that IBM's "downfall" isn't very related to Deep Blue in the first place)

Re: Gemini AI

#714

To me it doesn't look impressive at all. In this video: https://www.youtube.com/watch?v=LvGmVmHv69s , Google talked about solving a competitive programming problem using dynamic programming. But DP is considered only an intermediate level technique in National Olympiad in Informatics/USACO level competitions, which are targeted at secondary school students. For more advanced contests the tough questions usually requi…

DP?

Re: Gemini AI

#715
post #60

One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…

I find this video really freaky. It’s like Gemini is a baby or very young child and also a massively know it all adult that just can’t help telling how clever it is and showing off its knowledge. People speak of the uncanny valley in terms of appearance. I am getting this from Gemini. It’s sort of impressive but feels freaky at the same time. Is it just me?

No, there's an odd disconnect between the impressiveness of the multimodal capabilities vs the juvenile tone and insights compared to something like GPT-4 that's very bizarre in application.

It is a great example of what I've been finding a growing concern as we double down on Goodhart's Law with the "beats 30 out of 32 tests compared to existing models."

My guess is those tests are very specific to evaluations of what we've historically imagined AI to be good at vs comprehensive tests of human ability and competencies.

So a broad general pretrained model might actually be great at sounding 'human' but not as good at logic puzzles, so you hit it with extensive fine tuning aimed at improving test scores on logic but no longer target "sounding human" and you end up with a model that is extremely good at what you targeted as measurements but sounds like a creepy toddler.

We really need to stop being so afraid of anthropomorphic evaluation of LLMs. Even if the underlying processes shouldn't be anthropomorphized, the expressed results really should be given the whole point was modeling and predicting anthropomorphic training data.

"Don't sound like a creepy soulless toddler and sound more like a fellow human" is a perfectly appropriate goal for an enterprise scale LLM, and we shouldn't be afraid of openly setting that as a goal.

Re: Gemini AI

#716
post #362

Earlier quoted context omitted.

Investors are getting impatient! ChatGPT has already replaced Google for me and I wonder if Google starts to feel the pressure.

> "ChatGPT has already replaced Google for me" Would you mind elaborating more on this. Like how are you "searching" with ChatGPT?

I got some unbelievably better results searching in bing + chatgtp the full page newspaper ad that Trump bought in the 80s on the NYT and other newspapers to shit on nato (or something similar). With google I got absolutely nothing even rephrasing the search in multiple ways, with bing + chatgtp the first link was a website with the scanned newspaper page with the ad. I think that google search dominance is pretty much gone. The results are full of SEOd to the death websites rather than anything useful.

Re: Gemini AI

#717

Earlier quoted context omitted.

He's not wrong. DeepMind spends time solving big scientific / large-scale problems such as those in genetics, material science or weather forecasting, and Google has untouchable resources such as all the books they've scanned (and already won court cases about) They do make OpenAI look like kids in that regard. There is far more to technology than public facing goods/products. It's probably in part due to the cultura…

> They do make OpenAI look like kids in that regard. Nokia and Blackberry had far more phone-making experience than Apple when the iPhone launched. But if you can't bring that experience to bear, allowing you to make a better product - then you don't have a better product.

The thing is that OpenAI doesn't have an "iPhone of AI" so far. That's not to say what will happen in the future - the advent of generative AI may become a big "equalizer" in the tech space - but no company seems to have a strong edge that'd make me more confident in any one of them over others.

Re: Gemini AI

#718
post #60

One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…

He's not wrong. DeepMind spends time solving big scientific / large-scale problems such as those in genetics, material science or weather forecasting, and Google has untouchable resources such as all the books they've scanned (and already won court cases about) They do make OpenAI look like kids in that regard. There is far more to technology than public facing goods/products. It's probably in part due to the cultura…

They do not make Openai look like kids. If anything, it looks like they spent more time, but achieved less. GPT-4 is still ahead of anything Google has released.

Re: Gemini AI

#719
post #60

One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…

He's not wrong. DeepMind spends time solving big scientific / large-scale problems such as those in genetics, material science or weather forecasting, and Google has untouchable resources such as all the books they've scanned (and already won court cases about) They do make OpenAI look like kids in that regard. There is far more to technology than public facing goods/products. It's probably in part due to the cultura…

I thought that Google was based out of Silcon Valley/California/USA

Re: Gemini AI

#720
post #60

One observation: Sundar's comments in the main video seem like he's trying to communicate "we've been doing this ai stuff since you (other AI companies) were little babies" - to me this comes off kind of badly, like it's trying too hard to emphasize how long they've been doing AI (which is a weird look when the currently publicly available SOTA model is made by OpenAI, not Google). A better look would simply be to sh…

That was pretty impressive… but do I have to be “that guy” and point out the error it made? It said rubber ducks float because they’re made of a material less dense than water — but that’s not true! Rubber is more dense than water. The ducky floats because it’s filled with air. If you fill it with water it’ll sink. Interestingly, ChatGPT 3.5 makes the same error, but GPT 4 nails it and explains the it’s the air that…

I would've liked to see an explanation that includes the weight of water being displaced. That would also explain how a steel ship with an open top is also able to float.
Post reply on HN