Live data from Hacker News

Bard is much worse at puzzle solving than ChatGPT

twofergoofer.com

31–40 of 86 posts

Re: Bard is much worse at puzzle solving than ChatGPT

#31

Wow I had hoped for a more productive discussion than these 1-1 comparisons of Bard vs ChatGPT that I'm seeing everywhere. The model deployed with this version of Bard is clearly a smaller model than the biggest LaMDA/PaLM models Google has been working on for ages. Which, according to their publications, show unprecedented results on _proof writing_ of all things (see Minerva). While their strategic decisions may be…

Even the best Google models seem to be lagging for reasoning tasks vs OpenAI ones at the moment - see the graphs at https://github.com/suzgunmirac/BIG-Bench-Hard

Re: Bard is much worse at puzzle solving than ChatGPT

#32
post #27
post #19

Earlier quoted context omitted.

I know little more than what I've read about LLMs and other models. I just got it to do them again and it output them rather quickly: https://imgur.com/a/jfviCCe

are the numbers correct?

I spot checked an random assortment of them and so far yes, but it only takes one wrong one to set every subsequent number off.

Re: Bard is much worse at puzzle solving than ChatGPT

#33
Am I missing something? Most of TFA is about Bard failing to answer with rhyming words, but in the only prompts shown the author doesn't actually ask for rhyming words. He just says the hint and the name of the puzzle.

Is this not simply: "Bard is worse than ChatGPT at having seen the 'how-to-play' page for my side project during its training"?

Re: Bard is much worse at puzzle solving than ChatGPT

#35

Wow I had hoped for a more productive discussion than these 1-1 comparisons of Bard vs ChatGPT that I'm seeing everywhere. The model deployed with this version of Bard is clearly a smaller model than the biggest LaMDA/PaLM models Google has been working on for ages. Which, according to their publications, show unprecedented results on _proof writing_ of all things (see Minerva). While their strategic decisions may be…

It seems like they don't want to be the best, just good and cheap enough that they don't lose users therefore ad revenue.

They behave like Yahoo when Google took over.

Re: Bard is much worse at puzzle solving than ChatGPT

#37
post #6

It's so sad to me to see the downfall of google from the absolute coolest company on the planet to the one that's now trying to keep up.

s/Microsoft/Google/ http://www.paulgraham.com/microsoft.html

s/IBM/Microsoft/

History rhymes with itself.

Re: Bard is much worse at puzzle solving than ChatGPT

#38
post #32
post #27

Earlier quoted context omitted.

are the numbers correct?

I spot checked an random assortment of them and so far yes, but it only takes one wrong one to set every subsequent number off.

> it only takes one wrong one to set every subsequent number off.

Except its not actually doing that calculation, so one wrong one shouldn't truly affect the rest like "Real" math.

Re: Bard is much worse at puzzle solving than ChatGPT

#39
post #6

It's so sad to me to see the downfall of google from the absolute coolest company on the planet to the one that's now trying to keep up.

s/Microsoft/Google/ http://www.paulgraham.com/microsoft.html

> Microsoft

Good joke. MS has always been in the 'incompetent evil' quadrant, Newcomers just keep inexplicably giving them the benefit of the doubt or assuming/insisting they've "changed".

Re: Bard is much worse at puzzle solving than ChatGPT

#40
post #33

Am I missing something? Most of TFA is about Bard failing to answer with rhyming words, but in the only prompts shown the author doesn't actually ask for rhyming words. He just says the hint and the name of the puzzle. Is this not simply: "Bard is worse than ChatGPT at having seen the 'how-to-play' page for my side project during its training"?

Clicking through to the link next to 'last week's text' and then to 'full rules', it looks like the author is starting the chat sessions with a full explanation that isn't included in the screenshots. (Also, the last screenshot shows the author explicitly asking about rhymes.)
Post reply on HN