Live data from Hacker News

Bard is much worse at puzzle solving than ChatGPT

twofergoofer.com

61–70 of 86 posts

Re: Bard is much worse at puzzle solving than ChatGPT

#61

How do you navigate this blog to read the other articles? I couldn't find any way to read the one on gpt4 (clicking the underlined "wrote about" does nothing) and twofergoofer.com/blog goes to a 404.

Hah - we only made the blog over the weekend and don't have any nav or menu for now. But yep we link the prior article a few times in this article, that article goes into more detail!

Re: Bard is much worse at puzzle solving than ChatGPT

#62
post #8
post #6

Earlier quoted context omitted.

s/Microsoft/Google/ http://www.paulgraham.com/microsoft.html

But was Microsoft ever cool in the sense that Google was circa 2005? (Asking honestly, I'm too young to remember that period.)

I know that at that time I used to be like "what kind of pleb uses anything other than Windows? Get off your high horse and use the OS that actually works", because well... Windows XP was pretty good, and games worked on it.

Re: Bard is much worse at puzzle solving than ChatGPT

#63
post #15

"Bard is much worse than ChatGPT at solving an obscure word game I invented" would have been a more honest title, but would probably generate less clicks for the author. Bard may still be much worse than ChatGPT at solving all kinds of puzzles, but the article is click bait for promoting the author's word game, not an actual investigation that warrants that conclusion.

Having read through the word game, I agree with others that it's good that the game is less likely to be in the corpus. I think rhyming, while a challenging task, may be a poor benchmark for ability. The author doesn't seem to understand rhyming too well (cactus practice is a weak rhyme at best) I completely disagree with the "hasty rhyming test" - Skeleton and Gelatin don't rhyme (-ton vs -tin), and rhyme worse than…

In the article, I mention that Twofer Goofer requires perfect or strict rhyme. Perfect and strict rhyme require that all syllables are pronounced identically in the speaker's tongue (for me, American Midwest accent), except for the first sound of the word which can vary.

Hence pooh-teen and proh-tein do not rhyme. Skell-ih-tin and Gell-ih-tin do rhyme.

A game like this requires a pretty tight rhyming definition to not annoy players in a given day!

Thanks for reading: https://www.masterclass.com/articles/perfect-vs-imperfect-rh...

Re: Bard is much worse at puzzle solving than ChatGPT

#64
post #20

Earlier quoted context omitted.

what's wrong with that? the use of novel puzzles is frankly awesome because there's a much lower chance of contamination from previous puzzles so we get a chance to see how much generalization they've achieved.

I'm complaining about the title writing a check that the blog post can't cash.

Fair! But if I wrote Twofer Goofer in the title it would not resonate at at all. Alas, tradeoffs.

Re: Bard is much worse at puzzle solving than ChatGPT

#65

Wow I had hoped for a more productive discussion than these 1-1 comparisons of Bard vs ChatGPT that I'm seeing everywhere. The model deployed with this version of Bard is clearly a smaller model than the biggest LaMDA/PaLM models Google has been working on for ages. Which, according to their publications, show unprecedented results on _proof writing_ of all things (see Minerva). While their strategic decisions may be…

They knew the war they were entering, they knew their enemies, they knew how they'd get evaluated and still decided to get this model out in its current state, leading to the conclusion: Yes, this is really the best they can do and it's much worse than the state of the art.

In any case, it's a massive marketing blunder, the public opinion formed within the last hours was overwhelmingly "Bard sucks compared to ChatGPT."

Re: Bard is much worse at puzzle solving than ChatGPT

#66
post #20

Earlier quoted context omitted.

I'm complaining about the title writing a check that the blog post can't cash.

GPT-4 says: A more accurate and balanced title might be: "Comparing Bard and ChatGPT in Puzzle Solving: An Examination within the Context of a Word Game"

Sounds like GPT-4 will save us from clickbait

Re: Bard is much worse at puzzle solving than ChatGPT

#67
post #58

> Twofer Goofer HQ's adherence to strict "perfect" rhyme can be tricky for those slant rhyme-inclined. And yet one puzzle they hammer Bard for failing is "Cactus Practice". What accent do you have to have for that to be a perfect rhyme?

From Chicago ... and with more than a thousand solves on that puzzle we've never received a single complaint about the rhyme on that one (plus the stats say users find it an extremely easy rhyme to solve)! Curious how you pronounce that one such that they don't rhyme?

According to the dictionary (and matching my own non-native pronunciation), cactus is /ˈkæktʌs/, practice is /ˈpɹæktɪs/

The terminal /tʌs/ is not quite the same thing as /tɪs/; since they are both unstressed, the difference can be hard to notice in fast speech, but becomes clear when enunciating.

Re: Bard is much worse at puzzle solving than ChatGPT

#68

> Twofer Goofer HQ's adherence to strict "perfect" rhyme can be tricky for those slant rhyme-inclined. And yet one puzzle they hammer Bard for failing is "Cactus Practice". What accent do you have to have for that to be a perfect rhyme?

In standard Australian English, this is a perfect rhyme.

Re: Bard is much worse at puzzle solving than ChatGPT

#69
post #58

Earlier quoted context omitted.

From Chicago ... and with more than a thousand solves on that puzzle we've never received a single complaint about the rhyme on that one (plus the stats say users find it an extremely easy rhyme to solve)! Curious how you pronounce that one such that they don't rhyme?

According to the dictionary (and matching my own non-native pronunciation), cactus is /ˈkæktʌs/, practice is /ˈpɹæktɪs/ The terminal /tʌs/ is not quite the same thing as /tɪs/; since they are both unstressed, the difference can be hard to notice in fast speech, but becomes clear when enunciating.

Dictionaries often neglect common mergers like the weak vowel merger https://en.wikipedia.org/wiki/Phonological_history_of_Englis... where unstressed /ʌ/, /ə/ and /ɪ/ all end up being pronounced the same.
Post reply on HN