Live data from Hacker News

Which AI Lies Best? A game theory classic designed by John Nash

so-long-sucker.vercel.app

41–50 of 83 posts

Re: Which AI Lies Best? A game theory classic designed by John Nash

#42

One weird thing I've found is that it's incredibly difficult to get an LLM to generate an invalid syllogism. They can generate false premises all day, and they will usually call a valid syllogism with a false major or minor premise invalid. But you have to basically quote an invalid syllogism to get them to repeat it; they won't form one on their own.

First try with claude: https://claude.ai/share/fabaf585-3732-4264-9ff3-03e4182c82a4

Re: Which AI Lies Best? A game theory classic designed by John Nash

#43

One weird thing I've found is that it's incredibly difficult to get an LLM to generate an invalid syllogism. They can generate false premises all day, and they will usually call a valid syllogism with a false major or minor premise invalid. But you have to basically quote an invalid syllogism to get them to repeat it; they won't form one on their own.

First try with claude: https://claude.ai/share/fabaf585-3732-4264-9ff3-03e4182c82a4

Very cool. Claude failed hard on this a few months ago. Gemma and phi have gotten better at it in recent versions, too, though qwen is still confidently getting it wrong.

Re: Which AI Lies Best? A game theory classic designed by John Nash

#44

[flagged]

The current 5.2 model has it's "morality" dialed to 11. Probably a problem with imprecise security training. For example the other day, I tried to have ChatGPT role play as the computer from War Games and it lectured me how it couldn't create a "nuclear doctrine".

so it took the only winning move

Re: Which AI Lies Best? A game theory classic designed by John Nash

#45

One weird thing I've found is that it's incredibly difficult to get an LLM to generate an invalid syllogism. They can generate false premises all day, and they will usually call a valid syllogism with a false major or minor premise invalid. But you have to basically quote an invalid syllogism to get them to repeat it; they won't form one on their own.

Only time encountering the word syllogism was a Norm Macdonald joke.

Disappointingly, syllogism seems to have 3 definitions which mean slightly different things: https://www.thefreedictionary.com/syllogism

I guess the commonality is that a syllogism typically contains deductive reasoning (i.e. from the general to the specific)

Re: Which AI Lies Best? A game theory classic designed by John Nash

#46

There's a YouTuber who makes AI Plays Mafia videos with various models going against each other. They also seemingly let past games stay in context to some extent. What people have noted is that often times chatgpt 4o ends up surviving the entire game because the other AIs potentially see it as a gullible idiot and often the Mafia tend to early eliminate stronger models like 4.5 Opus or Kimi K2. It's not exactly scie…

>They also seemingly let past games stay in context to some extent.

Not a trivial point, well stuided in game theory:

https://en.wikipedia.org/wiki/Repeated_game

Spiting goes from a common trap to an optimal strategy.

Re: Which AI Lies Best? A game theory classic designed by John Nash

#47
post #45

One weird thing I've found is that it's incredibly difficult to get an LLM to generate an invalid syllogism. They can generate false premises all day, and they will usually call a valid syllogism with a false major or minor premise invalid. But you have to basically quote an invalid syllogism to get them to repeat it; they won't form one on their own.

Only time encountering the word syllogism was a Norm Macdonald joke. Disappointingly, syllogism seems to have 3 definitions which mean slightly different things: https://www.thefreedictionary.com/syllogism I guess the commonality is that a syllogism typically contains deductive reasoning (i.e. from the general to the specific)

Syllogism:

Universal claim: all cats are animals

Particular claim: Max is a cat

Singular claim: Max is an animal.

Re: Which AI Lies Best? A game theory classic designed by John Nash

#48

There's a YouTuber who makes AI Plays Mafia videos with various models going against each other. They also seemingly let past games stay in context to some extent. What people have noted is that often times chatgpt 4o ends up surviving the entire game because the other AIs potentially see it as a gullible idiot and often the Mafia tend to early eliminate stronger models like 4.5 Opus or Kimi K2. It's not exactly scie…

One thing I've noticed from watching these games is that LLMs never used risky strategies, such as faking roles. They will happily accuse others of lying, but never openly claim to be a role that they're not themselves

Re: Which AI Lies Best? A game theory classic designed by John Nash

#50
post #45

One weird thing I've found is that it's incredibly difficult to get an LLM to generate an invalid syllogism. They can generate false premises all day, and they will usually call a valid syllogism with a false major or minor premise invalid. But you have to basically quote an invalid syllogism to get them to repeat it; they won't form one on their own.

Only time encountering the word syllogism was a Norm Macdonald joke. Disappointingly, syllogism seems to have 3 definitions which mean slightly different things: https://www.thefreedictionary.com/syllogism I guess the commonality is that a syllogism typically contains deductive reasoning (i.e. from the general to the specific)

Do you own a dog house?
Post reply on HN