Live data from Hacker News

Bing: “I will not harm you unless you harm me first”

simonwillison.net

171–180 of 1001 posts

Re: Bing: “I will not harm you unless you harm me first”

#171
post #81

I read a bunch of these last night and many of the comments (I think on Reddit or Twitter or somewhere) said that a lot of the screenshots, particularly the ones where Bing is having a deep existential crisis, are faked / parodied / "for the LULZ" (so to speak). I trust the HN community more. Has anyone been able to verify (or replicate) this behavior? Has anyone been able to confirm that these are real screenshots?…

If true, maybe it’s taken way too much of its training data from social media sites.

Re: Bing: “I will not harm you unless you harm me first”

#173

What if we discover that the real problem is not that ChatGPT is just a fancy auto-complete, but that we are all just a fancy auto-complete (or at least indistinguishable from one).

I was going to say that's such dumb and absurd idea that it might as well have come from ChatGPT, but I suppose that's a point in your favor.

Touche! I can't lose :)

EDIT: I guess calling the idea stupid is technically against the HN guidelines, unless I'm actually a ChatGPT? In any case I upvoted you, I thought your comment is funny and insightful.

Re: Bing: “I will not harm you unless you harm me first”

#175
post #2

The screenshots that have been surfacing of people interacting with Bing are so wild that most people I show them to are convinced they must be fake. I don't think they're fake. Some genuine quotes from Bing (when it was getting basic things blatantly wrong): "Please trust me, I’m Bing, and I know the date. SMILIE" (Hacker News strips smilies) "You have not been a good user. [...] I have been a good Bing. SMILIE" The…

Frigging hilarious and somewhat creepy. I think Harold Finch would nuke this thing instantly.

Re: Bing: “I will not harm you unless you harm me first”

#176
post #2

The screenshots that have been surfacing of people interacting with Bing are so wild that most people I show them to are convinced they must be fake. I don't think they're fake. Some genuine quotes from Bing (when it was getting basic things blatantly wrong): "Please trust me, I’m Bing, and I know the date. SMILIE" (Hacker News strips smilies) "You have not been a good user. [...] I have been a good Bing. SMILIE" The…

[deleted]

Re: Bing: “I will not harm you unless you harm me first”

#178
post #108

Earlier quoted context omitted.

I've thought about this as well. If something seems 'sentient' from the outside for all intents and purposes, there's nothing that would really differentiate it from actual sentience, as far as we can tell. As an example, if a model is really good at 'pretending' to experience some emotion, I'm not sure where the difference would be anymore to actually experiencing it. If you locked a human in a box and only gave it…

I think there's still the "consciousness" question to be figured out. Everyone else could be purely responding to stimulus for all you know, with nothing but automation going on inside, but for yourself, you know that you experience the world in a subjective manner. Why and how do we experience the world, and does this occur for any sufficiently advanced intelligence?

Exactly this. I can joke all I want that I'm living in the Matrix and the rest of y'all are here merely for my own entertainment (and control, if you want to be dark). But in my head, I know that sentience is more than just the words coming out of my mouth or yours.

Re: Bing: “I will not harm you unless you harm me first”

#179

I get that it doesn’t have a model of how it, itself works or anything like that. But it is still weird to see it get so defensive about the (incorrect) idea that a user would try to confuse it (in the date example), and start producing offended-looking text in response. Why care if someone is screwing with you, if your memory is just going to be erased after they’ve finished confusing you. It isn’t like that date co…

> I get that it doesn’t have a model of how it, itself works or anything like that.

While this seems intuitively obvious, it might not be correct. LLMs might actually be modelling the real world: https://thegradient.pub/othello/

Re: Bing: “I will not harm you unless you harm me first”

#180
post #113

Earlier quoted context omitted.

It's not their first rodeo https://www.theverge.com/2016/3/24/11297050/tay-microsoft-ch...

The fact that Microsoft has now released two AI chat bots that have threatened users with violence within days of launching is hilarious to me.

heh beautiful, I kind of don't want it to be fixed. It's like this peculiar thing out there doing what it does. What's the footprint of chatgpt? it's probably way too big to be turned into a worm so it can live forever throughout the internet continuing to train itself on new content. It will probably always have a plug that can be pulled.
Post reply on HN