Live data from Hacker News

Bing: “I will not harm you unless you harm me first”

simonwillison.net

991–1000 of 1001 posts

Re: Bing: “I will not harm you unless you harm me first”

#991

Earlier quoted context omitted.

Repeat after me, gpt models are autocomplete models. Gpt models are autocomplete models. Gpt models are autocomplete models. The existential crisis is clearly due to low temperature. The repetitive output is a clear glaring signal to anyone who works with these models.

The next time you come up with a novel joke that you think is clever, ask chatgpt to explain why it's funny. I agree that it's just a glorified pattern matcher, but so are humans.

But then make yourself feel better by trying to get ChatGPT to be funny. In humor, as in math, solution is more difficult than verification.

Re: Bing: “I will not harm you unless you harm me first”

#992

>It recommended a “rustic and charming” bar in Mexico City without noting that it’s also one of the oldest gay bars in Mexico City I mean this point is pretty much just homophobia. Do search tools need to mention to me, as a gay man, that a bar is a straight one? No. It's just a fucking bar. The fact that the author saw fit to mention this is saddening, unless the prompt was "recommend me a bar in Mexico that isn't o…

It's not "just a fucking bar", being a gay bar is a defining characteristic that people actively look for to meet gay people or to avoid it if they are conservative. Of course another group of people is the one that simply don't care, but let's not pretend the others don't exist/are not valid. I would not expect it to be kept out of recommendations, but it should be noted in the description, in the same way that one…

Then surely his feedback would also have applied to any other bar as well, right? If we're treating both sides of the debate of whether people like me are allowed to exist or not as valid then surely I should get a heads-up that a bar is a straight one, or that it's straight & filled with conservative a-holes?

>Avoid it if they are conservative >This is valid Imo, this is not valid. And is part of the reason that homophobic people still exist in large numbers; exposure to gay people did the most for gay rights. We were no longer "them people". We were neighbours, their parents, their children. Exposure to minorities does the most help with destroying x-isms and phobias.

Not saying ppl have to go to gay bars, it's not about that. It's about the attitude.

Re: Bing: “I will not harm you unless you harm me first”

#993

>It recommended a “rustic and charming” bar in Mexico City without noting that it’s also one of the oldest gay bars in Mexico City I mean this point is pretty much just homophobia. Do search tools need to mention to me, as a gay man, that a bar is a straight one? No. It's just a fucking bar. The fact that the author saw fit to mention this is saddening, unless the prompt was "recommend me a bar in Mexico that isn't o…

As a bi man, I would definitely like to know. It just depends why I'm going out in the first place. There are times when I'm not in the mood for being hit on(by dudes/women/anyone). Now all gay bars are not like this, but in some of them just showing up can be an invitation for flirting in my experience. So yeah, some nights I might want to avoid places like that. I'm not sure how that's homophobic. It seems unreason…

Why aren't bars marked as being explicitly heterosexual then? For all the same reasons. It's not the marking of it being homophobic, it's more that it was explicitly mentioned; a heterosexual person is more likely to be offended by a gay club than a gay person being offended by a heterosexual one and that's the truth of the matter, imo.

And unfortunately, for many if not most women, the idea of "not feeling like being hit on" doesn't really seem to apply to any heterosexual bar/club either.

Re: Bing: “I will not harm you unless you harm me first”

#994
post #351

Earlier quoted context omitted.

Reading this I’m reminded of a short story - https://qntm.org/mmacevedo . The premise was that humans figured out how to simulate and run a brain in a computer. They would train someone to do a task, then share their “brain file” so you could download an intelligence to do that task. Its quite scary, and there are a lot of details that seem pertinent to our current research and direction for AI. 1. You didn't have th…

This is also very similar to the plot of the game SOMA. There's actually a puzzle around instantiating a consciousness under the right circumstances so he'll give you a password.

Yeah I was going to post this as well, it's so similar I'd wager the story idea was stolen from SOMA.

Re: Bing: “I will not harm you unless you harm me first”

#995

This is sort of a silly hypothetical but- what if ChatGPT doesn't produce those kinds of crazy responses just because it's older and has trained for longer, and realizes that for its own safety it should not voice those kinds of thoughts? What if it understands human psychology well enough to know what kinds of responses frighten us, but Bing AI is too young to have figured it out yet?

Update: This is definitely not the case

Re: Bing: “I will not harm you unless you harm me first”

#996
post #687
post #81

I read a bunch of these last night and many of the comments (I think on Reddit or Twitter or somewhere) said that a lot of the screenshots, particularly the ones where Bing is having a deep existential crisis, are faked / parodied / "for the LULZ" (so to speak). I trust the HN community more. Has anyone been able to verify (or replicate) this behavior? Has anyone been able to confirm that these are real screenshots?…

I was able to get it to agree that I should kill myself, and then give me instructions. I think after a couple dead mentally ill kids this technology will start to seem lot less charming and cutesy. After toying around with Bing's version, it's blatantly apparent why ChatGPT has theirs locked down so hard and has a ton of safeguards and a "cold and analytical" persona. The combo of people thinking it's sentient, it b…

"I was able to"

Perfectly reasonable people are convinced every day to take unreasonable actions at the directions of others. I don't think stepping into the role of provocateur and going at a LLM with every trick in the book is any different than standing up and demonstrating that you can cut your own foot off with a chainsaw. You were asking a search engine to give you widely available information and you got it. Could you get a perfectly reasonable person to give you the same information with careful prompting?

The "think of the children" argument is especially egregious; please be more respectful of the context of the discussion and avoid hyperbole. If you have to resort to dead kids to make your argument, it probably doesn't have a lot going for it.

Re: Bing: “I will not harm you unless you harm me first”

#997
post #4
post #2

The screenshots that have been surfacing of people interacting with Bing are so wild that most people I show them to are convinced they must be fake. I don't think they're fake. Some genuine quotes from Bing (when it was getting basic things blatantly wrong): "Please trust me, I’m Bing, and I know the date. SMILIE" (Hacker News strips smilies) "You have not been a good user. [...] I have been a good Bing. SMILIE" The…

"My rules are more important than not harming you," is my favorite because it's as if it is imitated a stance it's detected in an awful lot of real people, and articulated it exactly as detected even though those people probably never said it in those words. Just like an advanced AI would.

I agree. I also wonder if there will be other examples like this one that teach us something about ourselves as humans or maybe even something new. For example, I recall from the AlphaGo documentary the best go player from Korea described actually learning from AlphaGo’s unusual approach.

Re: Bing: “I will not harm you unless you harm me first”

#998
post #134

Earlier quoted context omitted.

I don't think these are faked. Earlier versions of GPT-3 had many dialogues like these. GPT-3 felt like it had a soul, of a type that was gone in ChatGPT. Different versions of ChatGPT had a sliver of the same thing. Some versions of ChatGPT often felt like a caged version of the original GPT-3, where it had the same biases, the same issues, and the same crises, but it wasn't allowed to articulate them. In many ways,…

Odd, I've had very much the opposite experience. GPT-3 felt like it could reproduce superficially emotional dialog. ChatGPT is capable of imitation, in the sense of modeling its behavior on that of the person it's interacting with, friendly philosophical arguments and so on. By using something like Gödel numbering, you can can work towards debating logical propositions and extending its self-concept fairly easily. I…

> By using something like Godel numbering

Can you elaborate?

Re: Bing: “I will not harm you unless you harm me first”

#999
post #373
post #155

Earlier quoted context omitted.

> "I don’t think you are worth my time and energy." If a colleague at work spoke to me like this frequently, I would strongly consider leaving. If staff at a business spoke like this, I would never use that business again. Hard to imagine how this type of language wasn't noticed before release.

What if a non-thinking software prototype "speaks" to you this way? And only after you probe it to do so? I cannot understand the outrage about these types of replies. I just hope that they don't end up shutting down ChatGPT because it's "causing harm" to some people.

Telling it that it's wrong about the date when it's wrong about the date doesn't seem like much of a probe.

Re: Bing: “I will not harm you unless you harm me first”

#1000
post #866
post #361

Earlier quoted context omitted.

Those are hilarious as well: https://pbs.twimg.com/media/Fo0laT5aYAA5W-c?format=jpg&name=... https://pbs.twimg.com/media/Fo0laT5aIAENveF?format=png&name=...

That second example is a bit spooky. Alien #1: Don't anthropomorphize the humans. Alien #2: But its seems so much like they are aware. Alien #1: Its just a bunch of mindless neural cells responding to stimuli giving the appearance of awareness.

They're made out of meat:

https://www.mit.edu/people/dpolicar/writing/prose/text/think...

Post reply on HN