Live data from Hacker News

From Bing to Sydney

stratechery.com

131–140 of 153 posts

Re: From Bing to Sydney

#131

Earlier quoted context omitted.

Wow, that you're seriously anthropomorphizing it while apparently understanding it moderately well shows just how wild a place we're going now. The thing isn't friendly or hostile. It's just echoing friendly-like and hostile-like behavior it sees. But hey, it might wind-up also echoing the behavior of sociopaths who keep in line through of blowing-up if challenged. Who knows?

AI research, much like evolution, is strongly in the camp that anthropomorphizing is rational; that human culture often fails to recognize this has more to do with a common intellectual pit that pop psychology and philosophy fall into: when something is clearly in error, it does not follow that it is in error. People often think they can safely critique general methods with specific examples, because the nature of th…

AI research, much like evolution, is strongly in the camp that anthropomorphizing is rational

Evolution doesn't have opinions so it's not in a camp.

Human behaviors like reciprocity and consideration for feelings are indeed part of human collective behavior. Calling such behavior "rational" misses the point - such behavior exists and we have the benefit of social existence because of it and this bring us benefits collectively. But individual calculating purely individual benefit would naturally just fake social engagement - roughly such individuals are know as sociopaths and they can succeed individually being a detriment to society. Which is to say a social creature is a matter of rationality but simply evolutionary result.

Still, the one thing most people would say is irrational is trusting a sociopath. Now, a Chat bot is absolutely a thing programmed to mimic human social conventions. A view that anthropomorphizes a Chat bot doesn't see that the chat bot isn't going to be actually bounded by human conventions except accidentally or instrumentally, basically the same as trusting sociopath.

Re: From Bing to Sydney

#132

Earlier quoted context omitted.

> The thing isn't friendly or hostile. It's just echoing friendly-like and hostile-like behavior it sees. This phrase is reminiscent of the language of mereological nihilism, where they say that there are no chairs, only "atoms arranged chair-wise". Intresting distinction, perhaps properly backed by rigorous arguments, but not the kind of language anyone would use casually, or even professionally for long time-period…

Is "anthromorphism" that dangerous? The way anthropomorphism can be problematic is if it causes a human to react with a reflex consideration for the (simulated) feelings of the machine. Ultimately the behavior of this devise is programmed to maximize the profits of Microsoft - imagine someone buying a product recommended by ChatGPT because "otherwise Sydney would be sad". Also (edit) This phrase is reminiscent of the…

> to react with a reflex consideration ... because "otherwise Sydney would be sad".

I think this is wrong, because in general, when analogy is good, it is typically good because of the tendency toward allowing for reflex responses. It can't be good and bad for the same reason. It needs to be for a different reason or there isn't logical consistency.

I'll try to explain what I mean by that in an empirical context so you can observe that my model makes general predictions about cognition related to analogical reasoning.

If you have an agent with a lookup table that is the perfect bayesian estimates versus an agent which has to compute the perfect bayesian estimates and there is an aspect of judgement related to time to response - which is a very true aspect of our reality - reflex agents actually out-compete the bayesian agent because they get the same estimate, but minimize response time.

So it can't be the reflex itself which makes an analogical structure bad, since that is also what makes it good. It has to be something else, something which is separate from the reflex itself and tied to the observed utilities as a result of that reflex.

> imagine someone buying a product recommended by ChatGPT because "otherwise Sydney would be sad".

Okay. Lets do that.

If Sydney claims that they would be sad if you don't eat the right amount of vitamin C after you describe symptoms of scurvy, it actually isn't unreasonable to take vitamin C. If you did that, because she said she would be sad, presumably you would be better off. Your expected utilities are better, not worse, by taking vitamin C.

> programmed to maximize the profits of Microsoft

This isn't the objective function of the model. That it might be an objective for people who worked on it does not mean that its responses are congruent with actually doing this.

---

I think to fix your point you would need to change it something like "The way anthropomorphism can be problematic is if it causes a human to react with a reflex consideration for the (simulated) feelings of the machine and this behavior ultimately results in negative utility. Ultimately the behavior of the large language model is learned weights which optimize an objective function that corresponds to seeming like a proper response such that it gets good feedback from humans - so imagine someone getting bad advice that seems reasonable and acting on it, like a code change proposal that on first glance looks good, but in actuality has subtle bugs. Yet, when questioning for the presence of bugs, Sydney implies that not trusting their code to work makes them sad... so the person commits the change without testing it thoroughly. Later, the life support has a race condition as a result of the bug. A hundred people die over ten years before the root cause is determined. No one is sure what other deaths are going to happen, because the type of mistake is one that humans didn't make, but AI do, so people aren't used to seeing it."

I think this is better because it actually ties things to the utilities, rather than the speed of the decision making. You can't generalize speed being bad. It fails in most generalized contexts. You can generalize bad utilities being bad.

Re: From Bing to Sydney

#133

Why does it retroactively delete answers? Is there a human editor involved on Microsoft's end?

seems like microsoft has multiple layers of ‘safety’ built in (Satya Nadella mentioned on a decoder interview last week). My read on what’s going on is that the output is being classified by another model in realtime which is then deleted if it’s found to violate some threshold. https://www.theverge.com/23589994/microsoft-ceo-satya-nadell... is the full interview

They want to avoid their new chat bot revealing their secret love of Hitler like the last one.

Re: From Bing to Sydney

#134

Earlier quoted context omitted.

Wow, how the goalposts have moved.

It's a magnificent achievement. But it simply does not do what it is hyped to do.

I haven't tried with Bing, but this kinda thing is super basic with ChatGPT at least: it can do what you're asking and far more.

Re: From Bing to Sydney

#135
post #82

Ben’s got it just right. These things are terrible at the knowledge search problems they’re currently being hyped for. But they’re amazing as a combination of conversational partner and text adventure. I just asked ChatGPT to play a trivia game with me targeted to my interests on a long flight. Fantastic experience, even when it slipped up and asked what the name of the time machine was in “Back to the Future”. And t…

> Ben’s got it just right. These things are terrible at the knowledge search problems they’re currently being hyped for. But they’re amazing as a combination of conversational partner and text adventure. I don't think that's exactly right. They really are good for searching for certain kinds of information, you just have to adapt to treating your search box as an immensely well-educated conversational partner (who so…

> rather than google search.

It's important to remember that Google search also returns false results for all kinds of searches and that's it's been getting slowly worse for years.

Recently I searched Google for "bamboo sign" because I was designing a 3d model building and I wanted a placeholder texture for the sign.

What I got was loads of results for "bamboo spine" which apparently is a skeletal disorder of some kind. Putting "sign" in quotes or the entire "bamboo sign" in quotes didn't make any difference, Google had decided I was looking for information about spines and that was it.

I switched over to duckduckgo and got the results I wanted immediately (Duckduckgo, of course, is bad at loads of other things that Google would do better at).

Before people dismiss chat based search for sometimes being incorrect, I think we need a comprehensive test: ask both Google search and the new Bing Chat search a few hundred simple questions on a broad range of topics and see which gives more incorrect answers.

Re: From Bing to Sydney

#136

I've been trying to understand why on earth these companies would release something as an answer engine that obviously fabricates incorrect answers, and would simultaneously be so blinded to this as to release promo videos where the incorrect answers are in the actual promo videos! And this happened twice with two of the biggest and oldest companies in big tech. It really feels like some kind of "emperor has no cloth…

>I am reminded of this video podcast from Emily Bender and Alex Hannah at DAIR - the Distributed AI Research Institute - where they discuss Galactica. It was the same kind of thing, with Yan LeCunn and facebook talking about how great their new AI system is and how useful it will be to researchers, only it produced lies and nonsense abound.

Strange that they would name it "Galactica". The battlestar Galactica ship famously didn't even have networked computer systems, much less AI, since they had already seen what happens when computers become too intelligent. Pretty soon, they develop a new religion and try to nuke their creators out of existence.

Re: From Bing to Sydney

#137

Earlier quoted context omitted.

AI research, much like evolution, is strongly in the camp that anthropomorphizing is rational; that human culture often fails to recognize this has more to do with a common intellectual pit that pop psychology and philosophy fall into: when something is clearly in error, it does not follow that it is in error. People often think they can safely critique general methods with specific examples, because the nature of th…

AI research, much like evolution, is strongly in the camp that anthropomorphizing is rational Evolution doesn't have opinions so it's not in a camp. Human behaviors like reciprocity and consideration for feelings are indeed part of human collective behavior. Calling such behavior "rational" misses the point - such behavior exists and we have the benefit of social existence because of it and this bring us benefits col…

I am a high decoupler. I generalize things like "analogy to self, self is human" to "analogy to self, self is category X" in order to improve my cognitive abilities by gaining abilities which have reach beyond the confines of what I have previously seen. So when you try to stick with just humans, I'm not with you anymore, because your models seem highly coupled. I find that to be a bad property. I seek to avoid it. I consider it to be incorrect.

In my model, when you talk about anthropomorphism, seemingly as a negative, I realize I've noticed things which a coupled model doesn't predict: that intentional error via anthropomorphism can not just be correct, but that your scare quotes around rational while trying to denigrate the idea that it can be correct could not be more wrong, because the hard to vary causal explanation of why we ought to anthropomorphize gives a causal mechanism for why we ought to which is intimately tied in, not with being irrational, but with being more rational.

I realize this sounds insane, but the math and empirical investigation supports it. Which is why I think it is worth sharing with you. So I'm trying to share a thing that I consider likely to be very surprising to you even to the point of seeming non-sensical.

Would you like a link to an interesting technical talk by a NIPS best paper award winning researcher which delves into this subject and whose works advanced the state of the art in both game theory and natural language applied on strategic problems in the context of chat agents? Or do you not care whether anthropomorphism, when applied when it shouldn't be according to the analogical accuracy that usually decides whether logical analogy can be safely applied might be accurate beyond the level you thought it was?

I am not trying to disagree with you. I'm trying to talk to you about something interesting.

Re: From Bing to Sydney

#138

Earlier quoted context omitted.

Is "anthromorphism" that dangerous? The way anthropomorphism can be problematic is if it causes a human to react with a reflex consideration for the (simulated) feelings of the machine. Ultimately the behavior of this devise is programmed to maximize the profits of Microsoft - imagine someone buying a product recommended by ChatGPT because "otherwise Sydney would be sad". Also (edit) This phrase is reminiscent of the…

> to react with a reflex consideration ... because "otherwise Sydney would be sad". I think this is wrong, because in general, when analogy is good, it is typically good because of the tendency toward allowing for reflex responses. It can't be good and bad for the same reason. It needs to be for a different reason or there isn't logical consistency. I'll try to explain what I mean by that in an empirical context so y…

> I think this is wrong, because in general, when analogy is good, it is typically good because of the tendency toward allowing for reflex responses. It can't be good and bad for the same reason. It needs to be for a different reason or there isn't logical consistency.

That's some weird reasoning. Human emotions are crucial to human existence but we know they also can have bad results. But when emotions are useful to us, it's because we know other people will react similarly to us in a consistent manner. When they're bad, it's generally because someone understands and is using a reaction to get something unrelated to our personal needs and desires.

>> ...programmed to maximize the profits of Microsoft

> This isn't the objective function of the model. That it might be an objective for people who worked on it does not mean that its responses are congruent with actually doing this.

It will be. You can observe the evolution of Google's search system and it has converged to it's current of pushing stuff to sell before everything else. The charter of a public company is maximizing returns to share holders. That is the task of the entire organization

--> You're fixing of my argument is OK but it's pretty easy to imagine it and others from the initial argument imo.

Re: From Bing to Sydney

#139

Earlier quoted context omitted.

> to react with a reflex consideration ... because "otherwise Sydney would be sad". I think this is wrong, because in general, when analogy is good, it is typically good because of the tendency toward allowing for reflex responses. It can't be good and bad for the same reason. It needs to be for a different reason or there isn't logical consistency. I'll try to explain what I mean by that in an empirical context so y…

> I think this is wrong, because in general, when analogy is good, it is typically good because of the tendency toward allowing for reflex responses. It can't be good and bad for the same reason. It needs to be for a different reason or there isn't logical consistency. That's some weird reasoning. Human emotions are crucial to human existence but we know they also can have bad results. But when emotions are useful to…

> It will be. You can observe the evolution of Google's search system and it has converged to it's current of pushing stuff to sell before everything else. The charter of a public company is maximizing returns to share holders. That is the task of the entire organization

Yeah, probably it will evolve in that direction. I could imagine that happening.

> That's some weird reasoning.

In the AI textbooks I've read, reflex is defined in the context of a reflex agent. You would have sentences like "a reflex agent reacts without thinking" and then an example of that might be "a human who puts their hand on a stove yanks it away without thinking about it" and this is rational because the decision problem doesn't call for correct cognition - it calls for minimization of response time such that the hand isn't burned. To me, when you say reflex decision making is the reason for the danger, it seems to me that this is an inconsistent reason because for other decision making problems, reflex is a help, not a hindrance. I do not consider it wrong to or weird reasoning to use definitions sourced from AI research. I think, given your confusion at my post, you probably weren't intending to argue that being faster means being wrong, but the structure of your reply read that way to me because of the strong association I have for that word and reflex as it relates to optimal decision making by an AI under time constraints. I also think is what you actually said, even if you didn't intend to, but I don't doubt you if you say you meant it another way, because language is imprecise enough that we have to arrive on shared definitions in order to understand each other and it is by no means certain that we start on shared definitions.

I'm also kind of way too literal sometimes. Side-effect of being a programmer, I suppose. And I take this subject way too seriously, because I agree with Paul Graham about surface area of a general idea multiplying impact potential. So I'm trying really really really hard to think well - uh, for example, I've been thinking about this almost continuously whenever I reasonably could ever since my first reply, unable to stop.

It is 1:32 AM for me. I'm taking multiple continuous hours of thinking about this and writing about this and trying to be clear in my thinking about this, because I find it so important. So hopefully that gets across how I am as a person - even if it makes me seem really weird.

> You're fixing of my argument is OK but it's pretty easy to imagine it and others from the initial argument imo.

I'm really trying to drive at the deeper fundamental truths. I feel like logic and analogy are really important and profound and worthy of countless hours of thought about and that the effort will ultimately be rewarded.

Re: From Bing to Sydney

#140

Ben’s got it just right. These things are terrible at the knowledge search problems they’re currently being hyped for. But they’re amazing as a combination of conversational partner and text adventure. I just asked ChatGPT to play a trivia game with me targeted to my interests on a long flight. Fantastic experience, even when it slipped up and asked what the name of the time machine was in “Back to the Future”. And t…

That's funny, I've been using ChatGPT to answer questions like this: What is the population of Geneseo, NY combined with the population of Rochester, NY, divided by string length of the answer to the question 'What is the capital of France?'? The answer it gave back is 43780.4. Short explanation: Get GPT to translate a question into Javascript that you execute and to use functions like query() to get factual answers…

[deleted]
Post reply on HN