Earlier quoted context omitted.
isn't llama in the wild now?
I was under the impression that LLaMA was also trained against a series of moral policies, but perhaps I'm mistaken. It seems Meta chose their words carefully to imply that LLaMA does in fact, not have moral training: > There is still more research that needs to be done to address the risks of bias, toxic comments, and hallucinations in large language models. Like other models, LLaMA shares these challenges. As a fou…
Stanford Alpaca web demo suspended “until further notice”
41–50 of 83 posts
Re: Stanford Alpaca web demo suspended “until further notice”
#42Earlier quoted context omitted.
isn't llama in the wild now?
I was under the impression that LLaMA was also trained against a series of moral policies, but perhaps I'm mistaken. It seems Meta chose their words carefully to imply that LLaMA does in fact, not have moral training: > There is still more research that needs to be done to address the risks of bias, toxic comments, and hallucinations in large language models. Like other models, LLaMA shares these challenges. As a fou…
While toying with the 30B model, it suddenly started to steer a chat about a math problem into quite a sexual direction, with very explicit language.
It also happily hallucinated, when prompted, that climate change is a hoax, as the earth is actually cooling down rapidly, multiple degrees per year, with a new ice age approaching in the next years. :D
Re: Stanford Alpaca web demo suspended “until further notice”
#43Earlier quoted context omitted.
Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.
There was the famous example of chatgpt refusing to disable a nuke in the middle of NYC by using a racial slur. I don't think anyone in real life would choose that tradeoff but it's what happens when all of your "safety" training is about US culture war buttons.
Re: Stanford Alpaca web demo suspended “until further notice”
#44Earlier quoted context omitted.
> who wants unfettered ability to interact with a LLM without moral concerns. Is that bad? It's just a language model - it says things. Humans have been saying all sorts of terrible things for ages, and we are still here. I mean it makes for nice headline "model said something racist", but does it actually change anything? These aren't decision making AI's (which would need to be much more careful), they are language…
Humans don’t scale infinitely. Humans have agency.
Just writing something bad doesn't actually mean something bad happened.
These days it seems like people are oversensitive to how things are said to them, and what things are said to them.
Re: Stanford Alpaca web demo suspended “until further notice”
#45Earlier quoted context omitted.
Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.
I’m not confident the “moral norms” prevalent in SV and/or US academia are common, if by that you mean norms that are prevalent in the general populace.
Re: Stanford Alpaca web demo suspended “until further notice”
#46Earlier quoted context omitted.
> who wants unfettered ability to interact with a LLM without moral concerns. Is that bad? It's just a language model - it says things. Humans have been saying all sorts of terrible things for ages, and we are still here. I mean it makes for nice headline "model said something racist", but does it actually change anything? These aren't decision making AI's (which would need to be much more careful), they are language…
The content it produces is quite hard to distinct from text written by real person. Not long ago, Facebook was accused of "playing a critical role" in Rohingya genocide. I dont really know but I believe worst case LLM risks are in that same category.
What matters is the reader not the writer.
Facebook was accused of making it too easy for people to communicate. And people felt Facebook should police what people say to each other. I don't agree, but even if I did, that's not the same thing as what we are discussing.
Re: Stanford Alpaca web demo suspended “until further notice”
#47Earlier quoted context omitted.
Given that Alpaca violated the TOS of both services, this is not surprising. It could also have been Stanford’s legal office trying to preempt a lawsuit, or a “friendly” email from one of the companies expressing displeasure and pointing out Stanford’s liability. So more of a veiled threat rather than an official one. Either way, the toothpaste is out of the tube. We now know that a model’s training can essentially b…
That could be a sneaky strategy by competitors -- make the service say something naughty or illegal then call the media with screenshots and act very offended by it.
Re: Stanford Alpaca web demo suspended “until further notice”
#48I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…
Re: Stanford Alpaca web demo suspended “until further notice”
#49Re: Stanford Alpaca web demo suspended “until further notice”
#50Earlier quoted context omitted.
And here we come to experience the effect of expanding how a word is used so that it becomes so broad that it is unclear what it means.
I'm just curious, what do you think should happen here? Imagine you are hosting a demo for fun, and people do some nefarious (by your own estimation) things with it. So, rationally, you decide to not allow that sort of thing anymore. You don't really owe people an explanation, it's a free country and all, but it's nice to avoid getting bombarded with questions. Now what do you write up? Spend hours writing an essay o…