I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…
> who wants unfettered ability to interact with a LLM without moral concerns. Is that bad? It's just a language model - it says things. Humans have been saying all sorts of terrible things for ages, and we are still here. I mean it makes for nice headline "model said something racist", but does it actually change anything? These aren't decision making AI's (which would need to be much more careful), they are language…
Stanford Alpaca web demo suspended “until further notice”
51–60 of 83 posts
Re: Stanford Alpaca web demo suspended “until further notice”
#52I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…
> if these models are trained on essentially raw WWW sourced data, I would personally want raw output. Llama is a very high-quality foundation LLM, you can already run it very easily using llama.cpp and will get the raw output you need. https://github.com/ggerganov/llama.cpp There's already instructions on how anyone can fine-tune it to behave similarly to ChatGPT for as little as $100: https://crfm.stanford.edu/2023…
If nothing else, I continue to be amazed and how uninteroperable certain technologies are.
I had to remove glibc and gcc to get llama to compile on my intel macbook. Masking/hiding them from my environment didn’t work, as it went out and found them and their header files instead of clang.
Which eventually worked fine.
Re: Stanford Alpaca web demo suspended “until further notice”
#53Earlier quoted context omitted.
> who wants unfettered ability to interact with a LLM without moral concerns. Is that bad? It's just a language model - it says things. Humans have been saying all sorts of terrible things for ages, and we are still here. I mean it makes for nice headline "model said something racist", but does it actually change anything? These aren't decision making AI's (which would need to be much more careful), they are language…
The content it produces is quite hard to distinct from text written by real person. Not long ago, Facebook was accused of "playing a critical role" in Rohingya genocide. I dont really know but I believe worst case LLM risks are in that same category.
Rohingya is a textbook example of blind optimisation and lack of context awareness. FB looked at a region, and people communicating in language they didn't understand. But they did see that certain symbols and/or combinations of symbols got a lot of engagement. If you're after money, you want to amplify the use of those symbols and hopefully generate lots more similar content.
Turns out that's a morally reprehensible thing when the people using those symbols were advocating genocide. (It was good for the revenue while it lasted, though.)
With LLMs and their hardcoded guard rails, I suspect we're going to see the danger emerge from the other side. Instead of actively spewing hatred, they will be used for mass sock-puppetry and opinion amplification on a massive scale. Think simple sabotage field manual for 21st century, but weaponised thousand-fold.
Re: Stanford Alpaca web demo suspended “until further notice”
#54Re: Stanford Alpaca web demo suspended “until further notice”
#55Earlier quoted context omitted.
Example of "moral policy" in practice: Midjourney appears to be banning making fun of the Chinese dictator for life because it's supposedly racist or something. With that kind of moral compass, I’m not sure I'd be missing its absence.
> Example of "moral policy" in practice: Midjourney appears to be banning making fun of the Chinese dictator for life because it's supposedly racist or something. > With that kind of moral compass, I’m not sure I'd be missing its absence. Please note that most forms of media and social media have no problem with politicians making credible threats of violence against entire groups of people. Politicians are subject t…
The actual issue is Midjourney not allowing regular users generate certain type of material solely because it makes fun of a political figure. What you are talking about is entirely tangential to the issue the grandparent comment is talking about.
Re: Stanford Alpaca web demo suspended “until further notice”
#56Earlier quoted context omitted.
I was under the impression that LLaMA was also trained against a series of moral policies, but perhaps I'm mistaken. It seems Meta chose their words carefully to imply that LLaMA does in fact, not have moral training: > There is still more research that needs to be done to address the risks of bias, toxic comments, and hallucinations in large language models. Like other models, LLaMA shares these challenges. As a fou…
It doesn't appear to be filtered in a significant way. While toying with the 30B model, it suddenly started to steer a chat about a math problem into quite a sexual direction, with very explicit language. It also happily hallucinated, when prompted, that climate change is a hoax, as the earth is actually cooling down rapidly, multiple degrees per year, with a new ice age approaching in the next years. :D
Re: Stanford Alpaca web demo suspended “until further notice”
#57Earlier quoted context omitted.
I was under the impression that LLaMA was also trained against a series of moral policies, but perhaps I'm mistaken. It seems Meta chose their words carefully to imply that LLaMA does in fact, not have moral training: > There is still more research that needs to be done to address the risks of bias, toxic comments, and hallucinations in large language models. Like other models, LLaMA shares these challenges. As a fou…
"Moral training" Just as dystopian it sounds. Fixing current subjective moral norms into the machine.
Re: Stanford Alpaca web demo suspended “until further notice”
#58Is this an indication that the biggest impact from LLMs will be on the edge? It's almost a certainty that a model as good (or better) than Alpaca's fine-tuned LLaMA 7B will be made public within the next or two. And it's been shown that a model of that size can run on a Raspberry Pi with decent performance and accuracy. With all that being the case, you could either use a service (with restrictions, censorship, etc)…
Even then, I feel like the play will be an enterprise service instead of licensing.
Re: Stanford Alpaca web demo suspended “until further notice”
#59Earlier quoted context omitted.
Humans don’t scale infinitely. Humans have agency.
And if the LLM scales infinitely it still does nothing unless a human reads and acts on it. And as you said: Humans don't scale, and have agency. Just writing something bad doesn't actually mean something bad happened. These days it seems like people are oversensitive to how things are said to them, and what things are said to them.
Re: Stanford Alpaca web demo suspended “until further notice”
#60Earlier quoted context omitted.
I was under the impression that LLaMA was also trained against a series of moral policies, but perhaps I'm mistaken. It seems Meta chose their words carefully to imply that LLaMA does in fact, not have moral training: > There is still more research that needs to be done to address the risks of bias, toxic comments, and hallucinations in large language models. Like other models, LLaMA shares these challenges. As a fou…
It doesn't appear to be filtered in a significant way. While toying with the 30B model, it suddenly started to steer a chat about a math problem into quite a sexual direction, with very explicit language. It also happily hallucinated, when prompted, that climate change is a hoax, as the earth is actually cooling down rapidly, multiple degrees per year, with a new ice age approaching in the next years. :D