Live data from Hacker News

Stanford Alpaca web demo suspended “until further notice”

alpaca-ai-custom4.ngrok.io

31–40 of 83 posts

Re: Stanford Alpaca web demo suspended “until further notice”

#31
post #24

Earlier quoted context omitted.

"Moral training" Just as dystopian it sounds. Fixing current subjective moral norms into the machine.

Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.

Example of "moral policy" in practice: Midjourney appears to be banning making fun of the Chinese dictator for life because it's supposedly racist or something.

With that kind of moral compass, I’m not sure I'd be missing its absence.

Re: Stanford Alpaca web demo suspended “until further notice”

#32
post #24

Earlier quoted context omitted.

"Moral training" Just as dystopian it sounds. Fixing current subjective moral norms into the machine.

Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.

There was the famous example of chatgpt refusing to disable a nuke in the middle of NYC by using a racial slur.

I don't think anyone in real life would choose that tradeoff but it's what happens when all of your "safety" training is about US culture war buttons.

Re: Stanford Alpaca web demo suspended “until further notice”

#34
post #22

I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…

> who wants unfettered ability to interact with a LLM without moral concerns. Is that bad? It's just a language model - it says things. Humans have been saying all sorts of terrible things for ages, and we are still here. I mean it makes for nice headline "model said something racist", but does it actually change anything? These aren't decision making AI's (which would need to be much more careful), they are language…

The content it produces is quite hard to distinct from text written by real person. Not long ago, Facebook was accused of "playing a critical role" in Rohingya genocide. I dont really know but I believe worst case LLM risks are in that same category.

Re: Stanford Alpaca web demo suspended “until further notice”

#35

I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…

[flagged]

Re: Stanford Alpaca web demo suspended “until further notice”

#37

Earlier quoted context omitted.

Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.

Example of "moral policy" in practice: Midjourney appears to be banning making fun of the Chinese dictator for life because it's supposedly racist or something. With that kind of moral compass, I’m not sure I'd be missing its absence.

> Example of "moral policy" in practice: Midjourney appears to be banning making fun of the Chinese dictator for life because it's supposedly racist or something.

> With that kind of moral compass, I’m not sure I'd be missing its absence.

Please note that most forms of media and social media have no problem with politicians making credible threats of violence against entire groups of people.

Politicians are subject to a different set of rules, and enjoy a lot more protection than you and I.

Re: Stanford Alpaca web demo suspended “until further notice”

#38
post #24

Earlier quoted context omitted.

"Moral training" Just as dystopian it sounds. Fixing current subjective moral norms into the machine.

Do you think public schools are inherently dystopian? I don't think you're using the right critique here. Picking a common system of moral norms is a lot better than no moral norms.

Yes

Re: Stanford Alpaca web demo suspended “until further notice”

#39
post #19

Earlier quoted context omitted.

Where do you get the weights?

https://github.com/antimatter15/alpaca.cpp has links

Interesting issue in that repo.. https://github.com/antimatter15/alpaca.cpp/issues/23

Re: Stanford Alpaca web demo suspended “until further notice”

#40

I think it's only a matter of time until one of these models gets into the hands of a hacker who wants unfettered ability to interact with a LLM without moral concerns. I think there's tremendous value in end user facing LLMs being trained against moral policies, but for internal or private usage, if these models are trained on essentially raw WWW sourced data, I would personally want raw output. I'm also finding it…

> if these models are trained on essentially raw WWW sourced data, I would personally want raw output.

Llama is a very high-quality foundation LLM, you can already run it very easily using llama.cpp and will get the raw output you need. https://github.com/ggerganov/llama.cpp

There's already instructions on how anyone can fine-tune it to behave similarly to ChatGPT for as little as $100: https://crfm.stanford.edu/2023/03/13/alpaca.html

Post reply on HN