Is that true that most small language models are fine tuned for erotic role-play?
OpenAI's new open-source model is basically Phi-5
61–70 of 233 posts
Re: OpenAI's new open-source model is basically Phi-5
#62Earlier quoted context omitted.
what's the problem with that? we have erotic texts dating back thousands of years, basically as old as the act of writing itself https://en.wikipedia.org/wiki/Istanbul_2461
I have no problem with it and I can understand why people don't want to say "I'm trying to pornify this model and it refuses to talk dirty!" in public. But if you're calling a model garbage maybe you should be honest about what the "problem" is.
Re: OpenAI's new open-source model is basically Phi-5
#63Earlier quoted context omitted.
You may be interested in the data surfaced by this large-scale survey[1] [1] https://aella.substack.com/p/fetish-tabooness-and-popularity...
This is interesting but there's a little more to it, especially with the erotic. If people were polled what they want to see on social media, few would say things that are inflammatory, upsetting, divisive, etc but those as we know are strong drivers of engagement. It's because you're polling for affinity or disclosed preference not for the actual engagement drivers. For instance, if a male says they watch male porno…
That's because those are two entirely different things. If you polled people and asked them "what causes you to spend more time on social media", then at least some self-aware folks would likely identify conflict, "someone is wrong on the Internet" (https://xkcd.com/386/), etc. That doesn't mean that's "what they want to see on social media", that means that's "what gets them to spend more time on social media".
Re: OpenAI's new open-source model is basically Phi-5
#64Earlier quoted context omitted.
Porn is always the frontier. It's a well-understood self-contained use-case without many externalities and simple business models. What more, with porn, the medium is the product probably more than the content. Having it on home-media in the 80s was the selling point. Getting it over the 1-900 phone lines or accessing it over the internet ... these were arguably the actual product. It might have been a driver of earl…
even if it is victim free, it can affect mental health in a way that a consumer will be more compelled to do a criminal act and create a real victim. let's say you publish a Steam game how to be a school shooter and shoot kids, wouldn't that lead to real school shootings ? who can definitely say that computer generated content about criminal behavior, won't lead to real crime with real victims? https://en.wikipedia.o…
Re: OpenAI's new open-source model is basically Phi-5
#65Earlier quoted context omitted.
Porn is always the frontier. It's a well-understood self-contained use-case without many externalities and simple business models. What more, with porn, the medium is the product probably more than the content. Having it on home-media in the 80s was the selling point. Getting it over the 1-900 phone lines or accessing it over the internet ... these were arguably the actual product. It might have been a driver of earl…
even if it is victim free, it can affect mental health in a way that a consumer will be more compelled to do a criminal act and create a real victim. let's say you publish a Steam game how to be a school shooter and shoot kids, wouldn't that lead to real school shootings ? who can definitely say that computer generated content about criminal behavior, won't lead to real crime with real victims? https://en.wikipedia.o…
Let's be specific: Rape, incest, necrophilia, bestiality, and pedophila ideation.
I think we can all agree (1) these are harmful, anti-social behaviors that we do not want in our society, (2) people don't choose to have these desires, (3) most people who have them have no desire to actually traumatize others, (4) people who have these struggle with it.
These multi-media AI role-play environments would allow that type of engagement without any harm.
Now given all this, I am not a psychologist and do not know if that's part of how someone unfortunate enough to have those inclinations can deal with it healthily.
But if it is, now it exists and hopefully we can see less of it in the real world. I'm all for harm reduction if this is a way to get there.
Re: OpenAI's new open-source model is basically Phi-5
#66If a model is trained only on synthetic data, is it still possible it will output things like this? https://x.com/elder_plinius/status/1952958577867669892
It's not particularly likely that the hidden information encoded in synthetic data would happen to include specific details for making LSD or VX, but it's much more plausible that synthetic data contains some information the model's trainers would prefer to not incorporate in the model.
Re: OpenAI's new open-source model is basically Phi-5
#67> for instance, they have broad general knowledge about science, but don’t know much about popular culture That seems like a good focus. Why learn details that can change within days of it being released? Instead, train the models to have good general knowledge, and be really good at using tools, and you won't have to re-train models from scratch just because some JS library now has a different API, instead the model…
Yeah, it always seemed like a sad commentary on our world that AIs are devoting their weights to encyclopedic knowledge of Harry Potter, Pokemon, and Reddit trolling.
Re: OpenAI's new open-source model is basically Phi-5
#68Earlier quoted context omitted.
Porn is always the frontier. It's a well-understood self-contained use-case without many externalities and simple business models. What more, with porn, the medium is the product probably more than the content. Having it on home-media in the 80s was the selling point. Getting it over the 1-900 phone lines or accessing it over the internet ... these were arguably the actual product. It might have been a driver of earl…
even if it is victim free, it can affect mental health in a way that a consumer will be more compelled to do a criminal act and create a real victim. let's say you publish a Steam game how to be a school shooter and shoot kids, wouldn't that lead to real school shootings ? who can definitely say that computer generated content about criminal behavior, won't lead to real crime with real victims? https://en.wikipedia.o…
Re: OpenAI's new open-source model is basically Phi-5
#69I saw a bunch of people complaining on Twitter about how GPT-OSS can't be customized or has no soul and I noticed that none of them said what they were trying to accomplish. "The main use-case for fine-tuning small language models is for erotic role-play, and there’s a serious demand." Ah.
Want a good use case? I am playing around with interactive workflow where the model suggests what can be wrong with a particular chunk of code, then the user selects one of the options, and the model immediately implements the fix. Biggest problem? Total Wild West in terms of what the models try to suggest. Some models suggest short sentences, others spew out huge chunks at a time. GPT-OSS really likes using tables e…
For example: make the suggestion output an object with multiple fields, naming one of them `concise_suggestion`. And make sure to take advantage of the `description` field.
For people not already using structured output, both OpenAI and Anthropic consoles have a pretty good JSON schema generator (give prompt, get schema). I'd suggest using one of those as a starting point.
Re: OpenAI's new open-source model is basically Phi-5
#70Earlier quoted context omitted.
I have no problem with it and I can understand why people don't want to say "I'm trying to pornify this model and it refuses to talk dirty!" in public. But if you're calling a model garbage maybe you should be honest about what the "problem" is.
Why? Is there any reason to believe problems in that context won't generalise?
Maybe nothing wrong with that, but it might mean that the perceived weaknesses don't generalize to an area of the model that hasn't been lobotomized.
* using safety the way OpenAI have been using the term, not looking to debate the utility of that.