Live data from Hacker News

OpenAI's new open-source model is basically Phi-5

seangoedecke.com

111–120 of 233 posts

Re: OpenAI's new open-source model is basically Phi-5

#111

Earlier quoted context omitted.

I can’t tell if the part of this post after they have made this “thing of questionable safety” and then bury their head in the sand is satirical or not.

Why is it controversial to state that medical expertise is a technical field that engineers should defer to experts in? It doesn't mean I can do things recklessly. Instead it's an acknowledgment of when I need to defer to somebody else just like I need to call up an attorney for legal stuff or an accountant for tax stuff.

It’s not controversial but it doesn’t address my question at all.

A: I’m curious about X.

B: We should trust the experts!

Sure, but what do the experts say? That was my entire question.

Re: OpenAI's new open-source model is basically Phi-5

#112
From the article:

> For the same reason that Microsoft probably continued to train Phi-style models: safety. Releasing an open-source model is terrifying for a large organization. Once it’s out there, your name is associated with it forever, and thousands of researchers will be frantically trying to fine-tune it to remove the safety guardrails.

I don't think this is really an issue in practice. Llama 2 and 3 were uncensored within a week. There's no bad press about this.

What does give a company a bad reputation is crap models. The llama 4 disappointment hurt meta's AI reputation a lot more than some community uncensoring.

Re: OpenAI's new open-source model is basically Phi-5

#113
post #20

I saw a bunch of people complaining on Twitter about how GPT-OSS can't be customized or has no soul and I noticed that none of them said what they were trying to accomplish. "The main use-case for fine-tuning small language models is for erotic role-play, and there’s a serious demand." Ah.

Most use of small models was erotic roleplay back when I closely followed r/LocalLlama. Made me wonder if I should even make one if that's what most used it for.

You might find this part funny. At first, I thought they had automated their coding for SAP or other ERP databases. Then, they started talking about how realistic the body parts were. I paused staring at the screen. The sad reality clicked.

Re: OpenAI's new open-source model is basically Phi-5

#114
post #26
post #20

I saw a bunch of people complaining on Twitter about how GPT-OSS can't be customized or has no soul and I noticed that none of them said what they were trying to accomplish. "The main use-case for fine-tuning small language models is for erotic role-play, and there’s a serious demand." Ah.

what's the problem with that? we have erotic texts dating back thousands of years, basically as old as the act of writing itself https://en.wikipedia.org/wiki/Istanbul_2461

[flagged]

Re: OpenAI's new open-source model is basically Phi-5

#115
post #68

Earlier quoted context omitted.

What about two consenting adults engaging in age play. By your logic, wouldn't that also lead to "real crime"?

There’s probably some difference between someone who is visibly an adult and mature vs more deeply entrenching pathways of arousal in response to someone who is visibly a child. I still find the adult fetish version repellent but it’s also really hard to police in a way that’s remotely ethically permissible. Ie yes it’s bad and in an ideal world nobody would do it. I see trying to restrict or ban it as the greater of…

An ideal world with less speeches and acts than we have today is a monastery, and monasteries are definitely not a model of an ideal world.

Re: OpenAI's new open-source model is basically Phi-5

#117

No training data, no open source. Don't fall for the company PR.

As long as it works who cares about training data. Obviously they can't open it for many reasons. License is one of them.

I don't care if it's open source or not. I care if people call something open source which it isn't.

Do you care about binary blobs in the kernel? No. Are binary blobs in the kernel open source? No.

But it is tedious to go through the same discussion every 10 years, with a relentless industry that wants to dupe people.

If there wasn't a benefit in for them, they would not call it open source.

Re: OpenAI's new open-source model is basically Phi-5

#118

From the article: > For the same reason that Microsoft probably continued to train Phi-style models: safety. Releasing an open-source model is terrifying for a large organization. Once it’s out there, your name is associated with it forever, and thousands of researchers will be frantically trying to fine-tune it to remove the safety guardrails. I don't think this is really an issue in practice. Llama 2 and 3 were unc…

> researchers will be frantically trying to fine-tune it to remove the safety guardrails.

It is a really weak excuse. They are more likely to take a reputation hit for having silly guardrails than for having someone to remove them.

Imagine if Bill Gates decided not to release MS Paint in 1985 because someone could have drawn something offensive with it.

Re: OpenAI's new open-source model is basically Phi-5

#119

This is really irresponsible, there's actual data on quality and this is just crazy to assert: "Any small online community for people who run local models is at least 50% perverts." --- I get what he's saying, there's def. communities where that's predominant, but it's simply not true in the vast majority of communities. Sort of like, in 1997, saying any small community on this here Internet thing is half perverts lo…

100%, what an insanely baseless article

Re: OpenAI's new open-source model is basically Phi-5

#120
post #109

Earlier quoted context omitted.

it's not erotic role-play, but I have a use case of making an AI-powered NetHack clone. specifically, to generate dungeon layouts, dialog for NPCs and to fill in the boatloads of minutae and interactions which NetHack is famous for. you kind of need soul for that, and a lot of background knowledge on mythology/fantasy lore, but also tool use to work the world systems.

Are you getting good results with this? Some time ago I build an "overworld simulator" by having a bunch if jsons that represented villages, characters, building, story hooks and plots. I just asked ChatGPT to "simulate the gameworld by a week". Technically this worked great, but everything was somewhat bland and generic. Noteable not-highlights: - there is a shimmer in the nearby forests that keep villagers up at ni…

Might be fixable with promoting and/or a much lower temperature?
Post reply on HN