There is a bigger reason why the end of open source AI might be close: as soon as training data becomes licensed, that’s it for open source AI. Poof. I wish I could be more eloquent on this point, but I’ve mostly just been depressed about this seeming inevitability. Hopefully it won’t be the case. But how could it be otherwise? Hundreds of thousands of people are mad at openai and midjourney for doing exactly what op…
Many AI safety orgs have tried to criminalize currently-existing open-source AI
11–20 of 405 posts
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#12Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#13There is a bigger reason why the end of open source AI might be close: as soon as training data becomes licensed, that’s it for open source AI. Poof. I wish I could be more eloquent on this point, but I’ve mostly just been depressed about this seeming inevitability. Hopefully it won’t be the case. But how could it be otherwise? Hundreds of thousands of people are mad at openai and midjourney for doing exactly what op…
The advantage that open AI (not the company) has is that if using copyrighted content as training data without licensing it is found to be illegal, they can just keep doing it. There's plenty of FOSS software basically designed to violate copyright law (comic readers, home media center servers/clients, torrent clients) that big tech cannot compete with lest they face legal consequences. Basically what I'm saying is t…
> open models that can use whatever data they like but must operate in the shadows without access to much compute for training
sounds pretty analogous to private trackers, where high-quality stuff is available but not in full public view. If rightsholders and big tech, abetted by states, crack down on open models I think you're right that the open source community will continue to train on liberated content, but it's not going to be as open and free-flowing as things are now. Going back to the compute problem, I can imagine analogues of private trackers where contributors, not wanting to expose themselves to whatever the law-firm letter/ISP strike analogue will be in this space, use invite-only closed networks to pool compute resources for training on encumbered content.
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#14There is a bigger reason why the end of open source AI might be close: as soon as training data becomes licensed, that’s it for open source AI. Poof. I wish I could be more eloquent on this point, but I’ve mostly just been depressed about this seeming inevitability. Hopefully it won’t be the case. But how could it be otherwise? Hundreds of thousands of people are mad at openai and midjourney for doing exactly what op…
The advantage that open AI (not the company) has is that if using copyrighted content as training data without licensing it is found to be illegal, they can just keep doing it. There's plenty of FOSS software basically designed to violate copyright law (comic readers, home media center servers/clients, torrent clients) that big tech cannot compete with lest they face legal consequences. Basically what I'm saying is t…
Consider that you just spend a few millions on training a model on copyrighted data. Release it would reveal that, problem.
I _guess_ you can try doing training in the public, like Seti @ Home or something like that, which distributes the risk? But no idea if this is even possible in this context.
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#15standard capitalist play to set up a moat by instilling fear
standard socialist play to use the government's hand to achieve their own objectives
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#16As a developer, do I need their certification or is it like a MADD situation where we if we don't observe and diminish their appeal/growth we will get laws with draconian measures that only benefit the big players? e.g. a PAC? (Don't get me wrong, drunk driving is "bad", but for a group like MADD to exist, eh and meh.)
As of now, there is no need to really submit these these safety orgs, so as long as we don't care for their approval - who cares, right?
Or is the optics far gone now that these orgs control the conversation?
Also, fun question: with AI safety orgs now attempting to police - who really polices the police / do other orgs/countries have the same rules, safety and artificial guard rails?
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#17Earlier quoted context omitted.
The advantage that open AI (not the company) has is that if using copyrighted content as training data without licensing it is found to be illegal, they can just keep doing it. There's plenty of FOSS software basically designed to violate copyright law (comic readers, home media center servers/clients, torrent clients) that big tech cannot compete with lest they face legal consequences. Basically what I'm saying is t…
As long as the cost for training a model is in the 7+ figures, that means that any such open model is bound to be tracked to someone with deep enough pockets to sue. Consider that you just spend a few millions on training a model on copyrighted data. Release it would reveal that, problem. I _guess_ you can try doing training in the public, like Seti @ Home or something like that, which distributes the risk? But no id…
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#18There is a bigger reason why the end of open source AI might be close: as soon as training data becomes licensed, that’s it for open source AI. Poof. I wish I could be more eloquent on this point, but I’ve mostly just been depressed about this seeming inevitability. Hopefully it won’t be the case. But how could it be otherwise? Hundreds of thousands of people are mad at openai and midjourney for doing exactly what op…
The advantage that open AI (not the company) has is that if using copyrighted content as training data without licensing it is found to be illegal, they can just keep doing it. There's plenty of FOSS software basically designed to violate copyright law (comic readers, home media center servers/clients, torrent clients) that big tech cannot compete with lest they face legal consequences. Basically what I'm saying is t…
The key difference is that those projects don't violate copyright themselves, but facilitate users doing so. If training without a license is infringement then projects that are doing so will struggle to host code/models publicly or access other parts of internet infrastructure.
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#19standard capitalist play to set up a moat by instilling fear
> Many AI safety orgs have tried to __criminalize__ currently-existing open-source AI standard socialist play to use the government's hand to achieve their own objectives
Re: Many AI safety orgs have tried to criminalize currently-existing open-source AI
#20AI safety people are hypocrites. If they practiced what they preached, they'd be calling for all AI to be banned, ala Dune. There are AI harms that don't care about whether or not the weights are available, and are playing out today. I'm talking about the ability of any AI system to obfuscate plagiarism[0] and spam the Internet with technically distinct rewords of the same text. This is currently the most lucrative u…
Compared to that, an open source LLM telling a curious teenager how to make gunpowder is... laughable.
This entire debacle is an example of disgusting "think of the children!" doublespeak, officially about safety, but really about locking shit down under corporate control.