I have deeply mixed feelings about the way LLMs slurp up copyrighted content and regurgitate it as something "new." As a software developer who has dabbled in machine learning, it is exciting to see the field progress. But I am also an author with a large catalog of writings, and my work has been captured by at least one LLM (according to a tool that can allegedly detect these things). Overall, current LLMs remind me…
Ehh LLMs have become a fundamental part of my work flow as a professional. GPT4 is absolutely capable of providing links to sources and citations. It is more reliable than most human teachers I have had and doesnt have an ego about its incorrect statements when challenged on them. It does become less useful as you get more technical or niche but its incredibly useful for learning in new areas or increasing the breadt…
The New York Times is suing OpenAI and Microsoft for copyright infringement
811–820 of 912 posts
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#812I hope this results in Fair Use being expanded to cover AI training. This is way more important to humanity's future than any single media outlet. If the NYT goes under, a dozen similar outlets can replace them overnight. If we lose AI to stupid IP battles in its infancy, we end up handicapping probably the single most important development in human history just to protect some ancient newspaper. Then another country…
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#813Earlier quoted context omitted.
"probably the single most important development in human history" is the kind of hyperbole you'd only find here. Better than medicine, agriculture, electrification, or music? That point of view simply does not jive with what I see so far from AI. It has had little impact beyond filling the internet with low-effort content. I feel like the crypto evangelists never got off the hype train. They just picked a new destina…
Also the assumption a publication that’s been around for 150 years is disposable, not the web application that was created a year ago. I’ve been saying for a while that people’s credulity and impulse to believe absolutely any storyline related to technology is off the charts.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#814Earlier quoted context omitted.
Well front page of HN right now is an article about how AI aided in the development of a new antibiotic
It wasn't LLM. It was a graph network.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#815Earlier quoted context omitted.
Can I ask what industries with what application? I've seen lots of task like summarizing articles or producing text. The image and video work seems too rudimentary to be taken seriously. Is there something out there that seems like a killer application? I was amazed at the idea of the block chain but we never found a use for it outside of cryptocurrency. I see a similariy with AI hype.
For me, thinking about it as a search engine on steroids is enough. The internet has changed the world. Economically, socially, technologically, psychologically, pretty much everything is now related to it in one or other way, in this sense the internet is comparable to books. AI is another step in that direction. There is a very real possibility that the day will come when you can get, say, personalized expert nutri…
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#816Earlier quoted context omitted.
> There is no way to know what sources have been memorized vs which have made their mark by affecting other types of functions in the neural net. But if it's possible for the neural net to memorize passages of text then surely it could also memorize where it got those passages of text from. Perhaps not with today's exact models and technology, but if it was a requirement then someone would figure out a way to do it.
Except it doesn’t memorize text. It generates text that is statistically likely. Generating a citation that is statistically likely wouldn’t really help the problem.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#817Earlier quoted context omitted.
The fact that copyright protection is far too long is entirely separate from the need for some kind of copyright protection to exist at all. All evidence suggests that it's completely impossible to live off your work unless you copyright it for some reasonable period, with the possible exception of performance art (music, theater, ballet). A writer or journalist just can't make money if any huge company can package t…
> All evidence suggests that it's completely impossible to live off your work unless you copyright it for some reasonable period Which evidence?
In hindsight, China wasn’t diligent in the enforcement of IP violations. However, it’s clear foreign presences and investment grew substantially in China during the early 90s upon the belief IP would be protected, or at the very least there would be recourse for violations.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#818Earlier quoted context omitted.
Why don’t they train their AI on non-copyrighted material? It’s only fair for the copyright owners to want a share of the pie. I’d want one as well for my work.
Because no one forced them to, and the copyrighted dataset is much larger? It's like trying to teach your kids using only non copyrighted textbooks. There's not much out there. Copyright is an ancient system that is a poor legal framework for the modern world, IMO. I don't think it should exist at all. Of course as a rightsholder you are free to disagree. If we can learn and recite information, and a robot can too, t…
Open AI is a business. NYT is a business. MS is a business. Neither will be happy when some other party takes something away from them without paying.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#819Earlier quoted context omitted.
Why don’t they train their AI on non-copyrighted material? It’s only fair for the copyright owners to want a share of the pie. I’d want one as well for my work.
Because they wouldn't have enough good quality training data then probably.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#820Earlier quoted context omitted.
Why do you expect an AI to cite it's source? Humans are allowed to use and profit on knowledge they've learned from any and all sources without having to mention or even remember their sources. Yes, we all agree that it's better if they do remember and mention their sources, but we don't sue them for failing to do so.
Quite simply, if you're stating things authoritatively, then you should have a source.