Live data from Hacker News

The New York Times is suing OpenAI and Microsoft for copyright infringement

theverge.com

811–820 of 912 posts

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#811

I have deeply mixed feelings about the way LLMs slurp up copyrighted content and regurgitate it as something "new." As a software developer who has dabbled in machine learning, it is exciting to see the field progress. But I am also an author with a large catalog of writings, and my work has been captured by at least one LLM (according to a tool that can allegedly detect these things). Overall, current LLMs remind me…

Ehh LLMs have become a fundamental part of my work flow as a professional. GPT4 is absolutely capable of providing links to sources and citations. It is more reliable than most human teachers I have had and doesnt have an ego about its incorrect statements when challenged on them. It does become less useful as you get more technical or niche but its incredibly useful for learning in new areas or increasing the breadt…

[deleted]

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#812

I hope this results in Fair Use being expanded to cover AI training. This is way more important to humanity's future than any single media outlet. If the NYT goes under, a dozen similar outlets can replace them overnight. If we lose AI to stupid IP battles in its infancy, we end up handicapping probably the single most important development in human history just to protect some ancient newspaper. Then another country…

I hope the nyt skullfucks this field. Humanity's future? You're doing statistics on stolen labor.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#813
post #417

Earlier quoted context omitted.

"probably the single most important development in human history" is the kind of hyperbole you'd only find here. Better than medicine, agriculture, electrification, or music? That point of view simply does not jive with what I see so far from AI. It has had little impact beyond filling the internet with low-effort content. I feel like the crypto evangelists never got off the hype train. They just picked a new destina…

Also the assumption a publication that’s been around for 150 years is disposable, not the web application that was created a year ago. I’ve been saying for a while that people’s credulity and impulse to believe absolutely any storyline related to technology is off the charts.

This is hackernews. Many people here work for startups and big tech companies. Their fortunes are tied to the perception that the technology they build is disruptive and valuable. They're not impartial.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#814
post #491

Earlier quoted context omitted.

Well front page of HN right now is an article about how AI aided in the development of a new antibiotic

It wasn't LLM. It was a graph network.

Almost like solving real problems requires enough domain knowledge to select an appropriate algorithm instead of relying on some magic black box trained by Microsoft on the whole internet.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#815
post #725

Earlier quoted context omitted.

Can I ask what industries with what application? I've seen lots of task like summarizing articles or producing text. The image and video work seems too rudimentary to be taken seriously. Is there something out there that seems like a killer application? I was amazed at the idea of the block chain but we never found a use for it outside of cryptocurrency. I see a similariy with AI hype.

For me, thinking about it as a search engine on steroids is enough. The internet has changed the world. Economically, socially, technologically, psychologically, pretty much everything is now related to it in one or other way, in this sense the internet is comparable to books. AI is another step in that direction. There is a very real possibility that the day will come when you can get, say, personalized expert nutri…

It kind of sucks ass at being a search engine though considering how often it straight up lies or makes things up.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#816
post #459

Earlier quoted context omitted.

> There is no way to know what sources have been memorized vs which have made their mark by affecting other types of functions in the neural net. But if it's possible for the neural net to memorize passages of text then surely it could also memorize where it got those passages of text from. Perhaps not with today's exact models and technology, but if it was a requirement then someone would figure out a way to do it.

Except it doesn’t memorize text. It generates text that is statistically likely. Generating a citation that is statistically likely wouldn’t really help the problem.

So it's just bullshit then.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#817

Earlier quoted context omitted.

The fact that copyright protection is far too long is entirely separate from the need for some kind of copyright protection to exist at all. All evidence suggests that it's completely impossible to live off your work unless you copyright it for some reasonable period, with the possible exception of performance art (music, theater, ballet). A writer or journalist just can't make money if any huge company can package t…

> All evidence suggests that it's completely impossible to live off your work unless you copyright it for some reasonable period Which evidence?

Chinas’s accession to the Universal Copyright Convention, and an alleged desire to comply with international IP law, led to an influx of OECD IP and foreign direct investment(FDI).

In hindsight, China wasn’t diligent in the enforcement of IP violations. However, it’s clear foreign presences and investment grew substantially in China during the early 90s upon the belief IP would be protected, or at the very least there would be recourse for violations.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#818

Earlier quoted context omitted.

Why don’t they train their AI on non-copyrighted material? It’s only fair for the copyright owners to want a share of the pie. I’d want one as well for my work.

Because no one forced them to, and the copyrighted dataset is much larger? It's like trying to teach your kids using only non copyrighted textbooks. There's not much out there. Copyright is an ancient system that is a poor legal framework for the modern world, IMO. I don't think it should exist at all. Of course as a rightsholder you are free to disagree. If we can learn and recite information, and a robot can too, t…

Ok, your reasoning escapes me. NYT has the right to sue and like any other business it’s holding onto their moat. Why would they let OpenAI train on their propery? Why wouldn’t they train their own AI on their own data?

Open AI is a business. NYT is a business. MS is a business. Neither will be happy when some other party takes something away from them without paying.

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#819

Earlier quoted context omitted.

Why don’t they train their AI on non-copyrighted material? It’s only fair for the copyright owners to want a share of the pie. I’d want one as well for my work.

Because they wouldn't have enough good quality training data then probably.

Too bad. Quality costs. Share the profits with everyone then and nobody would be unhappy

Re: The New York Times is suing OpenAI and Microsoft for copyright infringement

#820

Earlier quoted context omitted.

Why do you expect an AI to cite it's source? Humans are allowed to use and profit on knowledge they've learned from any and all sources without having to mention or even remember their sources. Yes, we all agree that it's better if they do remember and mention their sources, but we don't sue them for failing to do so.

Quite simply, if you're stating things authoritatively, then you should have a source.

Do you have a source for this claim?
Post reply on HN