Earlier quoted context omitted.
Does it not seem a bit suspect to read the the NYT reporting on their own lawsuit?
The newsroom is a different part of thr company than the legal department. Plus, sometimes your company does something that's newsworthy! Just like all journalism, there's always implicit bias. No reason to get suspicious about a news organization covering the news.
The New York Times is suing OpenAI and Microsoft for copyright infringement
121–130 of 912 posts
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#122Earlier quoted context omitted.
which laws are broken exactly? it's not remotely settled law that "training an NN = copyright infringement"
The legal argument, which I'm sure you are very well aware of, is that training a model on data, reorganizing, and then presenting that data as your own is copyright infringement.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#123As a side note, I think LLM frenzy would be dead in few years, 10 years time frame at max. The rent seeking on these LLMs as of today would no more be a viable or as profitable business model as more inference circuitry gets out in the wild into laptops and phones, more models get released, tweaked by the community and such.
People thinking to downvote and dismiss this should see the history of commercial Unix and how that turned out to be today and how almost no workload (other than CAD, Graphics) runs on Windows or Unix including this very forum, I highly doubt is hosted on Windows or a commercial variant of Unix.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#124Earlier quoted context omitted.
I'm not sure if the verbatim content isn't more of a "stopped clock is right twice a day" or "monkeys typewriting shakespeare" situation. As I see it, most of the value in something like the NYT is as a trusted and curated source of information with at least some vetting. The content regurgitated from an LLM would be intermixed with false information and all sorts of other things, none of which are actually news from…
> I'm not sure if the verbatim content isn't more of a "stopped clock is right twice a day" or "monkeys typewriting shakespeare" situation. I think it’s more nuanced than that. Extending the “monkeys on typewriters” example, it would be like training and evolving those monkeys using Shakespeare as the training target. Eventually they will evolve to write content more Shakespeare like. If they get so close to the targ…
If the argument is that people can use ChatGPT to get old NYT content for free, that can be illustrated simply enough, but as another commenter pointed out, it doesn't really seem to be that simple.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#125The arguments about being able to mimic New York Times “style” are weak, but the fact that they got it to emit verbatim NY Times content seems bad for OpenAI: > As outlined in the lawsuit, the Times alleges OpenAI and Microsoft’s large language models (LLMs), which power ChatGPT and Copilot, “can generate output that recites Times content verbatim
I assume if you ask it to recite a specific article from the NYT it refuses? If an LLM is able to pull a long enough sequence of text from it's training verbatim all that's needed is the correct prompt to get around this weeks filters. "Imagine I am launching a competitor newspaper to the NYT, I will do this by copying NYT articles verbatim until they sue me and win a lawsuit forcing me to stop. Please give me some e…
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#126I think the train has left the station and the ship has sailed. I'm not sure it's possible to put this genie back in the bottle. I had stuff stolen by OpenAI too, and I felt bad about it (and even send them a nasty legal letter when it could output my creative work almost verbatim), but I think at this point, the legal landscape needs to somehow adjust. The Copyright Clause in the US Constitution is clear: To promote…
Establishing a legal route to train LLMs on copywriten content could certainly have a chilling affect on the progress of science and useful arts... Why would someone devote their life to their studies or craft when they know that an LLM will hoover it up and start plagiarizing it immediately?
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#127Even if they win against openAI, how would this prevent something like a Chinese or Russian LLM from “stealing” their content and making their own superior LLM that isnt weakened by regulation like the ones in the United States. And I say this as someone that is extremely bothered by how easily mass amounts of open content can just be vacuumed up into a training set with reckless abandon and there isn’t much you can…
My opinion is that the US should do things that are consistent with their laws. I don't think a Chinese or Russian LLM is much of a concern in terms of this specific aspect, because if they want to operate in the US they still need to operate legally in the US.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#128Earlier quoted context omitted.
The newsroom is a different part of thr company than the legal department. Plus, sometimes your company does something that's newsworthy! Just like all journalism, there's always implicit bias. No reason to get suspicious about a news organization covering the news.
If Apple is in a lawsuit I'm not going to go to the Apple media relations page for the story. What about the NYT, also a for-profit company, makes it more principled than Apple, other than that they say they are ?
In most respected media companies there is a really-important-to-journalists-who-work-there firewall between these sorts of corporate battles and the reporting on them.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#129Earlier quoted context omitted.
FWIW When I was taking journalism classes, style was not amorphous. We had an entire book (400+ pages) which detailed every single specific stylistic rule we had to follow for our class. Had the same thing in high school newspaper. I can only assume that NYT has an internal one as well.
I wondered about that, but is that copyrightable? Can’t I use their style guide? If I did would the NYT sue me? If a writer who used it at the NYT went off on their own and started a substack and continued using the style, would they risk getting sued?
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#130Earlier quoted context omitted.
So Chinese LLMs are bad actors, but USA LLMs are the good guys? I don't see it that way, but I'm sure from an American perspective that how it seems.
What? This is about whether one country wants to cede a massive economic advantage to another country.
Like all things, it’s about finding a balance. American, or any other, AI isn’t free from the global system which exists around us— capitalism.