Meta torrented & seeded 81.7 TB dataset containing copyrighted data
11–20 of 981 posts
So according to some AI, the damages awarded per infringed work is ~$750 minimum in the US. 80TB of books, each let's say 10MB on average, would be 8 million works. So Meta should pay 6 billion USD for their copyright infringement?
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#12So if I torrented and seeded, I would be doing it for my own entertainment, not commercially. I expect big copy-write holders to come after myself. If Meta does it - I guess they have better lawyers ?
Could make interesting case law.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#13Earlier quoted context omitted.
I’ve got 70-80mb pirated books, I think because of the illustrations. Guess it depends on the book.
I don’t think they’re using picture heavy book for LLM training, no?
Even if they didn't use the illustration(which isn't clear given multimodal models), they'd still make use the text in the books.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#14Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#15Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#16Eye for an eye. Meta losses rights to 81.7 TB of IP. Transcribed into a text file
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#17Something tells me uncle Donald will exonerate his new favourite lapdog from any criminal or civil liability.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#18Earlier quoted context omitted.
I’ve got 70-80mb pirated books, I think because of the illustrations. Guess it depends on the book.
I don’t think they’re using picture heavy book for LLM training, no?
I don't think they need to be selective. It's not like Meta can run out of storage.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#19So according to some AI, the damages awarded per infringed work is ~$750 minimum in the US. 80TB of books, each let's say 10MB on average, would be 8 million works. So Meta should pay 6 billion USD for their copyright infringement?
Nice calculation, that’s actually quite doable for them, they have already been paying similar fines for a while.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#20A good chance for federal prosectutors to "send a message" as they did with Aaron Swartz but I don't see things going that way.