Earlier quoted context omitted.
True but lets take examples one by one to see what we can learn : Spotify was doing illegal things until they made a deal to become legal and not to be trialed over what they done. Seems like business deals is what saved them, not regulatory capture (the regulations around IP for music pre existed Spotify)
Sure that is what saved them initially, but following that early 2010’s period of hemorrhaging money and eventual recovery, then they started digging that moat. https://www.politico.com/story/2015/04/spotify-washington-lo... https://www.opensecrets.org/federal-lobbying/clients/issues?...
Meta torrented & seeded 81.7 TB dataset containing copyrighted data
141–150 of 981 posts
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#142Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…
" If you plug a laptop into a closet at MIT to download some scientific papers you forfeit your life." This is exactly what I immediately thought while reading the article. It almost feels like the legal system only punishes general public, while most of these guys are above it.
> There must be in-groups whom the law protects but does not bind, alongside out-groups whom the law binds but does not protect.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#143We all like hating big corporations, especially Meta, and people seem to use this as an opportunity to advocate for punishing them. I think it's wiser to advocate for changing our IP laws.
Also, change the law so this is legal for poor meta? smh..
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#144We all like hating big corporations, especially Meta, and people seem to use this as an opportunity to advocate for punishing them. I think it's wiser to advocate for changing our IP laws.
I truly hope Meta has a serious security issue that burns their company to the ground. That said, I want them to burn for the right reasons. Downloading data that should be available to the public is not one of them.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#145Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…
I guess the solution is to create a shell company for your illegal activities?
By the time the cheque comes, your illicit venture either went bust or you built a bilion dollar empire capable of buying the best lawyers and lobbying to walk away clean.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#146Earlier quoted context omitted.
Cain was severely punished. וְעַתָּ֖ה אָר֣וּר אָ֑תָּה מִן־הָֽאֲדָמָה֙ אֲשֶׁ֣ר פָּצְתָ֣ה אֶת־פִּ֔יהָ לָקַ֛חַת אֶת־דְּמֵ֥י אָחִ֖יךָ מִיָּדֶֽךָ׃ Therefore, you shall be more cursed than the ground, which opened its mouth to receive your brother’s blood from your hand. https://www.sefaria.org/Genesis.4.12
Jack the Ripper killed people and got away with it!!! Happy?
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#147Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…
> Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. I don't understand why you wouldn't just buy copies of the books. Seems like such a relatively inexpensive way to strengthen your legal case.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#148Earlier quoted context omitted.
For creating a backup of library genesis. No. They should be awarded a philanthropic prize.
There's evidence of them seeding back as little as possible. I'm not sure how that's "creating a backup".
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#149Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…
The limit is what you can actually get away with, not what the rules say you can get away with, and the system aggressively selects players who recognize this. It's amoral - there is no "ought", only "is". An actor gets punished or not, with absolutely no regard to whether it "should" get punished. One thing is consistent: following the rules as written means you lose.
You can see it in Y Combinator (and other) startups. The biggest ex-startups are things like AirBNB (hotels but we don't follow the rules but we don't get punished for not following them) and Uber (taxis but we don't follow the rules but we don't get punished for not following them).
One way to not get punished for not following the rules is to invent a variation of the game where the rules haven't been written yet. I again refer you to AirBNB and Uber; Omegle also comes to mind, although they didn't monetize.
Viewed in this light, Aaron Swartz's mistake was not the part where he downloaded journal articles, but the part where he got caught downloading journal articles. Shadow library sites are doing the same thing, minus the getting caught. So are Meta and Google and OpenAI. sci-hub is only involved in a lawsuit because it got caught and is now in the stage where it finds out whether it gets punished or not.
Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data
#150Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…
> Google itself got big by indexing other people's data without compensation Wrong. a) Robots.txt which defines what content you wish to make available to third parties predates every search engine including Google. Web site owners chose to make it available to Google and search engines have respected their wishes despite it not being in their best interest. b) The difference here is that OpenAI, Meta etc have not ev…