Live data from Hacker News

Meta torrented & seeded 81.7 TB dataset containing copyrighted data

arstechnica.com

321–330 of 981 posts

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#321
post #37

Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…

In Spotify’s defense, they used the pirated data only to show a proof of concept to the copyright holders, and that use was sanctioned by the local rights holders organization STIM. The copyright holders then approved their concept, and subsequently Spotify got the rights to offer their service to customers. Everybody won.

That’s not entirely true, in Spotify’s early days you could upload files to the service and listen to songs uploaded by other people. I think the majority of any song I wanted to listen to before they went Europe-only for a time was “pirated”.

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#322
post #37

Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…

Spotify was born as a response to piracy. Why do you say their catalog was pirated?

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#323
post #118

We have at least 4 types of ill-defined concepts of property in the 21st century , largely due to our laziness, intellectual inertia and lack of motivation to make forward-thinking definitions for the coming age of AI and ubiquitous access to all information and all communication. 1) the concept of copyright is as old as the word suggests (copies are the least of our worries going forward - it should be possible to d…

>we have an ill-defined concept of "personally identifying information" which gives people ownership to information that others have created via their own means - there should be better ways to ensure a level of privacy (but not absolute privacy) without overly-broad, nonsensical definitions of what is personally protected information What information about me could a corporation create via its own means that would b…

It's not just about corporations. Banking and government services e.g. are required to keep your personal information stored for years and years even against your will

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#324
post #166

Earlier quoted context omitted.

" If you plug a laptop into a closet at MIT to download some scientific papers you forfeit your life." This is exactly what I immediately thought while reading the article. It almost feels like the legal system only punishes general public, while most of these guys are above it.

Airbnb and Uber have showed us that laws matter only to the extent that the political will to enforce them exists. Throw enough lawyers and lobbying money at the problem and the laws can simply be re-written to be friendlier to your business model.

That's what makes a banana republic, and for all intents and purposes the U.S. are exhibit A.

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#325

Earlier quoted context omitted.

the reform needs to happen at the layer where whether a copyright is valid or not is decided upon, not before (at the point of "should copyright exist") and not after (enforcement). a world without copyright means those with the largest advertising budgets will reap nearly all the rewards from new IP created by small artists. BigCorp Inc. can just sit around and wait for talented musicians to post something interesti…

I don’t think this is true. At least in music, bands make far more money from touring and merch than they do from music sales. If copyright disappeared altogether, most smaller artists would be just fine because they have loyal fans and adjacent monetization strategies. See: Grateful Dead. They did just fine despite encouraging infringement of IP. IMO copyright mostly serves to protect the very biggest artists and co…

I think the point was that the big corporations get the money from selling music.

And saying that bands currently make more money from touring kind of proves the point. They get too low % cut of music sales.

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#326
post #166

Earlier quoted context omitted.

Airbnb and Uber have showed us that laws matter only to the extent that the political will to enforce them exists. Throw enough lawyers and lobbying money at the problem and the laws can simply be re-written to be friendlier to your business model.

The hotel and taxi industry were legit terrible before those two disrupted them. Laws are ment to be broken. Especially in cronist systems where incumbents write the laws.

Hotels were just fine.

Taxis were discriminatory and "uncool" to the point were Uber has saved thousands by preventing drunk driving.

Now if you go out with the boys and get drunk, it's a 30 second casual call to get an Uber and get home.

Live in a neighborhood Taxis are afraid to service,you can either make some extra income working for Uber or use it yourself. When Ubers used as its intended purpose, to basically make a quick buck, it's a lifeline to many low income people .

Say your rents it's going to be late, you can pick up 20 or 30 hours of Uber this month to make it happen. It's not really a career though...

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#327
post #37

Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…

Corporations are people. Just a notch above the regular kind.

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#328

Earlier quoted context omitted.

the english empire once tried to mantain a monopoly over steam loom machines the americans cheated their way to competition, heck, even before that, the english empire got jumpstarted by stealing gold from the spanish (who were themselves exploiting it away from aztec and other mexican natives) I'm saying it's business as usual, but also, culture doesn't work like tangible physical widgets so we must stop letting a f…

I don't think I've heard the term "English empire". Is it an attempt by the Scottish to pretend they weren't involved?

Is this an attempt to imply the Scots had imperial ambitions and have not been fighting to keep their homes free of invasion for several thousand years?

Fuck this sounds familiar right now

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#329
post #37

Based on the encyclopedic knowledge LLMs have of written works I assume all parties did the same. But I think there is a broader point to make here. Youtube was initially a ghost town (it started as a dating site) and it only got traction once people started uploading copyrighted TV shows to it. Google itself got big by indexing other people's data without compensation. Spotify's music library was also pirated in the…

Everyone on here is smart enough. Just do not participate and save your money. Do not pay for digital goods. If Netflix raises their prices, it doesn't matter because there is a torrent of all of their shows. If Spotify raises their prices, it doesn't matter because your favorite artist has their entire library in a torrent. If some game company ask you to pay real life prices for a digital costume, find the crack on…

I just can't get behind the sentiment that the unethical behavior by big companies means I get to access all the content I want for free.

Re: Meta torrented & seeded 81.7 TB dataset containing copyrighted data

#330
post #166

Earlier quoted context omitted.

Airbnb and Uber have showed us that laws matter only to the extent that the political will to enforce them exists. Throw enough lawyers and lobbying money at the problem and the laws can simply be re-written to be friendlier to your business model.

The reason there was no political will to punish Airbnb and Uber for violating the law was that initially they were subsidized with VC money and so were able to undercut traditional hotels and taxis on price. In the world of tradable goods, pricing below cost with the intent of putting competition out of business so you can raise prices later is known as "dumping" and is itself illegal.

Speed had a lot to do with this as well.

VC funding allowed them to move quickly enough that they got to a scale where they could afford legal and lobbying protection when challenges eventually happened.

Post reply on HN