Live data from Hacker News

Taking action against scraping for hire

about.fb.com

111–120 of 240 posts

Re: Taking action against scraping for hire

#111
post #15

>Octopus, a US subsidiary of a Chinese national high-tech enterprise, built a cloud-based platform designed to provide paying customers access to on-demand scraping software and services. It is interesting as how they try to position this as a Chinese attack on them.

it look like Zack is giving up on the Chinese market.

I guess after Winnie the Pooh rejected to name his children for him he got sour grapes for China.

Re: Taking action against scraping for hire

#112
post #85

Earlier quoted context omitted.

> the vast majority of Web scraping efforts are to build businesses on top of other organizations hard work and innovation. Not really. Scraping just gets data, not code, so it's hard to support this argument. The anti-scraping view is that the right to use the data rests with the company that collected it, but I don't think that view is held by most people.

If you are arguing that an organization's data is worthless but only their code has worth, then I'm not quite sure where to go from this point in this discussion, other than to say that is crazy .

I'm not saying that the data isn't valuable, but that possession of the data, valuable though it may be, is not related to the organization's hard work or innovation. For the most part, any control rights to the data likely rest or should rest with the people who provided it to the company.

Meta claiming that all of the photos on Instagram are Meta's property does not comport with current IP law or the views/opinions of most of the users on Instagram who do own the copyrights to those photos.

You really shouldn't be able to sue anyone for use or copying of data to which you do not hold copyright. The stuff on FB is licensed to FB by the people who own it (their users).

Re: Taking action against scraping for hire

#113

Earlier quoted context omitted.

Quoted post unavailable.

> I love to hate on Meta, but their actions here are spot on and make my morning very enjoyable as I sip my cup of coffee. You might want to reassess your intelligence there friend. It seems to be suffering from a common form of cogntive dissonance combined with some form of confirmation bias. How so? Well you clearly don't like scraping, otherwise you wouldn't be agreeing with a criminal... So there's the confirmati…

Who is the criminal here? Scraping is not illegal. This is a civil suit, so even if Meta wins, it's still not remotely criminal for anyone involved.

Also please explain to me how someone giving a company their Facebook credentials is an example of "people who are smart enough to make use of Zuckerbergs terrible security practices."

Re: Taking action against scraping for hire

#114

Earlier quoted context omitted.

I've reread the previous comment and I really don't see where there is any justification stated for acting in an unethical manner. While Facebook may be making an argument against unethical behavior by a few, using the language they do is detrimental to legitimate uses of crawling content available on the Web. Corporations, by nature, work in a way that individuals at those companies don't. They are literally "non-co…

We should all be wary of corporate control and claim to rights built from their user base, especially if those services are offered for "free". That's fine then. And I agree with you. But leave you with this. Do. Not. Give. The. Company. Your. Data. They are literally "non-corporal" entities and work toward increasing profit and stakeholder value Again, I agree. But if you think this is a bad thing, then you don't be…

Yeah, nothing says "commie" like trust busting and keeping markets competitive.

Re: Taking action against scraping for hire

#115
post #9

Earlier quoted context omitted.

Logged in vs not logged in data.

> Logged in Is this actually private data, or is it public stuff that's become annoyingly hard to view anonymously because Meta chose to stick it behind a login box?

Anything behind a login gate is private data for that registered user only.

Re: Taking action against scraping for hire

#116
post #88

Earlier quoted context omitted.

if a bot creates the account, who breaches the contract?

The person who ran the bot. Programs do not have agency, they are just tools. That's like saying "If the gun fires the bullet, who is liable for murder?" It's a silly question.

> That's like saying "If the gun fires the bullet, who is liable for murder?" It's a silly question.

I don't know I've seen several people unironically argue that it should be the gun's manufacturer.

Re: Taking action against scraping for hire

#117

Earlier quoted context omitted.

Quoted post unavailable.

> I love to hate on Meta, but their actions here are spot on and make my morning very enjoyable as I sip my cup of coffee. You might want to reassess your intelligence there friend. It seems to be suffering from a common form of cogntive dissonance combined with some form of confirmation bias. How so? Well you clearly don't like scraping, otherwise you wouldn't be agreeing with a criminal... So there's the confirmati…

This is a total non-sequitur argument here. You've gone from accusing me of lack of intelligence to suffering from cognitive dissonance and confirmation bias, to Facebook's terrible security practices: simply because I'm pleased that an organization has taken action against Web scrapers for violation of Terms of Service.

Yes, I've gone on record indicating that I believe Web scraping to be generally unethical, and that I'm pleased that some action was taken against those that make it their business to do so. And that is all that I have stated in my OP. You've decided to take me on some circular mental gymnastics journey I'm still trying to wrap my head around.

Re: Taking action against scraping for hire

#118
post #102

Earlier quoted context omitted.

If simp is supposed to be short for simpleton, you might want to consider how simple your thoughts are.

It's not, see https://www.urbandictionary.com/define.php?term=Simp

I can also link to a source that's going to be biased in my favor: https://www.etymonline.com/word/simp

Re: Taking action against scraping for hire

#119
post #14
post #3

Data harvesting is moral for me, but not for thee.

It's their platform. Do you really want some random companies scraping your facebook and instagram posts?

As others said, there is no “you” in the scheme. It's Facebook's data. When people access that data without paying, they are “bad guys”. When the very same people pay for it, they are “legal partners”. In both cases they can do anything with it, while Facebook can't be held responsible because of all the official agreements. So as long as there is no specifically bad publicity or money loss anything goes either way.

“You” only exist in numerous empty statements about “privacy”, “respect”, etc. If you are feeling artsy, you can make that hyped NFT thing out of those, and see whether those kilobytes of text really worth anything.

Re: Taking action against scraping for hire

#120

Earlier quoted context omitted.

I've reread the previous comment and I really don't see where there is any justification stated for acting in an unethical manner. While Facebook may be making an argument against unethical behavior by a few, using the language they do is detrimental to legitimate uses of crawling content available on the Web. Corporations, by nature, work in a way that individuals at those companies don't. They are literally "non-co…

We should all be wary of corporate control and claim to rights built from their user base, especially if those services are offered for "free". That's fine then. And I agree with you. But leave you with this. Do. Not. Give. The. Company. Your. Data. They are literally "non-corporal" entities and work toward increasing profit and stakeholder value Again, I agree. But if you think this is a bad thing, then you don't be…

What a pretty picture capitalism is. Break out the popcorn for the latest regular installment of “ok for me but not for thee”:

People You May Know employs tons of shady stuff Facebook doesn’t reveal and has saved their bacon early on from stagnating at around 100M users.

https://mashable.com/article/people-you-may-know-facebook-cr...

Facebook Beacon and others had a big outcry. They got hauled into Congress multiple times. And of course whenever they get caught, they always throw a “mea culpa” and do it all over again in a year under a different name. Here they are recording faces of their users secretly using camera permisions!!

https://www.independent.co.uk/tech/facebook-app-recording-ca...

Their entire business model is “Give us all your data for free.” Mark Z early on was flabbergasted himself when he realized he no longer needed to scrape sites on Harvard’s house websites and could just ask people to submit the data for each other: “They ‘trust me’, dumb fucks.”

https://www.esquire.com/uk/latest-news/a19490586/mark-zucker...

Proceeds to build entire business on this data…

BUT THEN. Someone else does it to them and they get mad. “You can’t scrape us!” LinkedIn tried this:

https://www.zdnet.com/google-amp/article/court-rules-that-da...

And it’s not like capitalist enterprises even try to be consistent in their legal complaints:

https://9to5mac.com/2022/04/14/apple-calls-out-meta-for-hypo...

Post reply on HN