Live data from Hacker News

A Facebook crawler was making 7M requests per day to my stupid website

coding.napolux.com

101–110 of 416 posts

Re: A Facebook crawler was making 7M requests per day to my stupid website

#102
post #74
post #69

How do you monetize a robot?

That's a good question. Even better if Facebook can pay me some money :P

dont know whether the facebook bots execute javascript - https://www.rdegges.com/2017/how-to-monetize-your-website-wi...

Re: A Facebook crawler was making 7M requests per day to my stupid website

#104
post #89

Earlier quoted context omitted.

Hey! Facebook engineer here. If you have it, can you send me the User-Agent for these requests? That would definitely help speed up narrowing down what's happening here. If you can provide me the hostname being requested in the Host header, that would be great too. I just sent you an e-mail, you can also reply to that instead if you prefer not to share those details here. :-)

I'm not sure I'd publicly post my email like that, if I worked at FB. But congratulations on your promotion to "official technical contact for all facebook issues forever".

Don't think I used my email for anything important doing my time at FB. If it gets out of hand he could just make a request to have a new primary email made and use the above one for "spam"

Re: A Facebook crawler was making 7M requests per day to my stupid website

#105

Earlier quoted context omitted.

When someone does this to Facebook its malicious and they go to jail. WHen Facebook does it to someone else... "oops".

From the point of view of the law, the intent (malicious or benevolent) is often as important as the action itself.

Nah, it only matters if you are a large corporation or not. If you are a large corporation you will ruin your opponent financial with legal fees (whether they are guilty or not). If you are a small timer going up against a large corporation they will delay the court case until you can not afford the legal fees.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#106

Earlier quoted context omitted.

When someone does this to Facebook its malicious and they go to jail. WHen Facebook does it to someone else... "oops".

Never forget Aaron Swartz. All he did was send download requests to JSTOR. https://www.youtube.com/watch?v=9vz06QO3UkQ

I'm going to get a ton of hate for this, but he was doing it with the intent to redistribute the content for free. He didn't own the content. There's a big difference. That being said it's very sad what came about of that. I really don't think the FBI needed to be involved.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#107
post #106

Earlier quoted context omitted.

Never forget Aaron Swartz. All he did was send download requests to JSTOR. https://www.youtube.com/watch?v=9vz06QO3UkQ

I'm going to get a ton of hate for this, but he was doing it with the intent to redistribute the content for free. He didn't own the content. There's a big difference. That being said it's very sad what came about of that. I really don't think the FBI needed to be involved.

> he was doing it with the intent to redistribute the content for free

There is actually no proof of this.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#109
post #106

Earlier quoted context omitted.

I'm going to get a ton of hate for this, but he was doing it with the intent to redistribute the content for free. He didn't own the content. There's a big difference. That being said it's very sad what came about of that. I really don't think the FBI needed to be involved.

> he was doing it with the intent to redistribute the content for free There is actually no proof of this.

Really? I felt that what happened to Aaron was a tragic injustice on many levels (from the fact that he was charged at all to the number of charges they threw at him), but for some reason I had always thought that it was fairly well-known that that was his intent.

I have no idea where I got that impression from, but I do recall reading several articles about him, as well as his blog around that time. It's possible I just internalized others' assumptions about his behavior, but it certainly seemed like an in-character thing for him to to.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#110

Make sure you file a bug - there are a myriad of sources internally, but we can often hunt it down easily enough (assuming it gets triaged to eng). Important info is the host of the urls being crawled, and the User Agent attached to the requests (headers are also good too). A timeseries graph of the hits (with date & timezone specified) can also help.

Why is this the owner's problem? Someone at Facebook should be filing the bug, or better yet instrumenting their systems so that incidents like this issue a wake-up-an-engineer alert. Fuck this culture of "it's up to the victim of our fuckup to file a bug report with us".

100%
Post reply on HN