Live data from Hacker News

A Facebook crawler was making 7M requests per day to my stupid website

coding.napolux.com

141–150 of 416 posts

Re: A Facebook crawler was making 7M requests per day to my stupid website

#142
post #140

Hi Napolux, It looks like your site is using a theme based on my website ( https://ruudvanasseldonk.com/ , source at https://github.com/ruuda/blog ). That is fine — it is open source after all, licensed under the GPLv3. But I can’t find the source code for your site, and I can’t find any prominent notices saying that you modified my source. Could you please add those?

Despite how it sounds, I ask this with zero judgment and pure curiosity.

Why do you care?

Re: A Facebook crawler was making 7M requests per day to my stupid website

#143
post #140

Hi Napolux, It looks like your site is using a theme based on my website ( https://ruudvanasseldonk.com/ , source at https://github.com/ruuda/blog ). That is fine — it is open source after all, licensed under the GPLv3. But I can’t find the source code for your site, and I can’t find any prominent notices saying that you modified my source. Could you please add those?

Kind of an odd request for an open source wordpress blog theme imo

Re: A Facebook crawler was making 7M requests per day to my stupid website

#144
post #140

Hi Napolux, It looks like your site is using a theme based on my website ( https://ruudvanasseldonk.com/ , source at https://github.com/ruuda/blog ). That is fine — it is open source after all, licensed under the GPLv3. But I can’t find the source code for your site, and I can’t find any prominent notices saying that you modified my source. Could you please add those?

They look similar at a glance (border-top + Calluna font), so he might have taken inspiration from yours, but doesn't seem to have used any of your assets - the styles are clearly different and based on the WP 'BlankSlate' theme.

(my personal blog had a top border like that a decade ago, when styles on the body were a novelty :))

Re: A Facebook crawler was making 7M requests per day to my stupid website

#145

Someone purporting to be working in SEO and getting hit by 7M requests, but offers not real proof and cannot link the webpage either? Am I just getting too suspicious? Edit: Was trying to make a (bad) joke.

> Am I just getting too suspicious?

Yes, I think so. What's the conspiracy supposed to be here?

Re: A Facebook crawler was making 7M requests per day to my stupid website

#146
post #111

Did anyone notice the branded IP address? 2a03:2880:20ff:d::face:b00c "face:b00c"

FB has 2a00::/12 (and probably other blocks?) and can make something that looks like its company name from [0-9a-f]. Why wouldn't it do something harmless and fun like this? It's not as if it requires special dispensation or is breaking any rules.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#147

Make sure you file a bug - there are a myriad of sources internally, but we can often hunt it down easily enough (assuming it gets triaged to eng). Important info is the host of the urls being crawled, and the User Agent attached to the requests (headers are also good too). A timeseries graph of the hits (with date & timezone specified) can also help.

Why is this the owner's problem? Someone at Facebook should be filing the bug, or better yet instrumenting their systems so that incidents like this issue a wake-up-an-engineer alert. Fuck this culture of "it's up to the victim of our fuckup to file a bug report with us".

Not just “file a bug report” but also compile a time series graph (lol) and then pray that Facebook triages it correctly, which they have no incentive to do.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#148

Someone purporting to be working in SEO and getting hit by 7M requests, but offers not real proof and cannot link the webpage either? Am I just getting too suspicious? Edit: Was trying to make a (bad) joke.

Ask me anything. That specific website is a personal project of mine which I'd like not to disclose. As you can see from my blog I'm not selling anything, I don't even have a banner on my blog, so where your suspect is coming from?

Sorry, I was mostly joking. I definitely do believe Facebook could be responsible for this.

Re: A Facebook crawler was making 7M requests per day to my stupid website

#149
post #140

Hi Napolux, It looks like your site is using a theme based on my website ( https://ruudvanasseldonk.com/ , source at https://github.com/ruuda/blog ). That is fine — it is open source after all, licensed under the GPLv3. But I can’t find the source code for your site, and I can’t find any prominent notices saying that you modified my source. Could you please add those?

Despite how it sounds, I ask this with zero judgment and pure curiosity. Why do you care?

Well I respect the question of GP if nothing then for principle. People should get their shit together and start respecting the license. The state today is horrendous and the culture is evolving into less and less respect of the license. If people don't follow them personally they wont care to follow them professionally either

Re: A Facebook crawler was making 7M requests per day to my stupid website

#150
post #140

Hi Napolux, It looks like your site is using a theme based on my website ( https://ruudvanasseldonk.com/ , source at https://github.com/ruuda/blog ). That is fine — it is open source after all, licensed under the GPLv3. But I can’t find the source code for your site, and I can’t find any prominent notices saying that you modified my source. Could you please add those?

Despite how it sounds, I ask this with zero judgment and pure curiosity. Why do you care?

Please read https://sfconservancy.org/copyleft-compliance/principles.htm.... My understanding is that if you do not enforce your copyright (or copyleft in this case) you can lose the copyright.
Post reply on HN