Can large companies not be faulted for ignoring robots.txt? Seems like something GDPR could enforce for personal(ly owned) sites?
Why do you think TikTok is ignoring robots.txt? This site is disallowing a lot of crawlers via User-Agent matching, but not ByteSpider. https://www.nerdcrawler.com/robots.txt The domain serving the images is allowing everything: https://res.cloudinary.com/robots.txt
Re: How we blocked TikTok's Bytespider bot and cut our bandwidth by 80%
#21I just assumed that would have been the first thing they tried…