Stay discoverable in search while disallowing AI training
21–30 of 59 posts
Re: Stay discoverable in search while disallowing AI training
#22A bit too late honestly (?). With so many people who have shifted over to reading AI summaries as a primary search response, those with AI-enabled sites will win by attrition. There is no going back from this. And the internet is a relatively new phenomenon. Recklessly, blindly applying ads to pages in hopes of generating revenue is a very silly thing to do. Technology with ad blockers and now AI summaries has taken…
All this attitude does is tear down the only viable income source for independent publishers and demonizes them for trying to make money, while everyone let's huge corporations off the hook for it because "well that's just what they do"
You aren't independent. You work for the BigTech company that serves ads on your site.
Re: Stay discoverable in search while disallowing AI training
#23Re: Stay discoverable in search while disallowing AI training
#24"Cloudflare classifies bots by behavior, and a single bot can exhibit more than one behavior." Is that really true CF classifies anyone not using a popular browser with Javascript enabled as a "bot" CF fingerprints www users As an example, look at CF's Permissions-Policy HTTP response header on a site with CF "bot protection", i.e., the "checking your browser" CAPTCHA nonsense (challenges.cloudflare.com). Then look a…
Re: Stay discoverable in search while disallowing AI training
#25Earlier quoted context omitted.
7 per minute. I use my high school’s website to test Internet connectivity bc the domain is short and they don’t do a TLS redirect (making it easy to detect WiFi portals).
neverssl.com
Re: Stay discoverable in search while disallowing AI training
#26If this admin is serious about AI growth they’d make anti-scrapping illegal.
Re: Stay discoverable in search while disallowing AI training
#27Earlier quoted context omitted.
7 per minute. I use my high school’s website to test Internet connectivity bc the domain is short and they don’t do a TLS redirect (making it easy to detect WiFi portals).
neverssl.com
I just want basically captive.apple.com with a shorter domain and the webserver not even listening on port 443 at all.
Surprised someone hasn't made this yet, it only requires one spare public IP.
Re: Stay discoverable in search while disallowing AI training
#28"Cloudflare classifies bots by behavior, and a single bot can exhibit more than one behavior." Is that really true CF classifies anyone not using a popular browser with Javascript enabled as a "bot" CF fingerprints www users As an example, look at CF's Permissions-Policy HTTP response header on a site with CF "bot protection", i.e., the "checking your browser" CAPTCHA nonsense (challenges.cloudflare.com). Then look a…
This is my biggest complaint about CF. They are implicitly supporting user-agent discrimination in favour of Big Browser, instead of discriminating on actual behaviour.
...and of course there are already companies running tons of VMs with "officially sanctioned" browser + OS stacks, that can get past all these "protections", for a fee.
"AI bots" is the newest boogeyman they came up with to take away freedom.
Re: Stay discoverable in search while disallowing AI training
#29"Accountable" is just a fancy word for "pinky promise, but with a label." Nothing stops the data from ending up in a training run once it's already been fetched.
[1] Analog hole and other workarounds aside, naturally.
Re: Stay discoverable in search while disallowing AI training
#30If this admin is serious about AI growth they’d make anti-scrapping illegal.