Earlier quoted context omitted.
Cows can be spherical.
And have uniform density.
Faking a JPEG
41–50 of 97 posts
Re: Faking a JPEG
#42I wonder if you could mess with AI input scrapers by adding fake captions to each image? I imagine something like: (big green blob) "My cat playing with his new catnip ball". (blue mess of an image) "Robins nesting"
A well-written scraper would check the image against a CLIP model or other captioning model to see if the text there actually agrees with the image contents.
Re: Faking a JPEG
#43[1]: https://www.ty-penguin.org.uk/robots.txt
[2]: https://www.ty-penguin.org.uk concatenated with /~auj/cheese (don't want to create links there)
Re: Faking a JPEG
#44Given that current LLMs do not consistently output total garbage, and can be used as judges in a fairly efficient way, I highly doubt this could even in theory have any impact on the capabilities of future models. Once (a) models are capable enough to distinguish between semi-plausible garbage and possibly relevant text and (b) companies are aware of the problem, I do not think data poisoning will be an issue at all.
There's no evidence that the current global DDoS is related to AI.
Re: Faking a JPEG
#45Love the effort. That said, these seem to be heavily biased towards displaying green, so one “sanity” check would be if your bot is suddenly scraping thousands of green images, something might be up.
Re: Faking a JPEG
#46> compression tends to increase the entropy of a bit stream. Does it? Encryption increases entropy, but not sure about compression.
I can see what was meant with that statement. I do think compression increases Shannon entropy by virtue of it removing repeating patterns of data - Shannon entropy per byte of compressed data increases since it’s now more “random” - all the non-random patterns have been compressed out. Total information entropy - no. The amount of information conveyed remains the same.
Re: Faking a JPEG
#47Love the effort. That said, these seem to be heavily biased towards displaying green, so one “sanity” check would be if your bot is suddenly scraping thousands of green images, something might be up.
Re: Faking a JPEG
#48Given that current LLMs do not consistently output total garbage, and can be used as judges in a fairly efficient way, I highly doubt this could even in theory have any impact on the capabilities of future models. Once (a) models are capable enough to distinguish between semi-plausible garbage and possibly relevant text and (b) companies are aware of the problem, I do not think data poisoning will be an issue at all.
There's no evidence that the current global DDoS is related to AI.
Re: Faking a JPEG
#49Reading about Spigot made me remember https://www.projecthoneypot.org/ I was very excited 20 years ago, every time I got emails from them that the scripts and donated MX records on my website had helped catching a harvester > Regardless of how the rest of your day goes, here's something to be happy about -- today one of your donated MXs helped to identify a previously unknown email harvester (IP: 172.180.164.102). Th…