There's a much more pressing issue.
I just asked ChatGPT to explain it, here's what it spouted:
"We need to securely record the knowledge created by humans up to this date using cryptographic hashes and a Merkle tree. This includes text, pictures, videos, and sound. This is vital for training future AIs and will ensure that the well is not poisoned. We can use SHA-3, SHA-256, Blake3, or any other cryptographic hash. We should store this in a large database, as well as sites and forums that will provide a "proof of existence" before a certain date."
It got a bit confused at the end. What I was suggesting was that there'd eventually be sites/forums/tools where every data you can access could check it's "proof of existence".
You'll then know for sure if the file was created before a certain date or not.
For example imagine I were to download a file, a tool could tell me (disclaimer for sensitive users, I'll mention the 'B' word):
"File xyz, with SHA-256 4f82d8dd...eb80fd7b, was first recorded on the Bitcoin blockchain in block 601382, in april 2023. We verified that using the Merkle tree XYZ3497 whose root hash was in that block."
It's not perfect but before we begin signing every new content we create, I think the most pressing matter is to "sign" as much of the past human knowledge as we can.
For what an AI will be able to do is to create and sign new content and use fake identities.
But what an AI will not be able to do is modify proof-of-existence of old documents. So we need, now, to establish proof-of-existence for as much old data as we can.