Large-Scale Online Deanonymization with LLMs
31–40 of 258 posts
Re: Large-Scale Online Deanonymization with LLMs
#32Re: Large-Scale Online Deanonymization with LLMs
#33Re: Large-Scale Online Deanonymization with LLMs
#34Re: Large-Scale Online Deanonymization with LLMs
#35Earlier quoted context omitted.
To be clear, we are making a clear concession here that the people weren't truly anonymous. But we did use an LLM to remove any identifying information from HN making them quasi-anonymous, this is more described in the appendix Table 2. We do also make a more real world like test in section 2. There we use the anthropic interviewer dataset which Anthropic redacted, from the redacted interviews our agent identified 9/…
But you also relied on people giving away too much personal information about themselves... which won't always be the case.
Re: Large-Scale Online Deanonymization with LLMs
#36so if they put their linkedin account on their HN account, we can figure out who they are.... genius stuff, AI really is changing the landscape all right
To be clear, we are making a clear concession here that the people weren't truly anonymous. But we did use an LLM to remove any identifying information from HN making them quasi-anonymous, this is more described in the appendix Table 2. We do also make a more real world like test in section 2. There we use the anthropic interviewer dataset which Anthropic redacted, from the redacted interviews our agent identified 9/…
Edit: actually I've re-upped your submission of that link and moved the links to the paper to the toptext instead. Hopefully this will ground the discussion more in the actual study.
Re: Large-Scale Online Deanonymization with LLMs
#37I'm not sure the practical implications are as dramatic as the paper suggests. Most adversaries who would want to deanonymize people at scale (governments, corporations) already have access to far more direct methods. The people most at risk from this are probably activists and whistleblowers in jurisdictions where those direct methods aren't available, not average users.
Easier methods probably means more adversaries.
Re: Large-Scale Online Deanonymization with LLMs
#38And surprise, a tool made for processing text did it quite well, explaining the kind of phrase constructions that revealed my native language.
So maybe this is a plus for passing any text published on the internet through a slopifier for anonymization?
EDIT: deanonymization -> anonymization
Re: Large-Scale Online Deanonymization with LLMs
#39What this tells me is that major social media sites, some of which claim to be developing frontier models, have no excuse for a bots waging influence campaigns on their sites.
Re: Large-Scale Online Deanonymization with LLMs
#40Is there a deployment of this tool so that I test it on myself? EDIT: please someone build this, vibe-code it. Thanks