Live data from Hacker News

A peek into Reddit's anti-spam internals

lyra.horse

61–70 of 88 posts

Re: A peek into Reddit's anti-spam internals

#61
post #9

Earlier quoted context omitted.

There is some sort of wink wink nudge nudge agreement going on with certain spam accounts. You will see them post article spam with hidden history, and if you look up their posts either via google or any other reddit crawling tool, they are posting all over various subreddits that same article maybe dozens of times. If they comment it is really basic and formulaic and found all over their post histories as well. I fe…

This practice was perfected by gallowboob years ago. He would spam a link/pic/post and monitor, if the post didn’t gain traction, he would delete and post again as to not trigger protections against the same link being posted. He was a cancer on Reddit and I’m sure he still exists under different monikers. But now there are 100s of gallowboobs.

gallowboob in particular was an interesting case because he was very much a real person. and oh hell did a subreddit I mod know that way all too well

i think the guy had a like a keyword alert on his username because like one of my co-mods on a subreddit would talk about the guy and then we'd get reports for "It's targeted harassment against me" (which are reports that are sent to the admins) like a few hours later. much to the dismay of him, we had a chat with the admins later and it was like "as long as you're not saying to do vote manipulate or harass the guy it's fine."

i think a lot of it came from the fact that so like if you're modding a subreddit, a lot of people spend their time in the modqueue view rather than the comments so you see the targeted harassment reports on "xyz is a meanie head" and just click "remove" because it already is on the edge at best for most subreddits. this is how context gets lost. so people would see "unfavorable treatment" (not that it didn't happen, gallowboob's company's domain was soft-banned on reddit yet his subreddits had automod rules set to approve them) when if more people were as trigger happy on the report button a similar thing would happen

the admin problems with this are much worse because the comments tend to be looked at in isolation so saying "i'm gonna kill you", in isolation, looks without context pretty bad, but might be part of a joke chain or song meme that reddit likes to do every so often. take into account the fact that admins get whiny sometimes if your AEO removals are too high. then take into account the AEO guy's Tarot card reading and whether Mercury is in retrograde and you get a lot of mods who are a bit trigger happy, esp when people've gotten banned for approving stuff the AEO removed for dumb reasons

this somewhat led to a bit of an inflated ego with regards to reddit but eventually from what i see he left... at least under that username anyway.

Re: A peek into Reddit's anti-spam internals

#62
post #51

Earlier quoted context omitted.

I'd be surprised if Reddit wasn't selling tools specifically aimed at seting up and managing operations like that

I would be surprised if Reddit was selling tools like that directly. Rather than just look the other way - because uncovering modern bot operations is a thankless threadmill, and user engagement metrics from fake users are still user engagement metrics.

to be fair, i once accidentally ended up on shreddit or new reddit or whatever they call it nowadays and i think there's something for managing your posts on reddit and seeing analytics about that or whatever

> uncovering modern bot operations

this significantly overestimates how sophisticated the spam waves are compared to like ability. the 80% of spam filtering basically never was really done as far as i can tell.

> a thankless threadmill, and user engagement metrics from fake users are still user engagement metrics.

that's probably it tho

Re: A peek into Reddit's anti-spam internals

#63
My takeaway:

> My test account (5 years old!) got banned immediately, and all of its post history got wiped too. RIP

I want to know the real string of the event I ever want to delete my account and content. This would be much faster than using a browser script to manually delete.

Re: A peek into Reddit's anti-spam internals

#64
post #10
post #6

Earlier quoted context omitted.

You're missing the second-chance pool https://news.ycombinator.com/pool which allows certain posts to reappear as if they were new.

Once again expressing my opinion that this is the worst anti feature of the site. Threads are like commenting in the void because most people are not going to be looking for replies to comments they made days or weeks ago. I see the true datestamp of the comment I replied upon upthread was not 3 hours ago, but 3 days ago. The fact that they change the timestamp is also very stupid (yes you can hover and still return…

[deleted]

Re: A peek into Reddit's anti-spam internals

#65

I cannot emphasize strongly enough just how deeply pervasive the spam is at Reddit. I'm a mod at the ecommerce subreddit, and I've only caught some of the AI-powered marketing operations because in one particular campaign that was making fictional claims about things I had direct knowledge about. Once I looked into the post history, and started to untangle the web of accounts that formed a self-supporting community o…

How did you manage to uncover the operation? Is there any things that tipped you off? Is it like, accounts only posting on a precise subset of subs, or how much a network of accounts reply to each other? Or too much linking/references to products, or perfect spelling/identical consistent typos?

I'm curious because it feels like it could be built into a tool to analyse - even if it does become a bit of an arms race.

Re: A peek into Reddit's anti-spam internals

#67

I cannot emphasize strongly enough just how deeply pervasive the spam is at Reddit. I'm a mod at the ecommerce subreddit, and I've only caught some of the AI-powered marketing operations because in one particular campaign that was making fictional claims about things I had direct knowledge about. Once I looked into the post history, and started to untangle the web of accounts that formed a self-supporting community o…

How did you manage to uncover the operation? Is there any things that tipped you off? Is it like, accounts only posting on a precise subset of subs, or how much a network of accounts reply to each other? Or too much linking/references to products, or perfect spelling/identical consistent typos? I'm curious because it feels like it could be built into a tool to analyse - even if it does become a bit of an arms race.

It's harder now that post history is hidden but I did something similar in the past.

Start with a single comment you think is a shill. Maybe 80% of their post history is vague inane generalities (like "aww so cute" in r/cats, or a reply to a top rated post that only paraphrases the existing context without adding anything new). You can use an LLM to identify every comment or post from that account that mentions a product or service. Take note of everyone who replies to that comment as well as the parent comment. Then use an LLM to identify every post from that original account asking for recommendations (hey r/bidet, what's your favorite bidet), and look at who responds. If you build this graph, draw directional edges based on who replied to who. The accounts with edges both ways across different posts are bots. Rinse and repeat by examing the post history of THOSE accounts. You will end up with a graph with a few loosely connected nodes (maybe false positives) but a tight web of spam accounts that frequently engage with each other.

That's your bot farm. This would be relatively trivial for reddit to implement, if they cared about reducing spam. I got a POC working in a few hours, back before they limited API access

Re: A peek into Reddit's anti-spam internals

#68
post #49

Earlier quoted context omitted.

My guess would be the second-chance pool ( https://news.ycombinator.com/pool , explained by dang at https://news.ycombinator.com/item?id=26998308 )

Exactly that, you can tell by hovering over the posted time "9 hours ago".. and see that it's really from June.

yeah but it was on the frontpage and quite popular last time! and the timestamp on the comments changed too, to be relative to the new "posted" timestamp.

Re: A peek into Reddit's anti-spam internals

#69

Earlier quoted context omitted.

I remember reading years ago about some corrupt mod in one of the image subreddits - he or his friend had started some image hosting site and had six different Reddit accounts that he used to upvote posts that used his site and downvote all other posts. It took people a long while to notice what he was up to.

And now automate and scale that with Claude/OpenAI/Gemini/whatever. It's insidious and terrible.

Why would you need an LLM to automate that?

Re: A peek into Reddit's anti-spam internals

#70

Earlier quoted context omitted.

How did you manage to uncover the operation? Is there any things that tipped you off? Is it like, accounts only posting on a precise subset of subs, or how much a network of accounts reply to each other? Or too much linking/references to products, or perfect spelling/identical consistent typos? I'm curious because it feels like it could be built into a tool to analyse - even if it does become a bit of an arms race.

It's harder now that post history is hidden but I did something similar in the past. Start with a single comment you think is a shill. Maybe 80% of their post history is vague inane generalities (like "aww so cute" in r/cats, or a reply to a top rated post that only paraphrases the existing context without adding anything new). You can use an LLM to identify every comment or post from that account that mentions a pro…

It's a good idea. But this would no longer work now that Reddit hides post history, right? Or does the API still provide a user's post history?

You'd basically need to be pulling all the data for all the subreddits, and then recreate a user's partial post/comment history from that.

Post reply on HN