Live data from Hacker News

New accounts on HN more likely to use em-dashes

marginalia.nu

181–190 of 643 posts

Re: New accounts on HN more likely to use em-dashes

#181
post #165

The data is available in a SQLite database on GitHub: https://github.com/vlofgren/hn-green-clankers You can explore the underlying data using SQL queries in your browser here: https://lite.datasette.io/?url=https%253A%252F%252Fraw.githu... (that's Datasette Lite, my build of the Datasette Python web app that runs in Pyodide in WebAssembly) Here's a SQL query that shows the users in that data that posted the most comm…

I still call voodoo on this. I use an iPhone, iPad, Mac to comment here—all of them autocorrect to em dashes at one point or another. Same goes for ellipsis.

Re: New accounts on HN more likely to use em-dashes

#184
post #50

I'm still salty that I can't use em-dashes anymore for fear of my writing being flagged as AI generated. Been using them for years—it's just `alt+shift+-` on a Mac keyboard and I find them more legible in many fonts compared to the simple dash on the typical numpad. It's so sad to me that good typographical conventions have been co-opted by the zeitgeist of LLMs.

LLM fatigue is real. It's not just em-dash — it's the overall tone of the writing that clues people in. But if your viewpoints and approach are unique, your typesetting won't raise suspicion of machine-generation, except in the most dull of readers. Just be you and it will be fine.

If you'd like more tips on writing I'd be happy to help.

Re: New accounts on HN more likely to use em-dashes

#185
post #48

One pattern I've noticed recently is sort of formulaic comments that look okish on their own, maybe a bit abstract/vague/bland, and not taking a particular side on good/bad in the way people like to do, but really obviously AI when you look at the account history and they're all the same formula: >this is [summary] >not just x, it's y >punchy ending, maybe question Once you know it's AI it's very obvious they told it…

Yeah, and some of them already have enough karma to downvote you if you call them out, which is infuriating…

Re: New accounts on HN more likely to use em-dashes

#186

Fwiw I did some more comparisons, looking for words disproportionately favored by noob comments: word noob new p-value ---------------------------- ai 14.93% 7.87% p=0.00016 actually 12.53% 5.34% p=1.1e-05 code 11.47% 6.04% p=0.00081 real 10.93% 2.95% p=2.6e-08 built 10.93% 2.11% p=2.1e-10 data 8.93% 3.51% p=6.1e-05 tools 7.6% 2.67% p=5.5e-05 agent 7.47% 2.95% p=0.00024 app 7.2% 3.09% p=0.00078 tool 6.8% 1.83% p=8.5e…

It's funny - some months ago I noticed that I use the word "actually" lot, and started trying to curb it from my writing. Not for any AI-related reason, but because it is almost always a meaningless filler word, and I find that being concise helps get my points across more clearly.

e.g. "The body of the template is parsed, but not actually type-checked until the template is used." -> "but not typechecked until the template is used." The word "actually" here has a pleasant academic tone, but adds no meaning.

Re: New accounts on HN more likely to use em-dashes

#187
post #48

One pattern I've noticed recently is sort of formulaic comments that look okish on their own, maybe a bit abstract/vague/bland, and not taking a particular side on good/bad in the way people like to do, but really obviously AI when you look at the account history and they're all the same formula: >this is [summary] >not just x, it's y >punchy ending, maybe question Once you know it's AI it's very obvious they told it…

The user [1] you've mentioned has 160 points being a poster of total four bland messages. This goes against a normal statistical distribution. And this gives away why they do it: the long-term aim is to cultivate voting rings to influence the narratives and rankings in the future. For now, this is only my theory but it may be a real monetization strategy for them.

[1] https://news.ycombinator.com/threads?id=snowhale

Re: New accounts on HN more likely to use em-dashes

#189

Earlier quoted context omitted.

What motivation is there to use AI to astroturf (if that's what this is) like this? Is it ideological? Is it product marketing in those relevant threads where someone is showcasing? Or is it pure technical testing, playing around?

In some cases, it's probably to establish aged accounts that are more trusted by users and spam algorithms. There's a market for old Reddit accounts, for example.

Yep. Like I said elsewhere on the thread, some of them already have enough karma to downvote.
Post reply on HN