Live data from Hacker News

We should promote more personal indexing, rather than algorithmic indexing

news.ycombinator.com

1–10 of 23 posts

We should promote more personal indexing, rather than algorithmic indexing

#1
There are high quality original sources writing information today, just like there was before Web 2.0, when people would go on the internet to learn things and spend actual quality time reading blogs and personal content as well (which wasn't written for virality, but for expression).

However, while original sources (NASA, Reuters, bloggers, authors, scholarly journals) still write and publish (including to social media), the viral content makers, a second tier of people who write about what the original author wrote about, specializing themselves for social media, just write the grabbiest headline, and the most engaging (infuriating, polarizing, salacious) version of a piece of the original content, and when we go to our platforms, these are the content examples we see and read.

The original source's publications (posts) become almost invisible. The source becomes almost unknown as a source.

Engagement algorithms can't help this, because this is actually their purpose, which is anti-original content and -quality content.

We should design platforms and systems that allow people to index their own content again (as was the case before Web 2.0 when people manually put links to other websites, blogs, organizations, and articles on their own websites. This would make original and quality content writers become visible and indexed (even just in people's worldviews, not just indexed on the internet) and make viral content makers more invisible.

It wouldn't be total, because many people prefer the emotionality of viral content, but it would at least create an internet where there was more value available.

Re: We should promote more personal indexing, rather than algorithmic indexing

#4
The problem isn't black and white. Discussion adds value, if you didn't think so you wouldn't have posted here. so if not black and white then your talking about subtle and degree which is hard for a machine to determine. Also, most people engage with narratives/stories not facts.

I'm trying to downplay your comments, just its hard to start a social media company. there's a reason it's considered a tarpit idea.

Re: We should promote more personal indexing, rather than algorithmic indexing

#6
Proposing a fake moralizing solution to a non problem is pointless. You ignore incentives and structures that make the world the way it is and falsely assert a superior past that never existed.

Inventing a fake metric like "quality" and deprioritizing "engagement" is just messaging.

Re: We should promote more personal indexing, rather than algorithmic indexing

#7
>(NASA, Reuters, bloggers, authors, scholarly journals) still write and publish (including to social media), the viral content makers, a second tier of people who write about what the original author wrote about,

I like original primary sources for some topics and secondary sources for others. For programing topics, I'm ok reading the original papers.

But for other topics ... say "civil engineering" ... I prefer a "popularizer" like Grady Hillhouse's "Practical Engineering". His 15-minute presentations are the right amount of depth for exposing me to various city infrastructure topics. I'm not going to pretend I'd be interested in reading original scholarly journals from civil engineers. I deliberately outsource that to Grady. Hardcore engineers may complain that infotainment/edutainment is "shallow learning" but people have to strategically limit themselves to "shallow" explanations of some topics so they can spend more time to deep dive into other specialized areas of interest.

The "viral content makers" serve a useful purpose in the ecosystem to satisfy varying levels of interest. Therefore, a search engine that optimized for original academic papers instead of Grady blog posts when I ask "How does a city manage stormwater runoff?" -- would not be helpful to me in most cases. I dare say a "general" search engine that didn't put academic papers on page 1 of search results would be preferred by most people.

Re: We should promote more personal indexing, rather than algorithmic indexing

#8
post #5

So once I have auto generated a 10k accounts using modern generative ai and then used them to make my content go viral with apparently personal indexing, what's your next step?

This sort of purposeful injection of "bad" information has been possible rather cheaply across plenty of platforms that are open to use like this, even before modern generative ai. Some have suffered, clearly, but others have resisted it rather well for years. And they continue to be useful even with the new tech. For instance, how do you think high value targets like Wikipedia continue to resist this sort of behavior?

Perhaps considering that may bring some creative ideas to mind for you and other hackers, other than the pessimism that seems to underlie your take here (it seems not uncommon these days).

Re: We should promote more personal indexing, rather than algorithmic indexing

#9
post #7

>(NASA, Reuters, bloggers, authors, scholarly journals) still write and publish (including to social media), the viral content makers, a second tier of people who write about what the original author wrote about, I like original primary sources for some topics and secondary sources for others. For programing topics, I'm ok reading the original papers. But for other topics ... say "civil engineering" ... I prefer a "p…

What I really want to resolve this time issue is a reliable NLP processing toolchain capable of distilling primary sources into a summarized source, something we're very close to achieving these days.

This gets you the reliability of reading from a primary source, and the objectivity of machine summarization, with none of the capitalistic intent of the "viral human summarizers".

Re: We should promote more personal indexing, rather than algorithmic indexing

#10
post #5

So once I have auto generated a 10k accounts using modern generative ai and then used them to make my content go viral with apparently personal indexing, what's your next step?

In other words, "how do we fight spam?" Same as always -- with authentication and data analysis that's commensurate with available technology (i.e. at least whatever tools spammers have, including generative AI).

If you get past modern spam heuristics and have 10,000 phone numbers or government IDs or whatever makes for good auth these days, then kuddos. We are all fodder for The Red Queen hypothesis. The race continues.

Post reply on HN