Live data from Hacker News

12k AI-generated blog posts added in a single commit

github.com

131–140 of 167 posts

Re: 12k AI-generated blog posts added in a single commit

#131
post #69

Earlier quoted context omitted.

This will end with the only way to authenticate the people we're talking to is meeting them at the coffeeshop in the morning.

Did you forget about Blade Runner?

... Did you ...?

Re: 12k AI-generated blog posts added in a single commit

#132
post #127

They even have a scrolling 5-star reviews section, clearly generated: https://oneuptime.com/#reviews-title https://github.com/OneUptime/oneuptime/commit/538e40c4ae724e... https://github.com/OneUptime/oneuptime/commit/2bc585df20e6bb... You can fabricate a professional business image in a few days with AI now. It's going to be hard to build an honest brand when everyone is going to point and say "vibe coded slop" becau…

"This enhancement improves the user experience by showcasing positive feedback from customers"

you can't make this up

Re: 12k AI-generated blog posts added in a single commit

#133

"Nawaz Dhandala"

I nominate Nawaz Dhandala as "the king of AI slop"

He's just an idiot doing it in public, because there are people generating hundreds of posts a day for years now without committing it on github under their real name.

Re: 12k AI-generated blog posts added in a single commit

#134

Earlier quoted context omitted.

> Who is "we"? Definitely not Google or any other major tech company, they're all actively encouraging this. Google has been fighting aggressively to replace its search results with snippets, now generated by LLMs, to avoid sending traffic to other websites. If they continue, they will basically lead Google Search to a tipping point where a good competitor can take this market by storm. Microsoft also believed Window…

The fact is what people really want from a search engine is a single perfect result that answers their query exactly. An LLM does the 'single result' bit, but it's dubious whether or not it's a perfect answer. Most of the time that's probably not very important so long as the answer satisfies the search enough that the user is happy. Google is trying to turn Search into that product e.g. the single answer to a given…

> Most of the time that's probably not very important

Well... Maybe, but what's the point of an answer if you can't trust it? For ultra-fast answers for unimportant stuff I keep Cerebras tab open.

Re: 12k AI-generated blog posts added in a single commit

#135

It's becoming much harder to determine on a daily basis what content is original, thought-out by a person, and trustworthy. Ironically, verifiably-old content is easier to trust now. Examples from recent personal experience: 1) Some time ago I was searching for growing information about a specific and uncommonly-grown plant, and was led to a top-ranked website with long pages containing everything about it, including…

You're making the classic mistake of looking for a trustworthy information source and then trusting it, instead of focusing on whether the information itself is trustworthy regardless of source. It's literally the same as my grandma saying "they said so on TV, therefore it must be true" while completely dismissing anything I've read on the internet because reasons. If you develop the skill of judging information by i…

Well, if not disclosed you could assume that somebody did due diligence for you, and could include sources. I don't even trust LLM even if all the information is included in the context window if I need reliable information. Trying to make money on slop is really bad manners. It's a scam, you can't call it otherwise. Btw, I like AI, it did a ton of value for me. We just need to find a way to live with it, without getting doomed in misinformation.

Re: 12k AI-generated blog posts added in a single commit

#136

It's becoming much harder to determine on a daily basis what content is original, thought-out by a person, and trustworthy. Ironically, verifiably-old content is easier to trust now. Examples from recent personal experience: 1) Some time ago I was searching for growing information about a specific and uncommonly-grown plant, and was led to a top-ranked website with long pages containing everything about it, including…

You're making the classic mistake of looking for a trustworthy information source and then trusting it, instead of focusing on whether the information itself is trustworthy regardless of source. It's literally the same as my grandma saying "they said so on TV, therefore it must be true" while completely dismissing anything I've read on the internet because reasons. If you develop the skill of judging information by i…

No it's not the same as your grandma. The point is that it's now more expensive to find the correct information to learn from. You don't know it's an LLM ahead of time, and you may spend hours until you figure out something is off. Hence why reputable sources will become more valuable.

> If you develop the skill of judging information by its merit rather than source..

Did you read example #1? I'm not talking about some piece of code from an LLM that you can verify or some political opinion that you can take with a grain of salt, but information that you can only gain and/or judge through expertise:

If you're not a physicist yourself, you can't judge "information by its merit" on specific physics topics, because you don't have a solid baseline.

Similarly, in growing plants, each plant has its own peculiarities, and only people experienced in growing it can tell you anything useful - it's knowledge accumulated by trial and error. Not knowledge that your "great discerning mind" can assess on its own. Even a botanist can't tell you the ideal growing conditions of a plant that they've never studied before.

Re: 12k AI-generated blog posts added in a single commit

#137
post #90

I suspect we'll address this by just going back to older ranking algorithms for search. We'll go back to the primary signal of good content being links from trusted sources. People gaming the content based algorithms will eventually cause their own downfall.

> I suspect we'll address this Who is "we"? Definitely not Google or any other major tech company, they're all actively encouraging this. > trusted sources. What trusted sources are there that haven't yet been taken over by AI?

[deleted]

Re: 12k AI-generated blog posts added in a single commit

#138

Ironically due to slop I feel like we are regressing as a civilization 2020, want to know how to use Redix for Redis connections in Elixir? Google it and the results were most likely high quality, written by senior engineers who knew what they were doing Today google that, and it will be endless amounts of slop

For some searches I've started to limit the date range to pre 2023. That drastically improves search results (DDG, but I imagine Google as well). As long as you're looking for more long term information/posts ofc.

Re: 12k AI-generated blog posts added in a single commit

#139
post #43

Earlier quoted context omitted.

Now that we have better ML, maybe we could take "link sentiment" into account too.

[dead]

Crawlers would need to use backlinks but also rank vector similarity to ensure the linked content matches the linked intent. Some kind of rainbow shades of how relevent the link is to the linkee and reverse.

Re: 12k AI-generated blog posts added in a single commit

#140
post #60

I suspect we'll address this by just going back to older ranking algorithms for search. We'll go back to the primary signal of good content being links from trusted sources. People gaming the content based algorithms will eventually cause their own downfall.

I don't have a ton of hope just yet because I think it's still an incentives problem rather than a technical one. I got tired of the increasing AI slop in my YouTube Music feed and switched to Deezer a few months ago. Since then, not a single AI artist I've been able to spot. If a relatively marginal player like that can manage it, why can't Spotify or YTM? My suspicion is simply that Deezer actually actually tries.…

Its not a technical problem.

Its a public good we refuse to turn into a government service for nebulous reasons.

Post reply on HN