Live data from Hacker News

Three sites made 215,128 “best software” pages for AI. Perplexity cites them

trellner.com

61–70 of 265 posts

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#61
post #18

If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I've also experienced that both Claude and Codex routinely include generated we…

> always picks its own

If you hate AI writing enough, this turns AI filters into a kind of humiliation ritual. AI will derank normal business writing for human readers, and uprank inflated, verbose, tic-heavy slop. So you have to put the heavy slop out with your name on it. Really perverse moment.

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#62
post #6

It's difficult to read more than a few sentences, when this itself is clearly a Claude artifact.

Weird, there are 2 negative articles on the HN front page right now about Perplexity, both from “research” sites that are clearly LLM slop themselves.

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#63

Earlier quoted context omitted.

The other day, I remember an article was posted to HN about something, but it came from a company that provides SEO services to companies by doing something like this: 1. For a given company, analyze their target audiences and the questions they are likely to ask LLMs about. 2. For each such question, ask it to each of the major LLMs, and compute the KL divergence between the pages they want to rank for the question…

And it's all because of ads. The incentives in an ad-funded internet are just always going to lead to this sort of thing. The most important thing is getting the user to load your page, not actually satisfying their query. Let's hope the LLM model continues to be paying for credits, because any that move to ad revenue will become useless for real work.

> And it's all because of ads.

Not really. A while ago there was a news piece stating that Israel was behind a series of fake think-tanks with very accessible websites which were created with the express purpose of feeding AI agents with alternative facts aligned with their foreign policy.

If anyone has the link at hand, please post it.

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#66
post #50

Earlier quoted context omitted.

> asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored [...] It always picks its own ...is not the same as claiming... > LLMs favor LLM-generated passages over human written ones Here, you're using the same LLM to both produce and judge the resulting work. If anything, I would expect an LLM to tend to prefer its own work given that the same training is produc…

It's not intuitive to me for why preference for its own writing would emerge, and during what type of training or tuning. Perhaps something like: learning to identify what source files it has worked on by the code style alone, because tasks may give human code (public repos, etc) and ask to make changes.

It would need to be researched, but I wonder if it ends up being something that happens at the token level?

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#68
post #18

If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I've also experienced that both Claude and Codex routinely include generated we…

> always picks its own If you hate AI writing enough, this turns AI filters into a kind of humiliation ritual. AI will derank normal business writing for human readers, and uprank inflated, verbose, tic-heavy slop. So you have to put the heavy slop out with your name on it. Really perverse moment.

disagree that there is one kind of ranking and one kind of engine analyzing that ranking; sort of de-facto true that one company does run the ad world; strongly agree that this is a nightmare possibility and directly dystopian

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#69
post #53

> The result covers Perplexity only. We have not measured ChatGPT, Gemini, Copilot or Google’s AI Mode Why only test Perplexity...? Isn't it the least popular among these?

I assume because promising trustworthiness by sourcing information from the web is specifically Perplexity's shtick. The fact that this study undermines the quality of random web sources hits Perplexity's value proposition the most.

Re: Three sites made 215,128 “best software” pages for AI. Perplexity cites them

#70

Earlier quoted context omitted.

And it's all because of ads. The incentives in an ad-funded internet are just always going to lead to this sort of thing. The most important thing is getting the user to load your page, not actually satisfying their query. Let's hope the LLM model continues to be paying for credits, because any that move to ad revenue will become useless for real work.

> And it's all because of ads. Not really. A while ago there was a news piece stating that Israel was behind a series of fake think-tanks with very accessible websites which were created with the express purpose of feeding AI agents with alternative facts aligned with their foreign policy. If anyone has the link at hand, please post it.

https://www.theguardian.com/world/2026/aug/26/fake-thinktank...
Post reply on HN