Live data from Hacker News

The revolt of the reader

bcantrill.dtrace.org

201–210 of 311 posts

Re: The revolt of the reader

#201
post #160

I was so optimistic about using LLMs for "write once, read many" English language documents, but the more I've used the tools, the more pessimistic I get. More and more, I try to ask it for low prose responses because its writing just seems like such a low signal to noise ratio I'm curious about why LLM writing fails. Particularly whether LLM writing is fundamentally flawed, or if it's just distinctive and since it o…

> I'm curious about why LLM writing fails. Apart from the tasteless manipulations of the providers, it's mostly training data. LLMs output the average of their training data, and the overwhelming majority of humans are bad writers.

Nah, the deepest problem is the lack of intent. LLMs don’t have a message they are trying to convey or a clear picture of who the intended audience is.

They don’t know what you want to say solely based off a prompt, as it can’t possibly convey enough detail. And they can’t read your mind to fill in the gaps.

If it was just training data, that would actually be a much easier problem to solve.

Re: The revolt of the reader

#202
post #43

Earlier quoted context omitted.

This is very handwavy and dismissive. It is pretty safe to assume that must of us catch it most of the time because the simple fact is so many people just copy and paste whatever the LLM outputs without even trying to edit it or mask that they used one. We’ve all seen so many examples of the exact same cadence and verbiage that we’ve learned how to identify it pretty reliably. The ones who are “slipping past us” are…

> It is pretty safe to assume that must of us catch it most of the time because the simple fact is so many people just copy and paste whatever the LLM outputs without even trying to edit it or mask that they used one. This statement does not logically cohere. "We can spot it because so many people make it easy to spot." You don't see how this is just petitio principii in action?

If something is obvious then by definition it’s obvious. Additionally the formula is simple: AI + effort = writing we can generally tolerate (unless you’re just a bad writer). AI + no effort = writing most people can’t tolerate regardless of your skill as a writer.

LLM writing, unless someone puts in the time to improve it, typically follows the exact same patterns and favors the same words. You can’t reas a thread here without people talking about “Claude speak.” It is readily apparent, I do not need to show you a 10 year study with n=100,000 to make this point. The article isn’t hallucinating a problem, we are all nodding along because we all see it every goddamn day lmao. He even cited a study and makes a pretty strong case for why it’s a useful metric here in TFA.

If the AI writing is indistinguishable from human writing, then it is not lazy copy and pasting of AI outputs and isn’t just raw LLM output with no work done on it. So in that case it’s no longer a problem.

LLM’s cannot write in a natural, human way that distinguishes it from the typical LLM output on the first try. If they could, we wouldn’t have this problem. Maybe one day they will. Hell maybe it’ll even be next week. But currently they do not so I do not understand why we are having this discussion.

Re: The revolt of the reader

#203
post #187

Earlier quoted context omitted.

This argument doesn't work, the average majority of humans are bad at math but recent LLM aren't.

But people who are bad at maths are unlikely to be writing about maths. A crude example might be if you search for “2+2=” in the training data, you’re much more likely to find “4” as the next character. Obviously llms are far more complex than this, but I think this proves the point. The fact you had to add the “recent” qualifier there highlights that llms in general were bad and had to be provided with corrective ta…

> And they still can’t count the R’s in strawberry!

Really? I do not have the time to survey the modern LLMs to see if your assertion is correct, but if it is then I'm surprised; I would have thought that that one would have shown up so often in their training data that they would be able to answer that question, even if they would then be unable to (for example) count the R's in raspberry, or in some other word where "count the R's in _____" was not widely found in recent online discussion.

Re: The revolt of the reader

#204
post #120

My revolt is against the cognitive stress of reading generated text. A trope typically indicates I’m in for an uphill read. I recently read this William Zinsser quote that inspired a nickname for this: Clotted Claude [1]. > Nobody has made the point better than George Orwell in his translation into modern bureaucratic fuzz of this famous verse from Ecclesiastes: > > I returned and saw under the sun, that the race is…

I think Orwell did a great job there actually. Despite its bizarre look, the sentence is evocative and eloquent. It does make me get a clear mental image from the very first word. It leaves little room for roaming and guessing, as it firmly nails elements one by one, and, by the time I reach the end of the sentence, I get the full meaning almost immediately. This sentence is not randomly written; this is crafted with…

I agree!

Both versions are great: the former is poetic and grand and would fit right in in a fantasy text; the latter is dry and informative and requires much less mental effort to translate and extract meaning from (though still more than "normal" text.

I could imagine a third version that's clearer than the second and still nearly as poetic as the first.

Re: The revolt of the reader

#205
post #48

Earlier quoted context omitted.

AI writing just means "writing I don't like" now. Just like Nazi means whatever and whoever I politically disagree with. Words have lost their meaning.

You're right that Nazi doesn't mean Nazi anymore. It means neonazi / white supremacist / white nationalist, which is a much broader group of people that, for some baffling reason, are under the impression that people don't care about their fascism and racism anymore.

"You guys are fucking nazis" - said the drunk young guy to the police when he was arrested for fighting outside a bar.

Re: The revolt of the reader

#206
post #120

Earlier quoted context omitted.

I think Orwell did a great job there actually. Despite its bizarre look, the sentence is evocative and eloquent. It does make me get a clear mental image from the very first word. It leaves little room for roaming and guessing, as it firmly nails elements one by one, and, by the time I reach the end of the sentence, I get the full meaning almost immediately. This sentence is not randomly written; this is crafted with…

When I read the Orwell's version, I immediately had the same feeling I had when reading mathematical proofs. I hate so much to hunt the preceding text for anaphora resolution... It's just such a bad and pretentious way of writing. It's hostile to the reader with the side of flaunting author's superiority. It's like having to sit through a party with acclaimed academics: every single one is so full of themselves, they…

I agree that it's not great style for just conveying information clearly and plainly (e.g I'd hate to read documentation written like this) but I'd argue that here, the medium is part of the message.

It's supposed to sound dry and depressing and somewhat sterile.

> It's hostile to the reader with the side of flaunting author's superiority.

I think if someone deliberately obfuscates meaning to sound fancier when the point is to communicate information directly, sure, I'd agree. But writing is often art, and I think demanding effort from the reader is fair in that case. I wouldn't make a blanket generalization like that.

Re: The revolt of the reader

#207
post #35

I would love to use Pangram but they simply don’t allow signing up with my custom email domain. The error was “This email address can't be used for signup. Please use a different email.” I’m not about to create a Gmail is to use your service. To me the attack on the decentralized nature on Internet infrastructure is no less serious than the attack on the human provenance of writing itself.

> but they simply don’t allow signing up with my custom email domain.

Tried 4 different domains. 3 of email services of various kind. 4th one my private domain which has absolutely no email reputation because I use it only for internal emails and sending is not even possible.

In the end I dug out some old gmail address and tried to use that.

The error was always the same, there had been suspicious activity from that domain. So the message is definitely incorrect. Well, there could have been suspicious activity from some gmail address, but if they don't allow gmail I guess they don't want many customers.

Yeah, did not cover my tracks. The could easily notice that I was the same one trying to sign up repeatedly with different emails.

Re: The revolt of the reader

#208
post #68

Earlier quoted context omitted.

Am I the only one that finds the second one much easier to parse?

The first is of course from the King James Bible which for centuries was essentially a standard that all English speaking peoples aspired to. If you find that version difficult I would expect much literary writing before the 1940s also seems difficult. This is just to say I recommend reading the King James even if you are an atheist, as I am. I also have to say that the first strikes me as being written by someone th…

My guess is that if you were betting on a race, you would put your money on the swift to win, and "sometimes fast runners trip over" wouldn't seem as smart.

The rewrite stinks of a consultant's report where humor is frowned upon, business is serious, and there's no way we could write in plain words that the CEO got there by chance. Objective consideration distances the author compared to the subjective original I have returned and [I saw]. The original gives examples, the rewrite's contemporary phenomena is vague enough to avoid calling out the board of directors. A compelled conclusion is one the author is - reluctantly, you understand - forced into. Innate capacity leaves an escape hatch for a education from a good school and life experience to excuse the board again. It's not an honest rewrite of the same sentiment - a subjective take that life isn't fair and it's not just here, and us.

A plenitude of observations undertaken in a multitude of geographically and culturally diverse locations has convinced this author of the incompleteness of the following claims: races are won by the swift, battles are won by the strong, bread is earned by wisdom, riches are earned through applied understanding, and favours return to skilled persons. Absent are the effects of time and chance on all situations, of which experience has made abundantly clear. Other phenomena may too have their inputs, e.g. underhanded manipulation.

(Swiftness is still the best available predictor of who will win a race, though, and training cardio, muscles, diet, electrolytes, mental endurance, is the best available way to increase your chances of winning and not dropping out, tripping over from tiredness, or getting cramp, even though you can't change your innate ability or age, and time and chance happeneth to ye regardless).

Re: The revolt of the reader

#209

My revolt is against the cognitive stress of reading generated text. A trope typically indicates I’m in for an uphill read. I recently read this William Zinsser quote that inspired a nickname for this: Clotted Claude [1]. > Nobody has made the point better than George Orwell in his translation into modern bureaucratic fuzz of this famous verse from Ecclesiastes: > > I returned and saw under the sun, that the race is…

I came to this post and saw that Orwell had forgotten to mention money as the prerequisite of success, which was very unlike him.

Remember he's transcribing a scripture into bad modern prose, instead of riffing on it.

Re: The revolt of the reader

#210

> do you think readers can’t tell? No. I have good anecdata: readers cannot reliably distinguish my own prose from LLM-written one apart from cases where LLMs use odd metaphors or one of their specific patterns. I've been specifically experimenting with that.

I’ve seen several false (or apparently false) accusations of LLM authorship on HN/Lobsters.

However, we have to distinguish a few hypotheses:

1. No careful readers will notice when a piece is AI written.

2. Careful readers will generally not notice AI writing.

3. Everyone who writes comments on HN will reliably classify writing as AI or not.

Yes, 3 is not true, but Bryan’s point depends on something in the area of 2.

The ability to distinguish AI writing depends on having a good ear. For people who lack it, they either don’t notice and don’t care, or they make paranoid accusations against anything that is remotely non-standard (“you used an em-dash, you must be AI!”).

Post reply on HN