Live data from Hacker News

Vomit: Clean up Claude 5's token output with a separate LLM

github.com

211–220 of 315 posts

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#211
post #204

Earlier quoted context omitted.

I've noticed that most people seem to consider the core problem of Claude's output as "too verbose" but I don't think this actually cuts to the heart of the matter at all. It's almost, in some weird way, the opposite: like the text is far too _dense_. It tries too hard to invent odd terminology to try to condense stuff, but it doesn't tell you up front that it is going to call your company wide error-handling mechani…

Kind of both. On the one hand, it is “verbose” in the sense that it will tell me every little nit that it can think of while doing a task, it will tell me a narrative about its thought process, and it will tell me every other detail it can think of. But it does so in a way that tries to be incredibly dense to the point that I have to struggle to figure out what it is saying. I wonder if there are any “legibility benc…

There are no "best" prompts. Its a random BS generation machine that you can at times direct enough to get stuff done for you. The output will almost always have varying levels of BS that you have to clean up with various levels of effort.

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#213
post #81

I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…

I'm probably going to be going against the grain here, but I think it's not as bad as it looks at first. I was similarly frustrated a few months ago, but have noticed I've started to learn the idiom. Its use of "dense jargon" and "stilted metaphor" is actually surprisingly consistent - it's speaking its own dialect, and you get used to it. After a while it gets much easier to read and even becomes somewhat efficient,…

Indeed I think much is shared across sessions and projects. I say we learn The Machine Vernacular [1].

[1] https://www.themachinevernacular.net/

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#214
post #204

Earlier quoted context omitted.

I've noticed that most people seem to consider the core problem of Claude's output as "too verbose" but I don't think this actually cuts to the heart of the matter at all. It's almost, in some weird way, the opposite: like the text is far too _dense_. It tries too hard to invent odd terminology to try to condense stuff, but it doesn't tell you up front that it is going to call your company wide error-handling mechani…

Kind of both. On the one hand, it is “verbose” in the sense that it will tell me every little nit that it can think of while doing a task, it will tell me a narrative about its thought process, and it will tell me every other detail it can think of. But it does so in a way that tries to be incredibly dense to the point that I have to struggle to figure out what it is saying. I wonder if there are any “legibility benc…

I quite like https://surgehq.ai/benchmarks/hemingway-bench

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#216
post #20

At some point one has to wonder if it's still worth using anthropic's models if we need to babysit 100% of its output with another vendor's model. Why not just use that other vendor's model for everything? I can't help but feel the circumstances that enable this kind of front page article are vestigial from the days when OAI was super bad and Anthropic was beyond reproach. This change-over-time is why I avoid getting…

Sol is like this too...

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#217

Looks like a wrapper around this prompt: You are an editor. You'll be given a message with strange characteristics: - Weird subject and verb combinations - Subjects that should be objects - Very roundabout reasoning, peppered with pseudo-epiphanies - A distracting beat to the flow of the message - Self-praise Remove these characteristics, and rewrite it in a clear, conversational style. Keep the intent of the message…

>every llm product that isn't a claude endpoint is going to be what you stuff into the claude etc endpoint

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#218

Earlier quoted context omitted.

Kind of both. On the one hand, it is “verbose” in the sense that it will tell me every little nit that it can think of while doing a task, it will tell me a narrative about its thought process, and it will tell me every other detail it can think of. But it does so in a way that tries to be incredibly dense to the point that I have to struggle to figure out what it is saying. I wonder if there are any “legibility benc…

There are no "best" prompts. Its a random BS generation machine that you can at times direct enough to get stuff done for you. The output will almost always have varying levels of BS that you have to clean up with various levels of effort.

" doesn't have a largest city because all of the cities in are small"

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#219

I suspect that sustained reading of Opus 5's unconscionably bad prose could actually cause psychological harm. We're strongly considering moving all of our Anthropic spend to Codex/open weight models. It's a mental health decision at this point.

I switched to GPT 5.6 Sol yesterday and it has been a joy going through and fixing up the Claude cruft. And being able to read everything the model says. A breath of fresh air for sure!

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#220

Earlier quoted context omitted.

So many vacuous statements at the seam. This is the hermetic load bearing part, which I confirmed rather than assuming.

Is this because they changed the word probabilities to allow for identifying AI text? If so, I don't need a computer to tell me when something is AI. It's crazy obvious from odd word choices.

It's mode collapse from RLVR.

That and if you talked to the same one person's frozen brain upload all day, you'd see the same catchphrases used too.

Post reply on HN