Live data from Hacker News

Vomit: Clean up Claude 5's token output with a separate LLM

github.com

21–30 of 315 posts

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#22

Which Claude 5? Opus 5 does seem to have diarrhea of the mouth. But Fable 5 hasn't been so bad for me. Or perhaps it is just better at adhering to my guidelines.

Pre Trump-castration Fable was verbose, but had a point, and used that wordiness to say or show the indeed intelligent things it reasoned about. This, whatever this is, is something else.-

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#23
I'm sorry but the whining over LLM output styles is embarrassing. Do Claude and GPT models always respond in exactly the way my most articulate coworker would? No. The overused jargon is absolutely annoying. But these things aren't my drinking buddies, they're professional tools. It's not _literally unreadable_. It's just not ideal. Most of my tooling is "not ideal". That's okay. That's what I'm paid for. I just work around it.

For me I added some instructions to speak clearly and it helped marginally and that's fine. There will be a new model out in a few weeks where I'm sure they've laser focused on this issue since nobody can shut the fuck up about it. The same thing happened with GPT if anyone can recall the ancient period of 4-6 months ago.

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#24
post #20

At some point one has to wonder if it's still worth using anthropic's models if we need to babysit 100% of its output with another vendor's model. Why not just use that other vendor's model for everything? I can't help but feel the circumstances that enable this kind of front page article are vestigial from the days when OAI was super bad and Anthropic was beyond reproach. This change-over-time is why I avoid getting…

individuals can blow in the wind, but if you are a company who bought a thousand seats and spent a ton of time training people up, establishing policies, vetting which extensions are allowed, the transition cost is much higher.

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#26
post #20

At some point one has to wonder if it's still worth using anthropic's models if we need to babysit 100% of its output with another vendor's model. Why not just use that other vendor's model for everything? I can't help but feel the circumstances that enable this kind of front page article are vestigial from the days when OAI was super bad and Anthropic was beyond reproach. This change-over-time is why I avoid getting…

> Why not just use that other vendor's model for everything?

Effectively all models can do style transfer reasonably well at this point, but not so much for "actual reasoning".

If the combination of two works better for you than each one by itself, why wouldn't you stack them like that?

Re: Vomit: Clean up Claude 5's token output with a separate LLM

#29
post #23

I'm sorry but the whining over LLM output styles is embarrassing. Do Claude and GPT models always respond in exactly the way my most articulate coworker would? No. The overused jargon is absolutely annoying. But these things aren't my drinking buddies, they're professional tools. It's not _literally unreadable_. It's just not ideal. Most of my tooling is "not ideal". That's okay. That's what I'm paid for. I just work…

Amen.

These things do work that previously would have taken expensive engineers months to do, at much lower quality, and what's our response? Ti nit pick on it being more verbose than we'd like?

Just like with humans, when someone is being too verbose, there's a skill to just filter through the noise and focus on the important parts.

This feels no different when I use an AI.

But I guess it's a good sign that we've from complaining about 'AI slop code' to, 'I don't like how it speaks to me'.

Post reply on HN