I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…
> The baked in communication style of these models is so obnoxious it's impacting my work. The best way I can describe it is that everything is optimized to impress the user and make the agent sound more authoritative, but the way this is done is through deliberate obfuscation, inserting inappropriate and extremely dense jargon, and bizarre, stilted metaphors. It's like they've been trained to produce output that's h…
Vomit: Clean up Claude 5's token output with a separate LLM
91–100 of 315 posts
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#92At some point one has to wonder if it's still worth using anthropic's models if we need to babysit 100% of its output with another vendor's model. Why not just use that other vendor's model for everything? I can't help but feel the circumstances that enable this kind of front page article are vestigial from the days when OAI was super bad and Anthropic was beyond reproach. This change-over-time is why I avoid getting…
I've been in the habit of pushing my claude-speak to codex to improve legibility, but only if I think someone is going to read it.
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#93I like the "Claudish to English" name better. https://github.com/gvzdv/claudish-to-english
Word, that one also includes an example which is great, shows really clearly what the issue is for those who might be less familiar
The rewrite did seem to lose the important fact about the ensure- pattern being idempotent.
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#94I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…
I'm probably going to be going against the grain here, but I think it's not as bad as it looks at first. I was similarly frustrated a few months ago, but have noticed I've started to learn the idiom. Its use of "dense jargon" and "stilted metaphor" is actually surprisingly consistent - it's speaking its own dialect, and you get used to it. After a while it gets much easier to read and even becomes somewhat efficient,…
…
No step here involves choosing based on meaning. It is a filter, a sort, and a slice.”
This is from Opus five minutes ago. I can certainly derive meaning from these kinds of statements in isolation, but paragraph upon paragraph of this is unintelligibly dense when trying to work with Claude to come up with a plan.
The worst part is that it can’t even make its responses make sense when asked to summarize in simple English or < 200 words. It simply cannot be steered to make its prose legible.
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#95I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…
Very well said. And when you say it like that, I have to wonder how much of this is a natural consequence of RHLF on such a grand scale, when you have millions of people pretty much much skimming chat responses or operating outside their depth and giving unqualified feedback to the models. Seems like a lot of people may be reinforcing what sounds smart over what is smart. Also as an aside: funny how much the LLMs con…
I wonder if the labs are sufficiently prepared to filter this kind of stuff out. I see a lot of non-developers asking development things of Claude, getting confused when they're in over their depth, and getting upset that they don't understand what the model is providing them, giving it bad feedback, and subsequently making the AI worse for the rest of us who know how to use the tool.
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#96Or just take full control of your agentic coding experience with Pi Coding Agent and picking and choosing your favorite model's API discounted on flex pricing on deepinfra.com instead. I highly recommend it. Claude and Codex usage limits cannot be trusted. Paying your own API bills in full is superior.
I use Pi but with my codex subscription, still preferable to paying the API cost (and I know I would be, as I track how much the cost 'should' be via token api pricing). Wish I could use my Claude subscription with pi too, much preferable to the endless command execution allow/deny prompts you have to do with CC, versus proper autonomous allow/deny lists defined ahead of time. Curious why you recommend the API? It's…
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#97I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…
Unfortunately this may only start to get worse as the AIs are trained on more and more AI generated content.
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#98Earlier quoted context omitted.
Pre Trump-castration Fable was verbose, but had a point , and used that wordiness to say or show the indeed intelligent things it reasoned about. This, whatever this is, is something else.-
Is the nerfed Fable worse or better than Opus 5?
I would opine:
- Nerfed Fable is worse than Fable
- ... and I would argue Opus 5 is worse than nerfed Fable, to the point I've found it unusable.-
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#99I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…
So going to continue trying that as a command structure going forwards...
Re: Vomit: Clean up Claude 5's token output with a separate LLM
#100I've been grappling with this for weeks, not just in Claude but in Codex as well, which isn't quite as bad but still annoying. AGENTS.md does very little, agents will consistently violate the communication preferences, especially as the session drags on. It's incredible to me that there's no good way to reliably change the way an LLM responds to you that a workaround like this would even be necessary. It seems like s…