/architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
31–40 of 47 posts
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#32LLM-written readmes love to use inscrutable jargon that means nothing outside of the context window that birthed it.
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#33Earlier quoted context omitted.
The problem is that there are a bunch of benchmarks, the model providers often don't even use the same benchmarks, a bunch of them have known problems, and it's expensive to do your own benchmarks. I am a GPT 5.x booster since to me it just feels smarter, and I generally felt like the benchmarks backed me up, but it's not every benchmark, so sadly we're mostly arguing about vibes. SWEBench-Pro was a big one, though a…
I find it fascinating that every time this kind of discussion comes up, people talk about night and day experiences between Claude and Codex, in both directions. I’m really wondering what people are doing to get such different outcomes. I’m currently working on two projects/clients one using Claude, one using Codex. I have a strong preference for the latter, but not because I think it is much more intelligent or writ…
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#34Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#35Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#36Last night I switched back to Codex for a minute having burned through my tokens for the week with Fable and oh boy I had a terrible experience. Running in circles over simple problems (which I ended up solving myself, like a peasant) and running "terraform apply" several times despite several instructions all over the place to never do that. The performance difference was stark.
I had a similar experience. So far Fable has been a game changer, at least for the work I used it for. Having said that, I think its writing is definitely worse than GPT 5.5. Ethan Mollick also observed the same. He called it more "Claudy." It generates worse academic prose than other frontier models.
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#37I know how to reduce Fable tokens by 100% ; https://www.anthropic.com/news/fable-mythos-access
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#38DESIGN.md: > Each rule below is enforced mechanically by the skill, not left to vibes. > R1. Repo docs are the memory; not in HANDOFF.md = didn't happen SKILL.md: > Not in docs/HANDOFF.md = didn't happen. Refuse to judge results that exist only in conversation or builder chat output. "Mechnical enforcement" just means "prompting the LLM a bit extra" these days? It (still) amazes me how much effort and tokens we expen…
Agents are in a wacky state, which makes projects like this fall into a weird spot. Eg I vaguely expect my agent to do two disparate things: manage dependency injection for tools, prompt modifications, etc, but also be the sort of “brain trust” that controls the flow of execution (can we stop now, do we keep going, etc). This project is meant to be the latter, but there’s not a clean way to integrate that into Claude…
Re: /architect: Reduce Fable tokens by 80%, Fable orchestrates/reviews, Codex builds
#39> freezes the gates LLM-written readmes love to use inscrutable jargon that means nothing outside of the context window that birthed it.