Live data from Hacker News

A week of using Codex more than Claude

allaboutcoding.ghinda.com

241–250 of 311 posts

Re: A week of using Codex more than Claude

#241
post #138

Earlier quoted context omitted.

Elon burned too many bridges to warrant ever supporting anything he is associated with ever again.

Sure, you can let politics dominate everything you do. Or you can realize that SpaceX is a massive (public) company with thousands of employees, and millions of shareholders, all of whom have their own opinions and goals, just like any other corporation. Competition is good. Excluding a leading player in the market because you don’t like Elon Musk is…something.

> Or you can realize that SpaceX is a massive (public) company with thousands of employees, and millions of shareholders, all of whom have their own opinions and goals, just like any other corporation.

Do you have ANY idea about the SpaceX corporate structure? Elon is basically SpaceX's Sun God and the other shareholders don't matter.

Plus SpaceX is incorporated in Texas where I'm fairly sure the legal system is arranged in such a way that it's supremely hard to contest anything in terms of corporate decisions.

As far as the average person cares, every SpaceX shareholder and employee is basically an Elon sharecropper and they matter less than Musk's toenails in terms of corporate decision making.

Re: A week of using Codex more than Claude

#242
post #185

Earlier quoted context omitted.

> I realize that there’s a population of people who just refuse to use anything associated with Musk because politics have eaten our brains Everyone has their own frameworks for risk assessments, its more of the historical incidents associated with it than politics

I mean, sure. If people tell me that they're avoiding Grok because of that bug, I sorta get it. I think it's silly (again: they're all hoovering up my data), but at least it's a rational basis related to the actual product. But let's be real: Musk exists as a polarizing political character, and his association with Grok just breaks some people's brains. A fair number of those people don't want to admit it, and just l…

Is he polarizing? Yes, and so are a lot more other people. He is probably comes more stronger.

On a personal level, everyone have their own rules, they may not be able to imposing them on others , however they do happen to evaluate their relative understanding of other people based upon those rules.

For me personally, I would probably put Anthropic, OpenAI and Grok in the same bucket, they are doing everything possible to make money. Ethics, morals, long term impacts, all of such things are not in their playbook. But again what these companies are doing just reflects the people who invested in them and what they want out of it. In some ways you can say its the money trying to maximize itself at all costs.

Re: A week of using Codex more than Claude

#243

I have completely moved away from Claude to Codex. OpenAI models are better at following instructions and skill use. Claude generates more code, tries to build things that are not planned or needed, repeats same mistakes over and over. Chatgpt on the other hand just does enough, within a project would not repeat the same mistakes, tries to guess your workflow, so you don’t have to ask it to run the same. If you need…

I think the conclusion is a little too strong but mostly agree, Codex feels like a chisel and Claude feels like a wrecking ball.

Your words are definitely a better way to put it

Re: A week of using Codex more than Claude

#244

I have completely moved away from Claude to Codex. OpenAI models are better at following instructions and skill use. Claude generates more code, tries to build things that are not planned or needed, repeats same mistakes over and over. Chatgpt on the other hand just does enough, within a project would not repeat the same mistakes, tries to guess your workflow, so you don’t have to ask it to run the same. If you need…

One more observation - Claude is better at delegation than Codex.

Chatgpt is a better planner.

Have a discussion in Chatgpt, have Claude to plan work chunks and have GLM to deliver, Codex to review and fix, delivers an overall a better version.

Problem - It’s just too much of a context switching.

Solution - I am thinking about a new product, a collaborative workspace where I can run this workflow.

Re: A week of using Codex more than Claude

#246
post #208

Earlier quoted context omitted.

There's some real research-backed evidence that this approach elevates output quality, and it was even discussed a few months back: https://news.ycombinator.com/item?id=47169627

Anthropic confirmed it in their recent report https://news.ycombinator.com/item?id=49316271 > We expect that agents coordinating in the wild will act in higher variance ways than we see here, because they’ll have different backgrounds and therefore different contexts. They also, presumably, won’t all be Claudes.

I do use a second Claude (both Opus) to review the first one's output - it almost always finds worthwhile (often critical) observations. They have different 'perspectives', one wants the job done, one wants to find problems, is how I had it explained.

Claude chat itself called it 'Claude-on-Claude action', which I found cute.

Re: A week of using Codex more than Claude

#247

I'm still using gpt-5.5 in copilot (visual studio proper) as a "custom" model. It feels like I've been enjoying an ~unpatched exploit while everyone else wallows in the proverbial mud. A large part of what pushes developers toward these products appears to be the billing model. Pre-paying for tokens is some kind of ideological red line for a lot of developers. I think this is a strategic error. The subscription model…

> The subscription models have so many more perverse incentives baked in

Yes. But, as your "so many more" implies, there are also perverse incentives in pay-per-token. And now, for the first time, the companies with the perverse incentives also happen to own the intelligence needed to, ad-nauseam, evade market and customer oversight. Potentially, this is a war where one side can inflict a thousand paper cuts in one second and the humans are on the other side. I think this is going to be an interesting test, a taste if you will, of what AGI means.

Re: A week of using Codex more than Claude

#248

I'm still using gpt-5.5 in copilot (visual studio proper) as a "custom" model. It feels like I've been enjoying an ~unpatched exploit while everyone else wallows in the proverbial mud. A large part of what pushes developers toward these products appears to be the billing model. Pre-paying for tokens is some kind of ideological red line for a lot of developers. I think this is a strategic error. The subscription model…

> Those paying $100/m+ for subscription access are almost certainly getting taken for a ride based upon my experience with prepaid tokens.

I don't understand what this means. Are they overpaying and getting less? Typically "taken for a ride" means, exactly "The seller got more out of the deal than usual sellers would".

Buying a burger for $3000 == "taken for a ride".

Paying $30 for all you can eat != "taken for a ride".

Re: A week of using Codex more than Claude

#249
post #100

Earlier quoted context omitted.

It's only helpful to Claude in that it wastes more tokens and costs more.

It does visibly use a few more tokens in the short term. But it hurts Claude's ability in the long term if removed because it has to either think more every time it encounters that section in the code (more invisible thinking tokens used), or there's an increased risk of making a maladaptive change without context. Heck I've been bitten enough times already by it, though more recently it's cross project concern where…

What you need is RAG and better context management. I use vornik (it has plugins for both codex and Claude) that gives both the long term memory you mention. Hope this helps.

Re: A week of using Codex more than Claude

#250
post #10

while everyone is somehow still stuck on and fascinated by claude, heres your quick update on the sota of coding models and harnesses mid august 2026 codex is good, both cli and desktop app, you get lots of usage on any plan. sol is good! and gets the job done, write or dictate a very long and thoughtful prompt, and leave sol xhigh or max fast working on it for an hour or so omp is an amazing harness, any feature cla…

I've been using all the SOTA models a lot at work, like serious amount of tokens. It's been really rare that I stick with one model and harness for too long... Except a month ago I started testing Kimi K3 and omp and I never went back.

Something with this combo works really well for Rust dev. The model doesn't really annoy me at all and I have not switched to Opus or SOL. And the monthly token bill is much lower...

Post reply on HN