Live data from Hacker News

Anthropic surpasses OpenAI to become most valuable AI startup

qazinform.com

111–120 of 512 posts

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#111
post #7

Earlier quoted context omitted.

GPT-5.5 is the better programmer but Opus 4.8 remains the better system architect and product designer. Codex is very "miss the forest for the trees", but is much better at successfully making large changes in large codebases. Claude Code makes more mistakes, but has more taste and a better grasp on idiomatic and elegant software development. If you can afford to, I recommend juggling both.

I find arguing that a complex weighted graph has a taste is interesting. This is not a jab, but a genuine curiosity of mine.

More interesting than arguing a jumble of electrochemical reactions have taste? That may seem more readily familiar but is no less strange if you prod at it. Nonetheless it’s difficult to argue either don’t produce output that has qualities of discernment (ie taste).

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#112

I never want to hear from developers again that they are not susceptible to marketing. I see meet ups specifically about Claude often. Modern tupperware party. A colleague was convinced Claude is better so we played a game. We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code. Couldn’t tell. Edit: i…

in my experience out of the box Claude Code is the better tool if you want to spend 0 time on config

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#113
post #82

Bernie Madoff would be jealous. Stealing all open source and reselling "git clone" + "sed" for $1 trillion is something he did not achieve. The chutzpah is remarkable.

They are selling shovels, not mining gold themselves though.

So it's more like selling a derivative on a promise to steal open source for you in a useful way.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#115
post #70

Earlier quoted context omitted.

Your argument is fine but different from the claim the OP is making. You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference in the output. Subjectively, people might still prefer one over due to anything from design to marketing, but that's very different from the claim that X is better than Y for coding (see: "A colleague was convinced Claude is better"). Ba…

> You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference in the output. You definitely can in principle; that’s the entire point of the comment you are responding to. If one tool completes it in 10 minutes with little hand holding, and the other does it in one hour at 4× the cost and while needing a lot of steering, the former is arguably better even if the e…

This obviously correct take will get pushback, so let me add some other examples:

- which tool required more detailed goal-setting in the prompt?

- did one tool ask follow-up questions up front vs spread out over implementation?

- did either tool match existing coding styles?

- did either tool remind you about potential conflicts between what you asked it to build and other parts of the codebase?

There are a lot of ways to compare agents besides just the code. (Similarly, working engineers are not evaluated just on their code output.)

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#116

I never want to hear from developers again that they are not susceptible to marketing. I see meet ups specifically about Claude often. Modern tupperware party. A colleague was convinced Claude is better so we played a game. We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code. Couldn’t tell. Edit: i…

If advertising is a multi-billion dollar industry then it has to be effective!

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#117
post #42

I never want to hear from developers again that they are not susceptible to marketing. I see meet ups specifically about Claude often. Modern tupperware party. A colleague was convinced Claude is better so we played a game. We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code. Couldn’t tell. Edit: i…

It's crazy hearing devs on this site claim Claude is 10x better than all other AI solutions. I think it is fomo. Claude $LATEST_VERSION is perceived as the best and anything else is "missing out". New version comes out? Suddenly the old version is worthless, how on earth did anyone get work done with that? Same reason people buy the RTX 4090 and 5090 cards - overpriced but they must have the "best". Never mind the di…

Opus 4.8 and GPT 5.5 are the best models, but people don't care about "best" anymore, until there is a big leap in capability I don't think anyone will care about point releases.

Vibes and tribalism will prevail until one of emerges as clearly and unambiguously superior to the other.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#118
post #104

I think Sam Altman is an asshole and I prefer to spend my money elsewhere. Frontier models being commoditize is inevitable. OpenAI thinks they're still competing on technology, and not user experience and market reputation otherwise they'd understand the continuous negative PR generated by Altman's chaos is going to cost them everything.

He must have done something personally to you.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#119
post #70

Earlier quoted context omitted.

Your argument is fine but different from the claim the OP is making. You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference in the output. Subjectively, people might still prefer one over due to anything from design to marketing, but that's very different from the claim that X is better than Y for coding (see: "A colleague was convinced Claude is better"). Ba…

> You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference in the output. You definitely can in principle; that’s the entire point of the comment you are responding to. If one tool completes it in 10 minutes with little hand holding, and the other does it in one hour at 4× the cost and while needing a lot of steering, the former is arguably better even if the e…

That's a fair callout and I agree my statement was too general in just mentioning 'output', as you correctly pointed out. To define 'better' you would indeed need to agree on the dimensions you would evaluate candidates against.

I think a more appropriate rephrasing would be 'You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference on dimensions you care about'. In the case of latest of claude code vs codex with gpt 5.5) both are similar enough in the dimensions people will care about in evaluating (vs. differing wildly in cost or time taken).

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#120
post #91
post #50

Earlier quoted context omitted.

I’ve been using DeepSeek V4 in OpenCode exclusively for about a month. I think it’s great, but coming from Claude Code it did feel like going back in time by ~6 months in model capabilities. This isn’t a big deal to me for what I do, but the difference is definitely there.

Deepseek v4 Pro is like Opus 4.5 or GPT 5.2, but costs pennies on the pound for API. Which is to say, I should definitely be using it more to let my Codex and Claude subs go further.

Opus 4.5 was definitely stronger than DeepSeek V4 for me, specifically with large context.

I’m being pedantic/splitting hairs, though. I’ve obviously switched to DeepSeek full-time because it makes more sense to me pragmatically — I spend a few more tokens to get the outcome I want, but the tokens are cheap as dirt and the API is faster.

Perhaps I should plug it into Claude Code and see how it performs? I haven’t tried that.

Post reply on HN