Live data from Hacker News

Anthropic surpasses OpenAI to become most valuable AI startup

qazinform.com

441–450 of 512 posts

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#441

Earlier quoted context omitted.

> We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code. > Couldn’t tell. Why would you expect them to be able to recognize the signature of a model from a pair of PRs? I don’t understand why you think this is a useful test for anything when we have numerous benchmarks that run 100s of tests on model…

I think the subscription pricing model kind of incentivizes developers (at least hobby developers) to pick one and go all in on it. For someone who has probably never paid $20/mo for a piece of software in their life, $20/mo is kind of a big commitment, and the pay-per-token schemes are reportedly much more expensive for the equivalent blob of coding they enable. So you "pick one," plonk down the $20, and use it as m…

> For someone who has probably never paid $20/mo for a piece of software in their life, $20/mo is kind of a big commitment

I could see this being true for a high school student or college freshman eating rice and beans in their dorm room. Many of us have been there.

For someone in a software development career, $20/month for tools is a trivial expense.

I think some people have a strong aversion to paying for any tooling, but I don't think the people carrying around their $3000 MacBook Pros are going to avoid paying $20 for a month to try something new if they're using it daily.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#442

Earlier quoted context omitted.

> At least Sam thinks it should work for everyone and has done literal experiments on UBI Where are you getting this? He goes out of his way to say how dangerous AI, and has implied before congress that only companies with special licenses should be able to develop it.

It's a rumor being spread by OpenAI employees. https://x.com/_aidan_clark_/status/2052089187659346047

Well not sure I can take any of them seriously if they think they are building AGI with LLM's. There is literally no thinking involved in an Large language model.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#443
[dupe]

Main discussion:

Anthropic raises $65B in Series H funding at $965B post-money valuation

https://news.ycombinator.com/item?id=48313048

Other submitted more-common source reports on this from 2 days ago that didn't need traction because of the above discussion:

https://www.nytimes.com/2026/05/28/technology/anthropic-tops... (https://news.ycombinator.com/item?id=48315537)

https://www.theguardian.com/technology/2026/may/28/anthropic... (https://news.ycombinator.com/item?id=48321498)

https://www.wsj.com/tech/ai/anthropic-valuation-openai-80bf2... (https://news.ycombinator.com/item?id=48315537)

https://www.businessinsider.com/anthropic-surpasses-openai-w... (https://news.ycombinator.com/item?id=48316994)

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#445

Earlier quoted context omitted.

> We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code. > Couldn’t tell. Why would you expect them to be able to recognize the signature of a model from a pair of PRs? I don’t understand why you think this is a useful test for anything when we have numerous benchmarks that run 100s of tests on model…

Kind of orthogonal to the discussion, but could you broadly describe the code you're working on that both models are bad at? One thing I'm still struggling with is figuring out what types of code LLMs can vs cannot write.

Rules of thumb:

The more your toolchain (compilers, linters, etc) can statically verify, the better agents will do.

The terser the code, the better agents will do.

The more often similar problems have been solved in open source, the better agents will do. Agents seem particularly good at plumbing together different pieces of software.

Anything that requires a judgement call, as opposed to having one obvious way to do it, will get worse results from an agent.

As the scope of the request grows, agents get worse at it. This can be mitigated somewhat using various techniques ("write a plan", "do step 1 of the plan", etc) but never fully resolved. At some point the task is so big that it's necessary to do large parts by hand.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#446
post #169

Earlier quoted context omitted.

I’ve heard this said, but why?

he pushes mysticism of the models he's starkly anti-China with a warlike posture that I find dangerous and unappealing Anthropic has a much more confused mission statement than OpenAI in interviews, Dario appears to care little for the well-being of common folk, while Sam at least pretends

Thanks for elaborating. Can you say more about mysticism?

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#447

The real problem is this: no cheap model right now produces a genuinely beautiful, usable UI when it comes to website building. Not one. And here’s the core tension. The models keep getting better. GPT 5.5 improved. But it also got more expensive. Opus 4.7 to 4.8 has become outrageously priced too, up 50%, and 4.6 was already brutally expensive to begin with. API pricing is a real pain. What’s missing is any meaningf…

I don't see a problem as long as Chinese companies are not banned; expensive providers like Anthropic will prove that value is being provided, encouraging competition, which will take care of problems.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#448

Earlier quoted context omitted.

> You cannot simply make a claim that (model + harness) X is better than Y, but then have no discernible difference in the output. You definitely can in principle; that’s the entire point of the comment you are responding to. If one tool completes it in 10 minutes with little hand holding, and the other does it in one hour at 4× the cost and while needing a lot of steering, the former is arguably better even if the e…

The colleague implicitly agreed that comparing the output was a valid way to settle the matter as they took part in the test, so they weren't using "better" in the way you propose.

I wasn’t really discussing the colleague, but either way, from:

> A colleague was convinced Claude is better so we played a game. We used the claude code and codex harness and I implemented some prs they needed with gpt5.5 and opus4.7 and asked them to identify which came from which only from the code.

I don’t think it’s obvious that they specifically agreed that losing the game meant that. They might just have thought “sure, it might be fun”, if they even gave it that much thought.

“So we played a game” is rather vague and I feel it’s a bit of a leap to read it as: “as an explicit outcome of their claim that Claude is better, we made a formal bet as to whether they could tell the difference in the output, the failure of which would mean a full retractation of their statement”.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#449
post #403

Earlier quoted context omitted.

Anthropic hasn't done anything customer hostile to me.

Notable and persistent and extraordinarily petty is their refusal to read AGENTS.md files, forcing the inclusion of their branding into the source of repos directly.

That is actually pretty wild to me. Do you have any sources regarding this?

I see my GPT5.x harness seek out and read CLAUDE.md files all the time.

Re: Anthropic surpasses OpenAI to become most valuable AI startup

#450
post #50
post #42

Earlier quoted context omitted.

It's crazy hearing devs on this site claim Claude is 10x better than all other AI solutions. I think it is fomo. Claude $LATEST_VERSION is perceived as the best and anything else is "missing out". New version comes out? Suddenly the old version is worthless, how on earth did anyone get work done with that? Same reason people buy the RTX 4090 and 5090 cards - overpriced but they must have the "best". Never mind the di…

I’ve been using DeepSeek V4 in OpenCode exclusively for about a month. I think it’s great, but coming from Claude Code it did feel like going back in time by ~6 months in model capabilities. This isn’t a big deal to me for what I do, but the difference is definitely there.

So my GPU comparison is pretty apt then. Paying 4-10x to be slightly ahead of the curve.
Post reply on HN