Live data from Hacker News

Claude Opus 4.6

anthropic.com

141–150 of 1001 posts

Re: Claude Opus 4.6

#141

This is huge. It only came out 8 minutes ago but I was already able to bootstrap a 12k per month revenue SaaS startup!

It only came out 35 minutes ago and GPT-5.3-codex already took the crown away!

Gee, it scored better on a benchmark I've never heard of? I'm switching immediately!

Re: Claude Opus 4.6

#142
post #102
post #59

Earlier quoted context omitted.

Is this a react feature or did they build something to translate react to text for display in the terminal?

They used Ink: https://github.com/vadimdemedes/ink I've used it myself. It has some rough edges in terms of rendering performance but it's nice overall.

Thats pretty interesting looking, thanks!

Re: Claude Opus 4.6

#143

This is huge. It only came out 8 minutes ago but I was already able to bootstrap a 12k per month revenue SaaS startup!

It only came out 35 minutes ago and GPT-5.3-codex already took the crown away!

Why are you posting the same message in every thread? Is this OpenAI astroturfing?

Re: Claude Opus 4.6

#144

This is huge. It only came out 8 minutes ago but I was already able to bootstrap a 12k per month revenue SaaS startup!

Amateur. Opus 4.6 this afternoon built me a startup that identifies developers who aren’t embracing AI fully, liquifies them and sells the produce for $5/gallon. Software Engineering is over!

Opus 4.6 agentically found and proposed to my now wife.

Re: Claude Opus 4.6

#146

I'm still not sure I understand Anthropic's general strategy right now. They are doing these broad marketing programs trying to take on ChatGPT for "normies". And yet their bread and butter is still clearly coding. Meanwhile, Claude's general use cases are... fine. For generic research topics, I find that ChatGPT and Gemini run circles around it: in the depth of research, the type of tasks it can handle, and the qual…

I kinda agree. Their model just doesn't feel "daily" enough. I would use it for any "agentic" tasks and for using tools, but definitely not for day to day questions.

Why? I use it for all and love it.

That doesn't mean you have to, but I'm curious why you think it's behind in the personal assistant game.

Re: Claude Opus 4.6

#150

5.3 codex https://openai.com/index/introducing-gpt-5-3-codex/ crushes with a 77.3% in Terminal Bench. The shortest lived lead in less than 35 minutes. What a time to be alive!

That's a massive jump, I'm curious if there's a materially different feeling in how it works or if we're starting to reach the point of benchmark saturation. If the benchmark is good then 10 points should be a big improvement in capability...
Post reply on HN