Live data from Hacker News

Claude Opus 4.6

anthropic.com

321–330 of 1001 posts

Re: Claude Opus 4.6

#321

Earlier quoted context omitted.

One aspect of this is that apparently most people can't draw a bicycle much better than this: they get the elements of the frame wrong, mess up the geometry, etc.

Yes, but obviously AGI will solve this by, _checks notes_ more TerraWatts!

The word is terawatts unless you mean earth-based watts. OK then, it's confirmed, data centers in space!

Re: Claude Opus 4.6

#322
post #40

The bicycle frame is a bit wonky but the pelican itself is great: https://gist.github.com/simonw/a6806ce41b4c721e240a4548ecdbe...

I suppose the pelican must be now specifically trained for, since it's a well-known benchmark.

Re: Claude Opus 4.6

#324

Earlier quoted context omitted.

Just for fun? Not everything has to be super serious… have a laugh, go for a walk, relax…

Mass-mass-mass-mass good comment. I mean. No I’m having an error - probably claud

happy happy happy sad sad sad err am robot no feeling err err happy sad err too many emotions 404 not found

Re: Claude Opus 4.6

#325
post #80

I'm disappointed that they're removing the prefill option: https://platform.claude.com/docs/en/about-claude/models/what... > Prefilling assistant messages (last-assistant-turn prefills) is not supported on Opus 4.6. Requests with prefilled assistant messages return a 400 error. That was a really cool feature of the Claude API where you could force it to begin its response with e.g. ` They suggest structured outputs o…

So what exactly is the input to Claude for a multi-turn conversation? I assume delimiters are being added to distinguish the user vs Claude turns (else a prefill would be the same as just ending your input with the prefill text)?

> So what exactly is the input to Claude for a multi-turn conversation?

No one (approximately) outside of Anthropic knows since the chat template is applied on the API backend; we only known the shape of the API request. You can get a rough idea of what it might be like from the chat templates published for various open models, but the actual details are opaque.

Re: Claude Opus 4.6

#326

Can we talk about how the performance of Opus 4.5 nosedived this morning during the rollout? It was shocking how bad it was, and after the rollout was done it immediately reverted to it's previous behavior. I get that Anthropic probably has to do hot rollouts, but IMO it would be way better for mission critical workflows to just be locked out of the system instead of get a vastly subpar response back.

Anthropic has good models but they are absolutely terrible at ops, by far the worst of the big three. They really need to spend big on hiring experienced hyperscalers to actually harden their systems, because the unreliability is really getting old fast.

Re: Claude Opus 4.6

#328
post #55

Earlier quoted context omitted.

It's not just that. Everyone is complacent with the utilization of AI agents. I have been using AI for coding for quite a while, and most of my "wasted" time is correcting its trajectory and guiding it through the thinking process. It's very fast iterations but it can easily go off track. Claude's family are pretty good at doing chained task, but still once the task becomes too big context wise, it's impossible to ge…

Cost wise, doesn’t that depend on what you could be doing besides steering agents?

Isn't the quote something like: "If these LLMs are so good at producing products, where are all those products?"

Re: Claude Opus 4.6

#329
post #62

Claude Code release notes: > Version 2.1.32: • Claude Opus 4.6 is now available! • Added research preview agent teams feature for multi-agent collaboration (token-intensive feature, requires setting CLAUDE_CODE_EXPERIMENTAL_AGENT_TEAMS=1) • Claude now automatically records and recalls memories as it works • Added "Summarize from here" to the message selector, allowing partial conversation summarization. • Skills defi…

> Claude now automatically records and recalls memories as it works Neat: https://code.claude.com/docs/en/memory I guess it's kind of like Google Antigravity's "Knowledge" artifacts?

If it works anything like the memories on Copilot (which have been around for quite a while), you need to be pretty explicit about it being a permanent preference for it to be stored as a memory. For example, "Don't use emoji in your response" would only be relevant for the current chat session, whereas this is more sticky: "I never want to see emojis from you, you sub-par excuse for a roided-out spreadsheet"

Re: Claude Opus 4.6

#330
post #185

This is the first model to which I send my collection of nearly 900 poems and an extremely simple prompt (in Portuguese), and it manages to produce an impeccable analysis of the poems, as a (barely) cohesive whole, which span 15 years. It does not make a single mistake, it identifies neologisms, hidden meaning, 7 distinct poetic phases, recurring themes, fragments/heteronyms, related authors. It has left me completel…

This sounds wayyyy over the top for a mode that released 10 mins ago. At least wait an hour or so before spewing breathless hype.

[deleted]
Post reply on HN