Live data from Hacker News

Claude Opus 4.7

anthropic.com

571–580 of 1001 posts

Re: Claude Opus 4.7

#571
> Opus 4.7 introduces a new xhigh (“extra high”) effort level

I hope we standardize on what effort levels mean soon. Right now it has big Spinal Tap "this goes to 11" energy.

Re: Claude Opus 4.7

#572
post #526

Working on some research projects to test Opus 4.7. The first thing I notice is that it never dives straight into research after the first prompt. It insists on asking follow-up questions. "I'd love to dive into researching this for you. Before I start..." The questions are usually silly, like, "What's your angle on this analysis?" It asks some form of this question as the first follow-up every time. The second obser…

I had a conversation right during the launch so not fully sure if it was Opus 4.7 but I also noticed the same behavior of asking questions that did not seem particularly useful to me, tho I still prefer that to not asking enough.

Re: Claude Opus 4.7

#573

This is a CC harness thing than a model thing but the "new" thinking messages ('hmm...', 'this one needs a moment...') are extraordinarily irritating. They're both entirely uninformative and strictly worse than a spinner. On my workflows CC often spends up to an hour thinking (which is fine if the result is good) and seeing these messages does not build confidence.

Could you say more about your workflow? I don’t think I’ve ever gotten close to an hour of thinking before. Always curious to learn how to get more out of agents.

I don't think it's something special about my workflow and more the application area--I'm writing a lot of Lean lately and particularly knotty proofs can take quite a lot of time. Long thinking intervals are more of a bug than a feature IMO: Even if Claude can one-shot the proof in 40-60 minutes I'd rather have a partial proof in 15 and fill in the gaps myself.

Re: Claude Opus 4.7

#574

Earlier quoted context omitted.

yeah they took "i pick the budget" and turned it into "trust us".

I keep saying even if there's not current malfeasance, the incentives being set up where the model ultimately determines the token use which determines the model provider's revenue will absolutely overcome any safeguards or good intentions given long enough.

This might be true, but right now everybody is like "please let me spend more by making you think longer." The datacenter incentives from Anthropic this month are "please don't melt our GPUs anymore" though.

Re: Claude Opus 4.7

#575
post #455

noticing sharp uptick in "i switched to codex" replies lately. a "codex for everything" post flocking the front page on the day of the opus 4.7 release me and coworker just gave codex a 3 day pilot and it was not even close to the accuracy and ability to complete & problem solve through what we've been using claude for. are we being spammed? great. annoying. i clicked into this to read the differences and initial exp…

[deleted]

Re: Claude Opus 4.7

#576

Too late, personally after how bad 4.6 was the past week I was pushed to codex, which seems to mostly work at the same level from day to day. Just last night I was trying to get 4.6 to lookup how to do some simple tensor parallel work, and the agent used 0 web fetches and just hallucinated 17K very wrong tokens. Then the main agent decided to pretend to implement tp, and just copied the entire model to each node...

I guess our conscience of OpenAI working with the Department of War has an expiry date of 6 weeks.

I quoted 2 weeks at the time. I think even that was generous.

Re: Claude Opus 4.7

#577

New model - that explains why for the past week/two weeks I had this feeling of 4.6 being much less "intelligent". I hope this is only some kind of paranoia and we (and investors) are not being played by the big corp. /s

I don't get it. Why would they make the previous model worse before releasing an update?

Just guessing, but it would seem like physical hardware constraints would dictate this approach. You'd have to allocate a growing percentage of resources to the new model and scale back access/usage of the old as you role it out and test it.

Re: Claude Opus 4.7

#578

Earlier quoted context omitted.

Does this mean Claude no longer outputs the full raw reasoning, only summaries? At one point, exposing the LLM's full CoT was considered a core safety tenet.

Anthropic always summarizes the reasoning output to prevent some distillation attacks

Proprietary pattern matcher proves there's no moat; promptly pre-covers other's perception.

Re: Claude Opus 4.7

#579
The adaptive thinking behavior change is a real problem if you're running it in production pipelines. We use claude -p in an agentic loop and the default-off reasoning summary broke a couple of integrations silently — no error, just missing data downstream. The "display": "summarized" flag isn't well surfaced in the migration notes. Would have been nice to have a deprecation warning rather than a behavior change on the same model version.

Re: Claude Opus 4.7

#580
post #455

noticing sharp uptick in "i switched to codex" replies lately. a "codex for everything" post flocking the front page on the day of the opus 4.7 release me and coworker just gave codex a 3 day pilot and it was not even close to the accuracy and ability to complete & problem solve through what we've been using claude for. are we being spammed? great. annoying. i clicked into this to read the differences and initial exp…

>> are we being spammed? great. annoying.

Yeah, very. Every single time this happens here, where there's a thread about an Anthropic model and people spam the comments with how Codex is better, I go and try it by giving the exact same prompt to Codex and Opus and comparing the output. And every single time the result is the same: Opus crushes it and Codex really struggles.

I feel like people like me are being gaslit at this point.

Post reply on HN