Live data from Hacker News

Claude Opus 4.7

anthropic.com

701–710 of 1001 posts

Re: Claude Opus 4.7

#701

> We stated that we would keep Claude Mythos Preview’s release limited and test new cyber safeguards on less capable models first. Opus 4.7 is the first such model: its cyber capabilities are not as advanced as those of Mythos Preview (indeed, during its training we experimented with efforts to differentially reduce these capabilities). We are releasing Opus 4.7 with safeguards that automatically detect and block req…

Curious how the safeguards work and what impact they will have.

In general I feel that over-engineering safeguards in training comes at a noticeable cost to general intelligence. Like asking someone to solve a problem on a white board in a job interview. In that situation, the stress slices off at least 10% of my IQ.

Re: Claude Opus 4.7

#702

OK 4.7 is a different animal altogether. - no longer a 10 year old autistic programming genius, but a confident programming genius basically taking the lead on what to do and truly putting you in your place. Slightly impatient but surprisingly confident, much more detailed in the tasks he does and double checks his work on the fly. - very little to no need to ask, have you rememebered to do this and that, its done. -…

slop

Re: Claude Opus 4.7

#703

So many messages about how Codex is better then Claude from one day to the other, while my experience is exactly the same. Is OpenAI botting the thread? I can't believe this is genuine content.

Why do you assume it's botted? Just open up Codex on GPT 5.4 and point it at your codebase.

Re: Claude Opus 4.7

#704
post #677
post #463

Earlier quoted context omitted.

Sounds like you will need to drink a(n identity) verification can soon [1] to continue as a security researcher on their platform. 1: https://support.claude.com/en/articles/14328960-identity-ver... Identity verification on Claude Being responsible with powerful technology starts with knowing who is using it. Identity verification helps us prevent abuse, enforce our usage policies, and comply with legal obligations. W…

Context for "please drink verification can": https://files.catbox.moe/eqg0b2.png

Yes, it's a stupid 4chan meme from 2013. I can only surmise those who quote it either don't know its origin, or they must be wholeheartedly 'embracing the cringe.'

Re: Claude Opus 4.7

#705
So far since continuing coding/debugging with 4.7 it's failed to fix 3 simple bugs after explaining it like 5 times and having a previous working example to look at...hmmmmmm....

Re: Claude Opus 4.7

#708

Serious question about using Claude for coding. I maintain a couple of small opensource applications written in python that I created back in 2014/2015. I have used Claude Code to improve one of my projects with features I have wanted for a long time but never really had the time to do. The only way I felt comfortable using Claude Code was holding its hand through every step, doing test driven changes and manually re…

Your suspicion is right.

Re: Claude Opus 4.7

#709

Early benchmark results on our private complex reasoning suite: https://gertlabs.com/?mode=agentic_coding Opus 4.7 is more strategic, more intelligent, and has a higher intelligence floor than 4.6 or 4.5. It's roughly tied with GPT 5.4 as the frontier model for one-shot coding reasoning, and in agentic sessions with tools, it IS the best, as advertised (slightly edging out Opus 4.5, not a typo). We're still running m…

Is there a page where I could read more? What's unintuitive at a glance is that Opus 4.7 has a lower success rate than Sonnet 4.6 (90% vs 100%) while having a higher Avg Percentile (87.2% vs 70.9%).

Re: Claude Opus 4.7

#710

> Opus 4.7 uses an updated tokenizer that improves how the model processes text. The tradeoff is that the same input can map to more tokens—roughly 1.0–1.35× depending on the content type. caveman[0] is becoming more relevant by the day. I already enjoy reading its output more than vanilla so suits me well. [0] https://github.com/JuliusBrussee/caveman/tree/main

What about some thing like

https://github.com/rtk-ai/rtk

Post reply on HN