Live data from Hacker News

Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

tokens.billchambers.me

271–280 of 620 posts

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#271

Earlier quoted context omitted.

I kind feel the same. I’m learning things and doing things in areas that would just skip due to lack of time or fear. But I’m so much more detached of the code, I don’t feel that ‘deep neural connection’ from actual spending days in locked in a refactor or debugging a really complex issue. I don’t know how a feel about it.

As someone who's switched from mobile to web dev professionally for the last 6 months now. If you care about code quality, you'll develop that neural connection after some time. But if you don't and there's no PR process (side projects), the motivation to form that connection is quite low.

> If you care about code quality, you'll develop that neural connection after some time.

No, because you can get LLMs to produce high quality code that has gone through an infinite number of refinement/polish cycles and is far more exhaustive than the code you would have written yourself.

Once you hit that point, you find yourself in a directional/steering position divorced from the code since no matter what direction you take, you'll get high quality code.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#272

Earlier quoted context omitted.

>we don't want a hard dependency on another multi-billion dollar company just to write software One of two main reasons why I'm wary of LLMs. The other is fear of skill atrophy. These two problems compound. Skill atrophy is less bad if the replacement for the previous skill does not depend on a potentially less-than-friendly party.

You can argu that you will have skill atrophy by not using LLMs. We have gone multi cloud disaster recovery on our infrastructure. Something I would not have done yet, had we not had LLMs. I am learning at an incredible rate with LLMs.

You're learning at your standard rate of learning, you're just feeding yourself over-confidence on how much you're absorbing vs what the LLM is facilitating you rolling out.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#273

Earlier quoted context omitted.

What open models are truly competing with both Claude Code and Opus 4.7 (xhigh) at this stage?

That's a lame attitude. There are local models that are last year's SOTA, but that's not good enough because this year's SOTA is even better yet still... I've said it before and I'll say it again, local models are "there" in terms of true productive usage for complex coding tasks. Like, for real, there. The issue right now is that buying the compute to run the top end local models is absurdly unaffordable. Both in ge…

> that are last year's SOTA

Early last year or late last year?

opus 4.5 was quite a leap

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#275

Earlier quoted context omitted.

Your qualification for if an LLM can write a novella is it has to be as good as The Metamorphosis ? Yes, those are examples of novellas, surely you believe an LLM could write a bad novella? I'm not sure what your point is. Either you think it can't string the words together in that length or your standard is it can't write a foundational piece of literature that stays relevant for generations... I'm not sure which.

I don't think it can write something that's of a fraction of the quality of Kafka. But GP's argument ("limit the space to text") could be taken to imply - and it seems to be a common implication these days - that LLMs have mastered the text medium, or that they will very soon. > it can't write a foundational piece of literature Why not, if this a pure textual medium, the corpus includes all the great stories ever wri…

I don't know what to tell you. It's more than a little absurd to make the qualification of being able to do something to be that the output has to be considered a great work of art for generations.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#276

Earlier quoted context omitted.

You lose some, you win some. The win could be short-term much higher, however imagine that the new tool suddenly gets ragged pulled from under your feet. What do you do then? Do you still know how to handle it the old way or do you run into skill atrophy issues? I’m using Claude/Codex as well, but I’m a little worried that the environment we work in will become a lot more bumpy and shifty.

> the new tool suddenly gets ragged pulled from under your feet If that happened at this point, it would be after societal collapse.

I don’t even wanna think about that scenario, maybe he gets averted somehow.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#278
post #46

We dropped Claude. It's pretty clear this is a race to the bottom, and we don't want a hard dependency on another multi-billion dollar company just to write software We'll be keeping an eye on open models (of which we already make good use of). I think that's the way forward. Actually it would be great if everybody would put more focus on open models, perhaps we can come up with something like the "linux/postgres/git…

> I think that's the way forward. Actually it would be great if everybody would put more focus on open models,

I'm still surprised top CS schools are not investing in having their students build models, I know some are, but like, when's the last time we talked about a model not made by some company, versus a model made by some college or university, which is maintained by the university and useful for all.

It's disgusting that OpenAI still calls itself "Open AI" when they aren't truly open.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#279

Earlier quoted context omitted.

I was worried about skill atrophy. I recently started a new job, and from day 1 I've been using Claude. 90+% of the code I've written has been with Claude. One of the earlier tickets I was given was to update the documentation for one of our pipelines. I used Claude entirely, starting with having it generate a very long and thorough document, then opening up new contexts and getting it to fact check until it stopped…

Yeah, +1. I will never be working on unsolved problems anyhow. Skill atrophy is not happening if you stay curious and responsible.

I used to speak Russian like I was born in Russia. I stopped talking Russian … every day I am curious ans responsible but I can hardly say 10 words in Russian today. if you don’t use it (not just be curious and responsible) you will lose it - period.

Re: Anonymous request-token comparisons from Opus 4.6 and Opus 4.7

#280

Earlier quoted context omitted.

What is this, some sort of cult?

No, it is an as snarky response to a person being snarky about usefulness of AI agents. It does seem like there is a cult of people who categorically see LLMs as being poor at anything without it being founded in anything experience other than their 2023 afternoon to play around with it.

Who cares? Why are people so invested in trying to “convert” others to see the light?

Can’t you be satisfied with outcompeting “non believers”? What motivates you to argue on the internet about it? Deep down are you insecure about your reliance on these tools or something, and want everyone else to be as well?

Post reply on HN