Live data from Hacker News

Viewing profile — ptnpzwqd

ptnpzwqd

HN member
Joined
Sat, Nov 04, 2023, 12:39 PM UTC
HN karma
146
Public activity
27 items

About ptnpzwqd

No profile information was provided.

Recent public activity

  1. comment
    Comment #48944776

    Be that as it may, it would seem absurd if we start calling distillation out as antagonistic, but don't do the same for the SOTA models being trained on human-created data.

  2. comment
    Comment #48421781

    This does not really match my observations. While it does feel to me the sentiment is shifting towards a more negative one, overall HN feels reasonably balanced between those that …

  3. comment
    Comment #48083316

    Responses seem to be very "either or" as usual on such topics. I think it should be possible to appreciate how impressive this is on one hand, while also discussing the limitations…

  4. comment
    Comment #48006484

    I think a notable difference is that the AI that is portrayed in most sci-fi (that I have read/watched anyway) tend to be "logical machines" that act deterministically based on the…

  5. comment
    Comment #47611036

    I feel the conclusions here are a bit thin. Code quality tends to have an impact on more than just aesthetics - and Claude Code certainly feels like a buggy mess from an end user's…

  6. comment
    Comment #47483512

    On the falling behind: I strongly doubt that is going to be the case - picking up these tools is not rocket science, even if you want to be able to use them fairly effectively. In …

  7. comment
    Comment #47325847

    I live in two realities too. One where articles like this talk about a 10x increase in productivity, where the sentiment on HN is that all software can now be vibe coded without re…

  8. comment
    Comment #47323171

    It was maybe not quite clear enough in my comment, but this is more of a hypothetical future scenario - not at all where I assess LLMs are today or will get to in the foreseable fu…

  9. comment
    Comment #47322601

    Yes, but LLM-based reviews are not nearly a compensation for human review, so it doesn't change much.

  10. comment
    Comment #47321272

    Yes, I agree. It was just me playing with a hypothetical (but in my view not imminent) future where vibe-coding without review would somehow be good enough.

  11. comment
    Comment #47321229

    At the moment verification at scale is an unsolved problem, though. As mentioned, I think this will act as a rough filter for now, but probably not work forever - and denying contr…

  12. comment
    Comment #47321068

    Sure - and I suspect we will see that soon enough. But it has downsides too, and finding the right way to vet potential contributors is tricky.

  13. comment
    Comment #47321052

    Even if we assume that LLMs become good enough for this to be true (some might feel that is the case already - I disagree, but that is beside the point), there is no reason why OSS…

  14. comment
    Comment #47320989

    The problem is the increasing review burden - with LLMs it is possible to create superficially valid looking (but potentially incorrect) code without much effort, which will still …

  15. comment
    Comment #47320859

    Not necessarily a bad idea, but I think the bigger issue here and now is the increasing assymmetry in effort between code submitter and reviewer, and the unsustainable review burde…

  16. comment
    Comment #47320836

    I suspect this is for now just a rough filter to remove the lowest effort PRs. It likely will not be enough for long, though, so I suspect we will see default deny policies soon en…

  17. comment
    Comment #47320789

    I think this is a reasonable decision (although maybe increasingly insufficient). It doesn't really matter what your stance on AI is, the problem is the increased review burden on …

  18. comment
    Comment #47240350

    Correct, but that has and probably always will be the case. You spend the time on what is needed for you to move ahead - if code review is now the most time consuming part, that is…

  19. comment
    Comment #47239977

    I of course cannot say what the future holds, but current frontier models are - in my experience - nowhere near good enough for such autonomy. Even with other agents reviewing the …

  20. comment
    Comment #47239782

    If reviewing has become the bottleneck, the obvious - albeit slightly boring - solution is to slow down spitting out new code, and spend relatively more time reviewing. Just going …

  21. comment
    Comment #47210361

    You can use this: hnthrowaway.outboard407@passmail.net

  22. comment
    Comment #47207322

    I don't expect that - am merely responding to the parent comments claim that Claude consistently one-shots production ready code (which does not at all match my observations).

  23. comment
    Comment #47205729

    Feel free to share it, would be very curious - ideally alongside the prompts.

  24. comment
    Comment #47205117

    I have used Claude (incl. Opus 4.6) fairly extensively, and Claude still spits out quality that is far below what I would call production ready - both littered with smaller issues,…

  25. comment
    Comment #47183859

    There are some really baffling takes here. And it doesn't really matter how good or bad coding agents are. Coding agents greatly reduce the barrier to contributing something that a…