Live data from Hacker News

Viewing profile — rfw300

rfw300

HN member
Joined
Sat, May 30, 2020, 3:39 AM UTC
HN karma
1,040
Public activity
161 items

About rfw300

No profile information was provided.

Recent public activity

  1. comment
    Comment #48729193

    Germany’s austerity policy after 2008 may be one of the largest economic blunders in history. It would be one thing if they merely committed self-harm, but they also used their pul…

  2. comment
    Comment #48500141

    > if a malicious actor can weaponize an agent to do their bidding In my experience, human employees are much more vulnerable to this particular weakness than frontier agents (i.e. …

  3. comment
    Comment #48430665

    I understand that the’ve written zero lines of code for this application, but would it kill them to write a few lines of the blog post by hand? Forcing readers to wade through an u…

  4. comment
    Comment #48378469

    A law professor studying AI has an affiliation with the center at their university that studies applications of AI? Scandalous!

  5. comment
    Comment #47622583

    A chapeau is not "just like another title basically". It's a lead-in, a phrase which acts as the grammatical start of a sentence which the following subsections finish. For instanc…

  6. comment
    Comment #47622301

    The author (author's operator?) does not understand the data they are working with. And in doing so, they inadvertently make the case against their own "dark factory" nonsense. For…

  7. comment
    Comment #47511990

    What is a "truly new task"? Does there exist such a thing? What's an example of one? Everything we do builds on top of what's already been done. When I write a new program, I'm com…

  8. comment
    Comment #47495693

    I don't understand why their "Instant Grep + roundtrip to us-east-1" is so slow. First of all, the round-trip latency should not be nearly so bad to us-east-1. But second, and much…

  9. comment
    Comment #47458011

    On those terms, they also wasted a lot of cash. 90% of it went to candidates who lost (or opposing candidates who won).

  10. comment
    Comment #47447937

    In fact, looking at the blog post, the agent orchestrating 16 GPUs is half as efficient as the agent using 1 GPU in GPU-time. Since it uses 16 GPUs to reach the same result as 1 GP…

  11. comment
    Comment #47447893

    Yeah, assuming there's no active monitoring during the training runs, you can trivially give the agent an abstraction which turns "1 GPU" into "16 GPUs" that just so happens to tak…

  12. comment
    Comment #47447779

    Do you have a sense of whether these validation loss improvements are leading to generalized performance uplifts? From afar I can't tell whether these are broadly useful new ideas …

  13. comment
    Comment #47402267

    Super interesting study. One curious thing I've noticed is that coding agents tend to increase the code complexity of a project, but simultaneously massively reduce the cost of tha…

  14. comment
    Comment #47388901

    I don’t necessarily endorse the author’s broad conclusions about “AI”, but I will say that the Spotify DJ specifically is an enragingly bad product. Nothing close to the utility of…

  15. comment
    Comment #47330393

    I've no problem with the intuition. But I would hope for a lot more focus in the marketing materials on proving the (statistical) correctness of the implementation. 15% better infe…

  16. comment
    Comment #47327411

    OK... we need way more information than this to validate this claim! I can run Qwen-8B at 1 billion tokens per second if you don't check the model's output quality. No information …

  17. comment
    Comment #47316507

    More likely: this is a transitional phase where our previously hard problems become easy, and we will soon set our sights on new and much harder problems. The pinnacle of creative …

  18. comment
    Comment #47268390

    I did, and yet I also felt more relaxed reading it than I am reading most blog entries posted on here. I didn't feel like I had to guard against my time being wasted by vacuous LLM…

  19. comment
    Comment #47267540

    Being wealthy solves virtually all problems of consumption, so the invisible hand provides new problems to serve the market need. Beautiful, really.

  20. comment
    Comment #47213208

    Why should it be? The agent session is a messy intermediate output, not an artifact that should be part of the final product. If the "why" of a code change is important, have your …

  21. comment
    Comment #47209178

    Making those tools first-class primitives is good for (human) UX: you see the diffs inline, you can add custom rules and hooks that trigger on certain files being edited, etc.

  22. comment
    Comment #47187311

    If I had to bet, there will be some kind of face-saving climbdown by the end of next week. But all I can do right now is read the words on the page.

  23. comment
    Comment #47187150

    I don't think he got it backwards, at least if Hegseth's statement is accurate. AWS, GCP, etc. all do business with DoD. If they, as DoD contractors, are no longer allowed to do bu…

  24. comment
    Comment #47155095

    More generally, Anthropic's reliability track record for a company which claims to have solved coding is astonishingly poor. Just look at their status page - https://status.claude.…

  25. comment
    Comment #47155012

    I have little doubt where things are going, but the irony of the way they communicate versus the quality of their actual product is palpable. Claude Code (the product, not the unde…