Live data from Hacker News

Viewing profile — practice9

practice9

HN member
Joined
Wed, Mar 28, 2018, 7:14 PM UTC
HN karma
539
Public activity
271 items

About practice9

No profile information was provided.

Recent public activity

  1. comment
    Comment #48157335

    It's because HN is in AI meta-psychosis :) Our experience is very similar except we didn't really have a review process before, and now LLMs find bugs before PRs get merged in main…

  2. comment
    Comment #45726248

    LLMs are getting quite good at reviewing the results and implementations, though

  3. comment
    Comment #45020509

    I find it hilarious/sad that the 0.5x cheaper Ergo M575 has much better design in that regard (just plastic that doesn’t degrade)

  4. comment
    Comment #44809417

    They should have used Claude Code for reviews

  5. comment
    Comment #44656486

    Humans cannot reason about code at scale. Unless you add scaffolding like diagrams and maps and … Things that most teams don’t do or half-ass

  6. comment
    Comment #44546130

    A variation of “no taxation without representation”?

  7. comment
    Comment #44288093

    The human is a bad co-author here really. I deployed lots of high performance, clean, well documented etc code generated by Claude or o3. I reviewed it wrt requirements, added test…

  8. comment
    Comment #43842510

    Well the system prompt is still the same for both models, right? Kinda points to people at OpenAI using o1/o3/o4 almost exclusively. That's why nobody noticed how cringe 4o has bec…

  9. comment
    Comment #43692246

    Kinda similar in a way to China or Russia “disappearances”

  10. comment
    Comment #43341043

    But who is the target group? Last time only some groups of enthusiasts were willing to work through bugs to even run the buggy release of Gemma Surely nobody runs this in productio…

  11. comment
    Comment #42977763

    I tried the square example from the paper mentioned with o1-pro and it had no problem counting 4 nested squares… And the 5 square variation as well. So perhaps it is just a questio…

  12. comment
    Comment #42930091

    With LLMs that even might be automated

  13. comment
    Comment #42706547

    Well none of the labs have good frontend or mobile engineers or even infra engineers Anthropic is ahead in this because they keep their UIs simplistic so the failure modes are also…

  14. comment
    Comment #42067214

    One of those guys needs to be fined for the pump & dump scheme (with SPCE: Virgin Galactic), and the other one should be investigated if he was receiving money from the Russian gov…

  15. comment
    Comment #41843477

    [flagged]

  16. comment
    Comment #41739800

    More like reparations

  17. comment
    Comment #41686224

    The interesting thing is that ship was damaged almost immediately after leaving the port, had a chance to stop in Russian ports along the way but instead is doing a tour near EU co…

  18. comment
    Comment #40347323

    It is cringe overenthusiastic, but a proper instructions/system prompt will fix that mostly

  19. comment
    Comment #40347302

    I always double-check even the most obscure facts returned by GPT-4 and have yet to see a hallucination (as opposed to Claude Opus that sometimes made up historical facts). I doubt…

  20. comment
    Comment #40259427

    Wasn't it pointing to https://x.ai just a few months ago? Interesting

  21. comment
    Comment #40098500

    [flagged]

  22. comment
    Comment #39952002

    teknium / Nous released Mistral finetunes (Hermes) that are quite great, and even published the datasets used for training. But for the worldsim I think they are really using Claud…

  23. comment
    Comment #39608844

    Interesting. I had a problem a few months ago with DNS not resolving Meta servers on my Starlink internet connection, but I was able to use the UI and the apps nonetheless, just co…

  24. comment
    Comment #39605479

    Unless it's a recent change, it works perfectly fine offline (wifi turned off). As for alternatives, there is Pico, but Quest 3 may be superior in games selection. Or go wired whic…

  25. story