Live data from Hacker News

Viewing profile — spindump8930

spindump8930

HN member
Joined
Thu, Mar 07, 2024, 7:39 PM UTC
HN karma
134
Public activity
50 items

About spindump8930

No profile information was provided.

Recent public activity

  1. comment
    Comment #48755413

    I agree with your recomendation, but converting a pdf to an image is by no means smaller. PDFs are much closer to SVGs then to jpegs.

  2. comment
    Comment #48667623

    > Claude and ChatGPT are both blocked in China So it's presumably cheaper than attempting to spin up your own method of circumventing the blocks.

  3. comment
  4. comment
    Comment #48415848

    Exactly. Good peer reviewers understand that you can also move down on the scaling curve, not just up. Also laughable to try a "yolo" run without validating a scaling ladder/curve.…

  5. comment
    Comment #48415830

    Can you share the specific part of this work that demonstrates better scaling than original transformers? Also note that many of the changes to that architecture, that have been pr…

  6. comment
    Comment #48415795

    That's why you do several small and medium scale tests, fit a curve, and ideally show that the trend persists at several scales. Not a single large or medium run - see the other co…

  7. comment
    Comment #48356527

    I think folks looking for more on this incident are better off reading the original threads linked elsewhere in the comments. This blog doesn't seem to add any information and is i…

  8. comment
    Comment #48136880

    Likely in this case the time vault was the collapse of Mt Gox, which has now recently been paying back holders.

  9. comment
    Comment #48068454

    Some combination of reporting bias given concerns about LLM security capabilities and actual new vulnerabilities found with LLM assistance. Even if exploits and outages are unrelat…

  10. comment
    Comment #48067495

    It's very common if you improperly seed, as others in the thread brought up! Or in your framing, as rare as earth getting hit if it were surrounded by a sci-fi density asteroid fie…

  11. comment
    Comment #47978421

    Sure, this is cute and interesting, but there's no validation or baselines and those examples are not particularly compelling. The o3 example just lists some terms!

  12. comment
    Comment #47948233

    Between the neo and the chances for privacy respecting local model inference, all the new apple hardware has me excited.

  13. comment
    Comment #47948184

    That artificial analysis page has some great references for this, thanks for sharing.

  14. comment
    Comment #47940085

    Remember that models on different inference platforms might not necessarily give exactly the same results, adding another axis of non-determinism to development. Things like quanti…

  15. comment
    Comment #47939938

    Any more context on the copilot training note? More pointers would be very interesting, but we'd need to keep in mind how many different underlying models were (are?) branded as co…

  16. comment
    Comment #47894542

    > The researchers tested five LLMs: OpenAI’s GPT-4o (before the highly sycophantic and since-sunset GPT-5) Interesting, I always thought the sycophancy peaked with 4o and the assoc…

  17. comment
    Comment #47894489

    Hopefully this money means more compute infrastructure to help Anthropic counter the efficiency changes that have created this perceived downtrend in claude quality.

  18. comment
    Comment #47894459

    Having known some folks who did recurse, I think places like this want to select for those who consider coding a type of craft or art or self-expression. You can use LLMs, but stan…

  19. comment
    Comment #47783580

    Not clear that they even have any GPUs yet: > Allbirds, which will be renamed “NewBird AI,” said it executed a $50 million deal with an unnamed institutional investor to acquire “h…

  20. comment
    Comment #47783543

    Yes, the paper itself tells a different story than the bullet points in this article.

  21. comment
    Comment #47783530

    The article seems quite editorialized, shifting between describing "large-scale AI models" and "neural network-based approaches". The underlying paper itself is more precise, compa…

  22. comment
    Comment #47694397

    Yes, it's far more certain that meta released this, which is less convincing on evals, as a result of the mythos previews.

  23. comment
    Comment #47694383

    Re: changes, there's been enormous turnover in AI organizations, and in theory this one was developed by a "new" org. Whether that means less or more benchmaxxing is anyone's guess…

  24. comment
    Comment #47694362

    Spending tons of money on Claude and the recent token benchmarks came WELL after Meta's huge investments in compute infrastructure for AI as well as the long history of language mo…

  25. comment
    Comment #47660603

    Only for poor quality systems. Unfortunately there are many systems that tried to make easy hype, but are the equivalent of an ML 101 classifier class project. If one measures for …