Live data from Hacker News

Viewing profile — nabakin

nabakin

HN member
Joined
Wed, Dec 06, 2017, 8:27 AM UTC
HN karma
2,183
Public activity
667 items

About nabakin

No profile information was provided.

Recent public activity

  1. comment
    Comment #48779207

    Are you running qwen3.6-27b on one 3090 with your KV cache at q4? Ime there is significant long-context recall accuracy degradation at that precision. I prefer putting the KV cache…

  2. comment
    Comment #48154149

    Then you can sandbox

  3. comment
    Comment #47909849

    I think they were mistaken or maybe they were just referring to inference because I don't see anyone making that claim and it would be quite the news.

  4. comment
    Comment #47897500

    Probably because you said you used DeepSeek. People don't want to see AI in the comments and don't trust AI responses.

  5. comment
    Comment #47893719

    Yes, that's the footnote from citation [5].

  6. comment
    Comment #47892830

    > Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. That is a huge claim to make with no evidence. I researched what you said, and I have found no sta…

  7. comment
    Comment #47820684

    > When it comes to information transfer and processing, light can do things that electricity can’t. Photons — particles of light — are far zippier than electrons at working their w…

  8. comment
    Comment #47620051

    Dunno but there's a PR for it. Probably also more performant than Modular.

  9. comment
    Comment #47619130

    If OP meant they have the fastest implementation of Gemma 4 on Blackwell at the moment, I guess that is technically true. I doubt that will hold up when TensorRT-LLM finishes their…

  10. comment
    Comment #47619074

    I know Arc AGI 2 has a private test set and they have a good amount of results[0] but it's not a conventional benchmark. Looking around, SWE Rebench seems to have decent protection…

  11. comment
    Comment #47617896

    It's easy to game and human evaluation data has its trade-offs, but it's way easier to fake public benchmark results. I wish we had a source of high quality private benchmark resul…

  12. comment
    Comment #47617646

    Faster than TensorRT-LLM on Blackwell? Or do you not consider TensorRT-LLM open source because some dependencies are closed source?

  13. comment
    Comment #47617258

    It's referring to the Lmsys Leaderboard/Lmarena/Arena.ai[0]. It's very well-known in the LLM community for being one of the few sources of human evaluation data. [0] https://arena.…

  14. comment
    Comment #47617034

    Public benchmarks can be trivially faked. Lmarena is a bit harder to fake and is human-evaluated. I agree it's misleading for them to hyper-focus on one metric, but public benchmar…

  15. comment
    Comment #47353103

    And nowadays we have Debian running in a VM on Android [1] [1] https://www.zdnet.com/article/how-to-use-the-new-linux-termi...

  16. comment
    Comment #47235140

    I would consider it reasonable if this was 4x TTFT and Throughput, but it seems like it's only for TTFT.

  17. comment
    Comment #46737241

    And right to repair

  18. comment
    Comment #46291868

    Ty this is great

  19. comment
    Comment #46284104

    TIL Europe still has some presence in the Americas. Thought all of that was gone with the Monroe Doctrine

  20. comment
    Comment #46039997

    Fyi it doesn't look like this post is listed on the frontpage anymore, even with the points it has. Not sure if it's intentional

  21. comment
  22. comment
    Comment #46036675

    I understand they are similar, but I think this post adds new information to the situation. Regardless, appreciate your help moderating the site.

  23. comment
    Comment #46036421

    Oops, that's correct, ty

  24. comment
    Comment #46036218

    Google Translate link: https://translate.google.com/translate?tl=en&hl=en&u=https:/... Additional context: https://grapheneos.social/deck/@GrapheneOS/11557599710445618... https://g…

  25. story