Viewing profile — nabakin
nabakin
HN member- Joined
- Wed, Dec 06, 2017, 8:27 AM UTC
- HN karma
- 2,183
- Public activity
- 667 items
- HN profile
- View on Hacker News ↗
About nabakin
No profile information was provided.
Recent public activity
-
comment
Comment #48779207
Are you running qwen3.6-27b on one 3090 with your KV cache at q4? Ime there is significant long-context recall accuracy degradation at that precision. I prefer putting the KV cache…
-
comment
Comment #48154149
Then you can sandbox
-
comment
Comment #47909849
I think they were mistaken or maybe they were just referring to inference because I don't see anyone making that claim and it would be quite the news.
-
comment
Comment #47897500
Probably because you said you used DeepSeek. People don't want to see AI in the comments and don't trust AI responses.
-
comment
Comment #47893719
Yes, that's the footnote from citation [5].
-
comment
Comment #47892830
> Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. That is a huge claim to make with no evidence. I researched what you said, and I have found no sta…
-
comment
Comment #47820684
> When it comes to information transfer and processing, light can do things that electricity can’t. Photons — particles of light — are far zippier than electrons at working their w…
-
comment
Comment #47620051
Dunno but there's a PR for it. Probably also more performant than Modular.
-
comment
Comment #47619130
If OP meant they have the fastest implementation of Gemma 4 on Blackwell at the moment, I guess that is technically true. I doubt that will hold up when TensorRT-LLM finishes their…
-
comment
Comment #47619074
I know Arc AGI 2 has a private test set and they have a good amount of results[0] but it's not a conventional benchmark. Looking around, SWE Rebench seems to have decent protection…
-
comment
Comment #47617896
It's easy to game and human evaluation data has its trade-offs, but it's way easier to fake public benchmark results. I wish we had a source of high quality private benchmark resul…
-
comment
Comment #47617646
Faster than TensorRT-LLM on Blackwell? Or do you not consider TensorRT-LLM open source because some dependencies are closed source?
-
comment
Comment #47617258
It's referring to the Lmsys Leaderboard/Lmarena/Arena.ai[0]. It's very well-known in the LLM community for being one of the few sources of human evaluation data. [0] https://arena.…
-
comment
Comment #47617034
Public benchmarks can be trivially faked. Lmarena is a bit harder to fake and is human-evaluated. I agree it's misleading for them to hyper-focus on one metric, but public benchmar…
-
comment
Comment #47353103
And nowadays we have Debian running in a VM on Android [1] [1] https://www.zdnet.com/article/how-to-use-the-new-linux-termi...
-
comment
Comment #47235140
I would consider it reasonable if this was 4x TTFT and Throughput, but it seems like it's only for TTFT.
-
comment
Comment #46737241
And right to repair
-
comment
Comment #46291868
Ty this is great
-
comment
Comment #46284104
TIL Europe still has some presence in the Americas. Thought all of that was gone with the Monroe Doctrine
-
comment
Comment #46039997
Fyi it doesn't look like this post is listed on the frontpage anymore, even with the points it has. Not sure if it's intentional
- comment
-
comment
Comment #46036675
I understand they are similar, but I think this post adds new information to the situation. Regardless, appreciate your help moderating the site.
-
comment
Comment #46036421
Oops, that's correct, ty
-
comment
Comment #46036218
Google Translate link: https://translate.google.com/translate?tl=en&hl=en&u=https:/... Additional context: https://grapheneos.social/deck/@GrapheneOS/11557599710445618... https://g…
- story