Viewing profile — polishgladiator
polishgladiator
HN member- Joined
- Sat, Aug 05, 2023, 11:56 PM UTC
- HN karma
- 18
- Public activity
- 9 items
- HN profile
- View on Hacker News ↗
About polishgladiator
No profile information was provided.
Recent public activity
-
comment
Comment #37027743
Something doesn't smell right. Such sloppy errors with measurement and comparison (from people who are supposedly experts?), and cageyness about answering technical questions, remi…
-
comment
Comment #37024483
Actually no -- that post shows they are not performing measurements and comparisons correctly. These are not serious people.
-
comment
Comment #37024136
OK, so this is a case of bad measurement and comparison. If you bothered to look at the llama.cpp output, you would see this line: llama_model_load_internal: offloaded 32/35 layers…
-
comment
Comment #37019545
> [...] llama.cpp is a fantastic framework to run models locally for the single-user case (batch=1) > [...] I don't think it would be particularly fair to compare and show that MKM…
-
comment
Comment #37019478
> If anyone has specific technical questions I'd be happy to answer as best I can. What is the context size for these measurements? Is it the full 4k for llama-2? And just to be cl…
-
comment
Comment #37018232
Based on the integration examples, I don't think they are simply repackaging llama.cpp Rather it looks like they are reimplementing their own quantization scheme, in such a way tha…
-
comment
Comment #37018150
I've been doing some hacking with Llama2 on an AMD 7900 XTX this weekend, using llama.cpp and q5_k_s quantization. Compared to MK600 on an RTX 4090 in their data, I am measuring hi…
-
comment
Comment #37017863
In the realm of text editors, Vim has been my unwavering companion for 25 years. It has become an integral part of my daily existence, an inseparable bond that defies the imaginati…
- comment