Viewing profile — parched99
parched99
HN member- Joined
- Sun, Apr 20, 2025, 3:04 PM UTC
- HN karma
- 3
- Public activity
- 5 items
- HN profile
- View on Hacker News ↗
About parched99
No profile information was provided.
Recent public activity
-
comment
Comment #43747892
I answered the question directly. IQ4_X_S is smaller, but slower and less accurate than Q4_0. The parent comment specifically asked about the QAT version. That's literally what thi…
-
comment
Comment #43746296
Resolving that issue, would help reduce (not eliminate) the size of the context. The model will still only just barely fit in 16 GB, which is what the parent comment asked. Best to…
-
comment
Comment #43746177
I'm aware. I was addressing the question being asked.
-
comment
Comment #43744568
I think Powershell is a bad test. I've noticed all local models have trouble providing accurate responses to Powershell-related prompts. Strangely, even Microsoft's model, Phi 4, i…
-
comment
Comment #43744249
I am only able to get the Gemma-3-27b-it-qat-Q4_0.gguf (15.6GB) to run with a 100 token context size on a 5070 ti (16GB) using llamacpp. Prompt Tokens: 10 Time: 229.089 ms Speed: 4…