Live data from Hacker News

Viewing profile — parched99

parched99

HN member
Joined
Sun, Apr 20, 2025, 3:04 PM UTC
HN karma
3
Public activity
5 items

About parched99

No profile information was provided.

Recent public activity

  1. comment
    Comment #43747892

    I answered the question directly. IQ4_X_S is smaller, but slower and less accurate than Q4_0. The parent comment specifically asked about the QAT version. That's literally what thi…

  2. comment
    Comment #43746296

    Resolving that issue, would help reduce (not eliminate) the size of the context. The model will still only just barely fit in 16 GB, which is what the parent comment asked. Best to…

  3. comment
    Comment #43746177

    I'm aware. I was addressing the question being asked.

  4. comment
    Comment #43744568

    I think Powershell is a bad test. I've noticed all local models have trouble providing accurate responses to Powershell-related prompts. Strangely, even Microsoft's model, Phi 4, i…

  5. comment
    Comment #43744249

    I am only able to get the Gemma-3-27b-it-qat-Q4_0.gguf (15.6GB) to run with a 100 token context size on a 5070 ti (16GB) using llamacpp. Prompt Tokens: 10 Time: 229.089 ms Speed: 4…