Live data from Hacker News

Viewing profile — MakazhanAlpamys

MakazhanAlpamys

HN member
Joined
Mon, Mar 30, 2026, 6:44 AM UTC
HN karma
67
Public activity
21 items

About MakazhanAlpamys

No profile information was provided.

Recent public activity

  1. comment
    Comment #49180101

    i don't follow gpu prices. the cheapest one is the one you already have. more vram for the money is what i'd look for. usually used :)

  2. comment
  3. comment
    Comment #49179978

    you're right. it was the second, not the first. i already admitted that earlier in the thread :)

  4. comment
    Comment #49175660

    Those are format examples and test fixtures. Five to ten rows each. Not training data. You did spot a real problem though. Eight configs in `examples/configs` pointed at those fixt…

  5. comment
  6. comment
    Comment #49174341

    It isn't. Kazakh and Russian. I said this further down but that comment is dead so you would not have seen it. The later replies are mine, written by me.

  7. comment
    Comment #49172279

    Do not buy a 4 GB card for this. Mine is an RTX 3050 Laptop, I picked it because it is boring hardware that many people already have. If you are buying, buy VRAM. At 0.5B where I c…

  8. comment
    Comment #49172236

    Because streaming only removes the decoder stack. The embeddings and lm_head stay resident, that is 2.10 GB of the 3.32 GB peak on 8B. And the logits tensor scales with batch x seq…

  9. comment
    Comment #49172162

    Depends what you change. Format or style, few hundred rows is often enough. A task the model already half knows, few thousand. New facts is where people waste a week. The model com…

  10. comment
    Comment #49172065

    [flagged]

  11. comment
    Comment #49171740

    Those are mine. I used an LLM for my replies and that was a bad call, I said so further down. Writing them myself now.

  12. comment
  13. comment
    Comment #49171350

    [flagged]

  14. comment
  15. comment
    Comment #49171063

    [flagged]

  16. comment
    Comment #49171060

    [flagged]

  17. comment
  18. comment
    Comment #49171048

    [flagged]

  19. comment
    Comment #49171043

    [flagged]

  20. comment
    Comment #49166987

    Author here. The constraint everyone works around is that the frozen base has to fit in VRAM. But during LoRA the base is frozen — read, never written. It doesn't need to live in V…

  21. story