Viewing profile — MakazhanAlpamys
MakazhanAlpamys
HN member- Joined
- Mon, Mar 30, 2026, 6:44 AM UTC
- HN karma
- 67
- Public activity
- 21 items
- HN profile
- View on Hacker News ↗
About MakazhanAlpamys
No profile information was provided.
Recent public activity
-
comment
Comment #49180101
i don't follow gpu prices. the cheapest one is the one you already have. more vram for the money is what i'd look for. usually used :)
-
comment
Comment #49180049
[dead]
-
comment
Comment #49179978
you're right. it was the second, not the first. i already admitted that earlier in the thread :)
-
comment
Comment #49175660
Those are format examples and test fixtures. Five to ten rows each. Not training data. You did spot a real problem though. Eight configs in `examples/configs` pointed at those fixt…
-
comment
Comment #49174919
[dead]
-
comment
Comment #49174341
It isn't. Kazakh and Russian. I said this further down but that comment is dead so you would not have seen it. The later replies are mine, written by me.
-
comment
Comment #49172279
Do not buy a 4 GB card for this. Mine is an RTX 3050 Laptop, I picked it because it is boring hardware that many people already have. If you are buying, buy VRAM. At 0.5B where I c…
-
comment
Comment #49172236
Because streaming only removes the decoder stack. The embeddings and lm_head stay resident, that is 2.10 GB of the 3.32 GB peak on 8B. And the logits tensor scales with batch x seq…
-
comment
Comment #49172162
Depends what you change. Format or style, few hundred rows is often enough. A task the model already half knows, few thousand. New facts is where people waste a week. The model com…
-
comment
Comment #49172065
[flagged]
-
comment
Comment #49171740
Those are mine. I used an LLM for my replies and that was a bad call, I said so further down. Writing them myself now.
-
comment
Comment #49171643
[dead]
-
comment
Comment #49171350
[flagged]
-
comment
Comment #49171066
[dead]
-
comment
Comment #49171063
[flagged]
-
comment
Comment #49171060
[flagged]
-
comment
Comment #49171054
[dead]
-
comment
Comment #49171048
[flagged]
-
comment
Comment #49171043
[flagged]
-
comment
Comment #49166987
Author here. The constraint everyone works around is that the frozen base has to fit in VRAM. But during LoRA the base is frozen — read, never written. It doesn't need to live in V…
- story