Viewing profile — Lwerewolf
Lwerewolf
HN member- Joined
- Fri, Jul 17, 2015, 9:44 PM UTC
- HN karma
- 95
- Public activity
- 68 items
- HN profile
- View on Hacker News ↗
About Lwerewolf
Recent public activity
-
comment
Comment #49173619
Poolside's stuff (US) is pretty good.
-
comment
Comment #49168595
You don't have to. Chatgpt-sol-high already does that for me. Could be the extra instructions or the base system prompt or whatever, point is - it already does it.
-
comment
Comment #49167129
The MI350p exists and should run a decent quant (say, the ~96GB antirez mix) well, but you can get two rtx pro 6000s for one of these, or 8x (actually more) r9700 + probably the ge…
-
comment
Comment #49162555
Almost reads like state policy-induced FOMO.
-
comment
Comment #49125069
Just tried the preview on my little test codebase and a "check this out and tell me what you think" prompt used over double the tokens of the previous iteration, but it was a lot m…
-
comment
Comment #49123998
I've had it run to ~400k when debugging "obscure" (to it) sequences. Wouldn't recommend more.
-
comment
Comment #49053589
Might wanna try it now, seems to have been largely fixed. Check huggingface threads and reddit. Works for me, very memory hungry and PP speed drops off a cliff around 200k context …
-
comment
Comment #49026903
...they'd messed something up in that quant, apparently fixed promptly. ds4 now supports it as well and it's... well, context-limited (<=250k on a 128gb machine) but it positively …
-
comment
Comment #49021154
Tool calling in pi completely broke with the updated q4 gguf (spinquant-less). Guessing it'll take some time.
-
comment
Comment #49013560
Right: https://huggingface.co/poolside/Laguna-S-2.1-FP8/discussions... Wait time.
-
comment
Comment #49011678
Well, just ran said gguf on the GeneralsX codebase with a pretty open-ended "Explain this codebase to me, and the general game loop." prompt, and... Let me also look at the GameLog…
-
comment
Comment #49008123
Deleted earlier, didn't see you post, pasting here: /* Just started testing with the gguf (with gpu offload, m5 max 128gb), q4_k_m, running seemingly well. Speed is initially sligh…
- comment
-
comment
Comment #48999012
This: https://github.com/Blaizzy/mlx-lm/tree/pc/add-lg ...and this is what I should probably wait for (not sure why it's in vlm): https://github.com/Blaizzy/mlx-vlm/tree/pc/laguna-…
-
comment
Comment #48997674
nvfp4 mlx, literally barebones pi. edit: on bigger tests, got it to loop pretty easily unfortunately, probably local settings.
-
comment
Comment #48997436
Almost like a built-in heavyweight harness.
-
comment
Comment #48997399
Testing it now. At the very least, competitive with DS4-Flash indeed. On my small (and per Sol's words, _very_ semantically dense) C test codebase, it found things that only gpt-5.…
-
comment
Comment #48976957
Pretty sure this might be a duplicate. Regardless, tried the 1bit bonsai 27b gguf three different ways - their llama.cpp fork (prism, was it) on 2 machines (1255u/16g, m5 max/128g)…
-
comment
Comment #48976629
Not too sure about the engines as of late. Bigger and heavier vehicles - yes, but still mostly rural/highway, and at way lower speeds than the EU. Overall, IMO it's just the distan…
-
comment
Comment #48926249
All show no go is the trend these days - and not just with cars.
-
comment
Comment #48892979
Only guarantee that you can get is the sandbox in which it operates. The model itself is a slot machine and can result in anything, and if its sandbox is nonexistent... here's one …
-
comment
Comment #48864354
187 N/A BSFC @ 2000rpm and open throttle. Tried emulating a DI 2GR-FXE. Seems a bit optimistic, but still fun to play with.
-
comment
Comment #48789298
Re: energy specifically - I think "psychological state" is the main thing. Once there's a will, there's a way. Make it not draining on your psyche to explore new stuff and you're g…
-
comment
Comment #48784294
The ones in cars need to be heated up quite a bit in order to work, and you still need reference air. Otherwise, I'm pretty sure that CO2 isn't a problem but rather an indication o…
-
comment
Comment #48745741
Same with ds4-flash. Low power is 13t/s vs 26t/s, power usage is ~30w vs ~100-120w. I still use high power because my m5 max is basically a server with built-in UPS and screen (and…