Viewing profile — jakswa
jakswa
HN member- Joined
- Sun, Aug 19, 2012, 11:46 PM UTC
- HN karma
- 237
- Public activity
- 122 items
- HN profile
- View on Hacker News ↗
About jakswa
Recent public activity
-
comment
Comment #49246456
I'll back up your smaller claim, but be specific that it's UD-Q4_K_XL size: - muse glimmer: 15.9GB - qwen 3.6 27B: 17.6GB My video card is so close to its limit that these GB thres…
-
comment
Comment #49245649
it's here! https://news.ycombinator.com/item?id=49245575
-
comment
Comment #49244694
I'm listening to pelican sounds on youtube while I wait for Simon.
-
comment
Comment #49244573
I like the tabletop RPG use case, and wanted to say: If your hardware likes it you should check out Gemma 4 for creative DMing use case. I found it to be much better at holding the…
-
comment
Comment #49243793
some support already merged, and I verified in a local build that it runs (cannot get MTP params working tho, about ~40 tok/s on my beefy 800GB/s 7900XT w/ 20GB VRAM). https://gith…
-
comment
Comment #49243748
Q3 results: unsloth/Muse-Glimmer-30B-GGUF:UD-Q3_K_XL gets down to 15.6GB VRAM and full context (131k) on the 4 parallel slots. Prompt/generation speeds about the same. Overall feel…
-
comment
Comment #49243581
Another candidate for the 7900XT (20GB VRAM) I got sitting around. I pulled latest llama.cpp (targeting vulkan during build) after seeing a muse PR merged a few hours ago, and unsl…
-
comment
Comment #49198429
> Every source file is summarized once into a short description of what it does. I admit I haven't had a chance to read the whole README, but wanted to get down my hesitation after…
-
comment
Comment #49198292
I remember trying out a zombie running app ~15yrs ago and am a fan of the concept. That was a podcast pretty much, audio-only storyline in your ears describing proximity/urgency/et…
-
comment
Comment #49197927
I can't begin to picture how much AI slop they are having to filter out. I wonder if they are going to entertain some UX to save people disappointment. "The project should be non-t…
-
comment
Comment #49196076
Hell yeah I got people to move in. Nice job getting it to walk me through those intro steps.
-
comment
Comment #49186908
In case anyone else is curious, local Gemma4 12B Q6 XL really struggles to make use of this for simple goals like a daily briefing after an MCP tool call. I don't think there's a s…
-
comment
Comment #49184820
Seems like a good candidate to replace my home-grown llama.cpp web UI clone. This comment used to be a gripe about their provider UX not giving an example URL (realized /v1 is expe…
-
comment
Comment #49112911
my guess: their new-ish "PR stacks" feature? edit: yes i bet it's https://github.blog/changelog/2026-07-30-stacked-pull-reques...
-
comment
Comment #49078159
what quantization? FP4?
-
comment
Comment #48929659
They've got an openai + anthropic compatible endpoints. I got far enough to run some tests on the openai endpoint, albeit with some finagling (their /models list is empty, my tool …
-
comment
Comment #48927057
seems pretty dang snappy and I like it's tone/personality so far. > look at today's hackernews frontpage and generate me a daily briefing report (create an artifact) to read later …
-
comment
Comment #48897849
why does this website absolutely wreck my browser
-
comment
Comment #48859777
what model did you try it with? I agree and also push back a bit: How will we know when the LLMs reach the point of handling it if no one takes the leap? I applaud more people slud…
-
comment
Comment #48849472
I'll agree and expand on "weird restrictions" -- I used to check the claude usage graphs multiple times a day to see where I'm at on my weekly budget. With gpt 5.5 I don't think I'…
-
comment
Comment #48836659
Probably not _as_ good, but you can run gemma 4 for the ears/brain (accepts audio input), and kokoro TTS for the mouth. You want something like silero VAD sitting in front of the L…
-
comment
Comment #48827862
The best project I found to throw it at was cloning llama-server's web UI essentially in one shot. I'm not sure what I'll do with 5 extra days, maybe try to imagine some complex fe…
-
comment
Comment #46620907
https://jake.town
-
comment
Comment #41916652
Lately I've been following https://loco.rs/ as it aims for a rails-like experience, complete with generators for workers, controllers, etc. I've only had time to experiment but it'…
-
comment
Comment #41642178
The car prototype reminded me of Spy Hunter graphics, but I couldn't remember that NES game's name at first. Sent me on a nice nostalgia dive!