Live data from Hacker News

Viewing profile — jakswa

jakswa

HN member
Joined
Sun, Aug 19, 2012, 11:46 PM UTC
HN karma
237
Public activity
122 items

About jakswa

https://jake.town

Recent public activity

  1. comment
    Comment #49246456

    I'll back up your smaller claim, but be specific that it's UD-Q4_K_XL size: - muse glimmer: 15.9GB - qwen 3.6 27B: 17.6GB My video card is so close to its limit that these GB thres…

  2. comment
    Comment #49245649

    it's here! https://news.ycombinator.com/item?id=49245575

  3. comment
    Comment #49244694

    I'm listening to pelican sounds on youtube while I wait for Simon.

  4. comment
    Comment #49244573

    I like the tabletop RPG use case, and wanted to say: If your hardware likes it you should check out Gemma 4 for creative DMing use case. I found it to be much better at holding the…

  5. comment
    Comment #49243793

    some support already merged, and I verified in a local build that it runs (cannot get MTP params working tho, about ~40 tok/s on my beefy 800GB/s 7900XT w/ 20GB VRAM). https://gith…

  6. comment
    Comment #49243748

    Q3 results: unsloth/Muse-Glimmer-30B-GGUF:UD-Q3_K_XL gets down to 15.6GB VRAM and full context (131k) on the 4 parallel slots. Prompt/generation speeds about the same. Overall feel…

  7. comment
    Comment #49243581

    Another candidate for the 7900XT (20GB VRAM) I got sitting around. I pulled latest llama.cpp (targeting vulkan during build) after seeing a muse PR merged a few hours ago, and unsl…

  8. comment
    Comment #49198429

    > Every source file is summarized once into a short description of what it does. I admit I haven't had a chance to read the whole README, but wanted to get down my hesitation after…

  9. comment
    Comment #49198292

    I remember trying out a zombie running app ~15yrs ago and am a fan of the concept. That was a podcast pretty much, audio-only storyline in your ears describing proximity/urgency/et…

  10. comment
    Comment #49197927

    I can't begin to picture how much AI slop they are having to filter out. I wonder if they are going to entertain some UX to save people disappointment. "The project should be non-t…

  11. comment
    Comment #49196076

    Hell yeah I got people to move in. Nice job getting it to walk me through those intro steps.

  12. comment
    Comment #49186908

    In case anyone else is curious, local Gemma4 12B Q6 XL really struggles to make use of this for simple goals like a daily briefing after an MCP tool call. I don't think there's a s…

  13. comment
    Comment #49184820

    Seems like a good candidate to replace my home-grown llama.cpp web UI clone. This comment used to be a gripe about their provider UX not giving an example URL (realized /v1 is expe…

  14. comment
    Comment #49112911

    my guess: their new-ish "PR stacks" feature? edit: yes i bet it's https://github.blog/changelog/2026-07-30-stacked-pull-reques...

  15. comment
    Comment #49078159

    what quantization? FP4?

  16. comment
    Comment #48929659

    They've got an openai + anthropic compatible endpoints. I got far enough to run some tests on the openai endpoint, albeit with some finagling (their /models list is empty, my tool …

  17. comment
    Comment #48927057

    seems pretty dang snappy and I like it's tone/personality so far. > look at today's hackernews frontpage and generate me a daily briefing report (create an artifact) to read later …

  18. comment
    Comment #48897849

    why does this website absolutely wreck my browser

  19. comment
    Comment #48859777

    what model did you try it with? I agree and also push back a bit: How will we know when the LLMs reach the point of handling it if no one takes the leap? I applaud more people slud…

  20. comment
    Comment #48849472

    I'll agree and expand on "weird restrictions" -- I used to check the claude usage graphs multiple times a day to see where I'm at on my weekly budget. With gpt 5.5 I don't think I'…

  21. comment
    Comment #48836659

    Probably not _as_ good, but you can run gemma 4 for the ears/brain (accepts audio input), and kokoro TTS for the mouth. You want something like silero VAD sitting in front of the L…

  22. comment
    Comment #48827862

    The best project I found to throw it at was cloning llama-server's web UI essentially in one shot. I'm not sure what I'll do with 5 extra days, maybe try to imagine some complex fe…

  23. comment
    Comment #46620907

    https://jake.town

  24. comment
    Comment #41916652

    Lately I've been following https://loco.rs/ as it aims for a rails-like experience, complete with generators for workers, controllers, etc. I've only had time to experiment but it'…

  25. comment
    Comment #41642178

    The car prototype reminded me of Spy Hunter graphics, but I couldn't remember that NES game's name at first. Sent me on a nice nostalgia dive!