Live data from Hacker News

Viewing profile — kpw94

kpw94

HN member
Joined
Fri, Apr 21, 2023, 6:03 AM UTC
HN karma
997
Public activity
142 items

About kpw94

No profile information was provided.

Recent public activity

  1. comment
    Comment #49261600

    Interesting project. 2 gut feeling concerns: - Strict subset of go might be confusing to an agent actually (trying to use unavailable go features) - So -> c11 source to source comp…

  2. comment
    Comment #49253396

    Agree that go is the best due to its main design goal: A language that's simple for any programmer fitting that definition https://news.ycombinator.com/item?id=30688969 . > "They’r…

  3. story
  4. comment
    Comment #49044701

    > OpenAI and Anthropic, which are gearing up for potentially massive IPOs, did not sign the letter. Not anymore, OpenAI did sign it: https://www.microsoft.com/en-us/corporate-respo…

  5. story
  6. comment
    Comment #48811421

    Yeah definitely. I've recently commented on that: https://news.ycombinator.com/item?id=48557890

  7. comment
    Comment #48810299

    In the context of local LLMs on limited hardware I've ran to the exact same conclusion: "tok/s" isn't the most useful metric when my personal North star metric, given my fixed hard…

  8. comment
    Comment #48722290

    > What it does: > > --jinja for tool calling support Pretty sure this flag hasn't done anything for a while. It's enabled by default since ~November of last year

  9. comment
    Comment #48675696

    The huge spike of "lk-99" in science & frontier tech is amusing... This is cool concept, would love a positive/negative sentiment computed for each comment that refers to a given w…

  10. story
  11. comment
    Comment #48558655

    Thanks! Super helpful. I do use it the same way as you're describing on personal projects at home, in a very crude manner (pasting code snippets in llama server web UI prompt. Next…

  12. comment
    Comment #48557890

    > About the generation speed: ~100-150 t/s on the RTX 5090 and ~40 t/s on the Mac Curious if you can share the prefill speed too? I run locally on a crappy desktop (some AMD iGPU w…

  13. comment
    Comment #48545489

    > gemma (unsloth/gemma-4-26B-A4B-it-GGUF) models Since you're running quantized (at UD-Q4_K_XL) , check out the "qat" models (unsloth/gemma-4-26B-A4B-it-qat-GGUF) ! - https://huggi…

  14. comment
    Comment #48543903

    I did the opposite switch: In ~2015 got an Xbox one, as a media center it was an awesome experience: Kinect voice control to play/pause and other things way before Google home/Amaz…

  15. comment
    Comment #48534109

    > What's hard to figure out here? Negative externalities are hard to figure out. Since parent mentions "toxic byproduct": Say you're the company that invented Teflon pans. you made…

  16. story
  17. comment
    Comment #48152386

    That seems a very risky assumption for any car (self driving or human driver) during flash floods. "Turn around don't drown": You think you know how deep it is under because you've…

  18. comment
    Comment #48124064

    > And the author is correct (while the phrasing is a bit weird.) Right, that's just a description of the https://en.wikipedia.org/wiki/Baumol_effect

  19. comment
    Comment #48038350

    Speculative execution techniques in software & hardware exist everywhere, - Speculative multi threading - Data Value Speculation - Speculative Memory Disambiguation - Runahead Exec…

  20. comment
    Comment #47868480

    My non-controversial theory: It's all the attention-span-shortening stuff. - tech apps starting with infinite scroll (facebook, 9gag, Instagram, etc.) - media/tech shortened conten…

  21. comment
    Comment #47866765

    But isn't the prefill speed the bottleneck in some systems* ? Sure it's order of magnitude faster (10x on Apple Metal?) but there's also order of magnitude more tokens to process, …

  22. comment
    Comment #47866110

    When you say tok/s here are you describing the prefill (prompt eval) token/s or the output generation tok/s? (Btw I believe the "--jinja" flag is by default true since sometime lat…

  23. comment
    Comment #47855693

    Right, they're not the only FAANG company for which we know they're doing it: https://news.ycombinator.com/item?id=46318494

  24. comment
    Comment #47721139

    Some might be tempted to brush aside that Server Linux threat model is very different from Desktop Linux (to snarkily reply "we'll it's powering a vast majority of GDP via all of A…

  25. comment
    Comment #47629587

    > I don't know how to force this issue as a European. There are just too many levels of abstraction between me and Brussels. > EU moves so much faster when it comes to regulations …