Live data from Hacker News

Viewing profile — BoredomIsFun

BoredomIsFun

HN member
Joined
Thu, Nov 13, 2025, 1:32 PM UTC
HN karma
63
Public activity
150 items

About BoredomIsFun

No profile information was provided.

Recent public activity

  1. comment
    Comment #49247736

    yes. you can even parallelize two cards and get 1.7 times the speed.

  2. comment
    Comment #49247722

    It is very bad with any svgs.

  3. comment
    Comment #49247709

    Throw in 5060ti. By the way 5090 is 32 GiB.

  4. comment
    Comment #49247694

    True, but local setups can run LLM requests in parallel too. In this case efficiency gap is much narrower.

  5. comment
    Comment #49247669

    There is a finetune Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO which seems to be as good at coding as vanilla Qwen, but way, way better at creative writing than Qwen an…

  6. comment
    Comment #49240478

    A good plenty of counterexamples on eqbench.com.

  7. story
  8. comment
    Comment #49121026

    eerily looks like output of LLM with wrong chat template, when EOM is not detected and all kind plausible trash starts streaming out of it.

  9. comment
    Comment #48988993

    they, rightfully so, have no faith in LLMs.

  10. comment
    Comment #48988974

    West has though a history of supporting regimes far more hideous than Chinese, and ready to push narrative whenever possible. It is well known idea BTW in political thought that de…

  11. comment
    Comment #48938180

    > especially base ones Did you actually try them? I did.They generated even more "slopey" text than instruction-tuned ones.

  12. comment
    Comment #48733744

    /r/localllama is not like that at all.

  13. comment
    Comment #48722954

    > I get hallucinated tool call parameters and bizarre invocations tweaking sampler might help

  14. comment
    Comment #48092197

    It would be true, if model providers did not throttle their models. I do not have definitive proof they do but the rumors are abundant.

  15. comment
    Comment #47997293

    > Chemical reactions are just math. No, it is quantum mechanics. Physical world is not reducible to math, it has been long proven since early 20th century.

  16. comment
    Comment #47984424

    LFM models I've tried all seemed to be suffering from serious coherence issues. I found Gemmas the best at tasks requiring rock solid coherent output; even Qwen's not comparable.

  17. comment
    Comment #47974887

    > Qwen 3.6 burns it to the ground. Not for creative writing or NLP.

  18. comment
    Comment #47873942

    It feels like a pointless conversation, if no sampler settings (min_p, temperature etc.) mentioned.

  19. comment
    Comment #47623270

    > An LLM is a router and completely stateless aside from the context you feed into it. Not the latest SSM and hybrid attention ones.

  20. comment
    Comment #47623259

    good old illustrtation: https://www.ml6.eu/en/blog/large-language-models-to-fine-tun... The it- one is the yellow smiling dot, the pt- is the rightmost monster head.

  21. comment
    Comment #47593155

    > If I offend anyone I will not be apologising for it. What you said is simply counterfactual, so no reason to be offended.

  22. comment
    Comment #47593089

    Asimov is a widespread lastname in ex-USSR, esp. Central Asia. I personally know three unrelated Asimovs.

  23. comment
    Comment #47540601

    > Local model enthusiasts often assume that running locally is more energy efficient than running in a data center, It is a well known 101 truism in /r/Localllama that local is rar…

  24. comment
    Comment #47527291

    Hmm...no. These two things are orthogonal. Regardless, Olmo are opensource.

  25. comment
    Comment #47527215

    1 and 3 contradict each other. Last thing people need is anti-AI hysteria.