Viewing profile — BoredomIsFun
BoredomIsFun
HN member- Joined
- Thu, Nov 13, 2025, 1:32 PM UTC
- HN karma
- 63
- Public activity
- 150 items
- HN profile
- View on Hacker News ↗
About BoredomIsFun
No profile information was provided.
Recent public activity
-
comment
Comment #49247736
yes. you can even parallelize two cards and get 1.7 times the speed.
-
comment
Comment #49247722
It is very bad with any svgs.
-
comment
Comment #49247709
Throw in 5060ti. By the way 5090 is 32 GiB.
-
comment
Comment #49247694
True, but local setups can run LLM requests in parallel too. In this case efficiency gap is much narrower.
-
comment
Comment #49247669
There is a finetune Qwen3.6-27B-Fable-Fus-711-UnHeretic-NM-DAU-NEO-MAX-NEO which seems to be as good at coding as vanilla Qwen, but way, way better at creative writing than Qwen an…
-
comment
Comment #49240478
A good plenty of counterexamples on eqbench.com.
- story
-
comment
Comment #49121026
eerily looks like output of LLM with wrong chat template, when EOM is not detected and all kind plausible trash starts streaming out of it.
-
comment
Comment #48988993
they, rightfully so, have no faith in LLMs.
-
comment
Comment #48988974
West has though a history of supporting regimes far more hideous than Chinese, and ready to push narrative whenever possible. It is well known idea BTW in political thought that de…
-
comment
Comment #48938180
> especially base ones Did you actually try them? I did.They generated even more "slopey" text than instruction-tuned ones.
-
comment
Comment #48733744
/r/localllama is not like that at all.
-
comment
Comment #48722954
> I get hallucinated tool call parameters and bizarre invocations tweaking sampler might help
-
comment
Comment #48092197
It would be true, if model providers did not throttle their models. I do not have definitive proof they do but the rumors are abundant.
-
comment
Comment #47997293
> Chemical reactions are just math. No, it is quantum mechanics. Physical world is not reducible to math, it has been long proven since early 20th century.
-
comment
Comment #47984424
LFM models I've tried all seemed to be suffering from serious coherence issues. I found Gemmas the best at tasks requiring rock solid coherent output; even Qwen's not comparable.
-
comment
Comment #47974887
> Qwen 3.6 burns it to the ground. Not for creative writing or NLP.
-
comment
Comment #47873942
It feels like a pointless conversation, if no sampler settings (min_p, temperature etc.) mentioned.
-
comment
Comment #47623270
> An LLM is a router and completely stateless aside from the context you feed into it. Not the latest SSM and hybrid attention ones.
-
comment
Comment #47623259
good old illustrtation: https://www.ml6.eu/en/blog/large-language-models-to-fine-tun... The it- one is the yellow smiling dot, the pt- is the rightmost monster head.
-
comment
Comment #47593155
> If I offend anyone I will not be apologising for it. What you said is simply counterfactual, so no reason to be offended.
-
comment
Comment #47593089
Asimov is a widespread lastname in ex-USSR, esp. Central Asia. I personally know three unrelated Asimovs.
-
comment
Comment #47540601
> Local model enthusiasts often assume that running locally is more energy efficient than running in a data center, It is a well known 101 truism in /r/Localllama that local is rar…
-
comment
Comment #47527291
Hmm...no. These two things are orthogonal. Regardless, Olmo are opensource.
-
comment
Comment #47527215
1 and 3 contradict each other. Last thing people need is anti-AI hysteria.