I find the situation the big LLM players find themselves in quite ironic. Sam Altman promised (edit: under duress, from a twitter poll gone wrong) to release an open source model at the level of o3-mini to catch up to the perceived OSS supremacy of Deepseek/Qwen. Now Qwen3’s release makes a model that’s “only” equivalent to o3-mini effectively dead on arrival, both socially and economically.
Qwen3: Think deeper, act faster
241–250 of 412 posts
Re: Qwen3: Think deeper, act faster
#242Earlier quoted context omitted.
Without structural assumptions, there is no necessity - only observed regularity. Necessity literally does not exist. You will never find it anywhere. Hume figured this out quite a while ago and Kant had an interesting response to it. Think the lack of “necessity” is a problem? Try to find “time” or “space” in the data. Data by itself is useless. It’s interesting to see peoples’ reaction to this.
@whatnow37373 — Three sentences and you’ve done what a semester with Kritik der reinen Vernunft couldn’t: made the Hume-vs-Kant standoff obvious. The idea that “necessity” is just the exhaust of our structural assumptions (and that data, naked, can’t even locate time or space) finally snapped into focus. This is exactly the kind of epistemic lens-polishing that keeps me reloading HN.
Re: Qwen3: Think deeper, act faster
#243Earlier quoted context omitted.
You really had me until the last half of the last sentence.
The plural of anecdote is data.
Had an outright genuine guffaw at this one, bravo.
Re: Qwen3: Think deeper, act faster
#244I’m most excited about Qwen-30B-A3B. Seems like a good choice for offline/local-only coding assistants. Until now I found that open weight models were either not as good as their proprietary counterparts or too slow to run locally. This looks like a good balance.
It would be interesting to try, but for the Aider benchmark, the dense 32B model scores 50.2 and the 30B-A3B doesn't publish the Aider benchmark, so it may be poor.
Re: Qwen3: Think deeper, act faster
#245Earlier quoted context omitted.
Except, very literally, data is a collection of single points (ie what we call "anecdotes").
Except that the plural of anecdotes is definitely not data, because without controlling for confounding variables and sampling biases, you will get garbage.
Re: Qwen3: Think deeper, act faster
#246As is now traditional for new LLM releases, I used Qwen 3 (32B, run via Ollama on a Mac) to summarize this Hacker News conversation about itself - run at the point when it hit 112 comments. The results were kind of fascinating, because it appeared to confuse my system prompt telling it to summarize the conversation with the various questions asked in the post itself, which it tried to answer. I don't think it did a g…
Qwen3 is impressive in some aspects but it thinks too much!
Qwen3-0.6b is showing even better performance than Llama 3.2 3b... but it is 6x slower.
The results are similar to Gemma3 4b, but the latter is 5x faster on Apple M3 hardware. So maybe, the utility is to run better models in cases where memory is the limiting factor, such as Nvidia GPUs?
[1] github.com/hbbio/nanoagent
Re: Qwen3: Think deeper, act faster
#247Re: Qwen3: Think deeper, act faster
#248Something that interests me about the Qwen and DeepSeek models is that they have presumably been trained to fit the worldview enforced by the CCP, for things like avoiding talking about Tiananmen Square - but we've had access to a range of Qwen/DeepSeek models for well over a year at this point and to my knowledge this assumed bias hasn't actually resulted in any documented problems from people using the models. Asid…
Right now these models have less censorship than their US counterparts. With that said, they're in a fight for dominance so censoring now would be foolish. If they win and establish a monopoly then the screws will start to turn.
Re: Qwen3: Think deeper, act faster
#249Earlier quoted context omitted.
I similarly have a small, simple spatial reasoning problem that only reasoning models get right, and not all of them, and which Qwen3 on max reasoning still gets wrong. > I put a coin in a cup and slam it upside-down on a glass table. I can't see the coin because the cup is over it. I slide a mirror under the table and see heads. What will I see if I take the cup (and the mirror) away?
Yup, it flunked that one. I also have a question that LLMs always got wrong until ChatGPT o3, and even then it has a hard time (I just tried it again and it needed to run code to work it out). Qwen3 failed, and every time I asked it to look again at its solution it would notice the error and try to solve it again, failing again: > A man wants to cross a river, and he has a cabbage, a goat, a wolf and a lion. If he le…
Re: Qwen3: Think deeper, act faster
#250Earlier quoted context omitted.
I don't really want it added to the training set, but eh. Here you go: > Assume I have a 3D printer that's currently printing, and I pause the print. What expends more energy, keeping the hotend at some temperature above room temperature and heating it up the rest of the way when I want to use it, or turning it completely off and then heat it all the way when I need it? Is there an amount of time beyond which the ans…
What kind of answer do you expect? It all depends on the hotend shape and material, temperature differences, how fast air moves in the room, humidity of the air, etc.
No it doesn't.