Viewing profile — naasking
naasking
HN member- Joined
- Fri, Sep 09, 2011, 2:29 PM UTC
- HN karma
- 12,667
- Public activity
- 8,642 items
- HN profile
- View on Hacker News ↗
About naasking
No profile information was provided.
Recent public activity
-
comment
Comment #49245331
People say Qwen overthinks because they analyzed the thinking traces, and Qwen finds the answer relatively quickly but then second guesses itself multiple times for another 20,000+…
-
comment
Comment #49212723
The continued vagueness suggests this isn't going anywhere, so I'll just conclude by saying that you've given nobody any real reason to believe this.
-
comment
Comment #49198528
No such thing has been shown.
-
comment
Comment #49198411
Typical LLM pretraining is very inefficient with data. NanoGPT slowrun shows that data can be used much more efficiently.
-
comment
Comment #49196951
I don't understand all this dismissal of Einstein. What you just described is that he published a unified theory of Brownian motion, developed a new way of conceptualizing scientif…
-
comment
Comment #49196170
I use LLMs regularly for real work. I understand the current limitations, but you're talking about existing limitations that will never be overcome, because they cannot be overcome…
-
comment
Comment #49193942
Yes, it's worse because processors clocks are much faster than memory clocks.
-
comment
Comment #49188099
Yes. That's why law is a different discipline from linguistics requiring its own rules of analysis, pedagogy, etc.
-
comment
Comment #49188088
That's just vibes then. How is that supposed to be convincing?
-
comment
Comment #49185768
> but there are many things which they are not good at which it is not cost effective or meaningful to improve Can you name a few such things so I can keep an eye on them in the co…
-
comment
Comment #49185732
> We know exactly how attention layers work and how they produce the next word as well as draw them from larger feature spaces. This is not what's meant by the statements that we d…
-
comment
Comment #49185690
The search space is far too large for a mere order of magnitude to make any difference at all.
-
comment
Comment #49185654
> we know everything about how LLMs work No we don't. That we understand the low level mechanics of a system doesn't mean we understand how any high level phenomena emerge from tho…
-
comment
Comment #49185587
I don't think that's correct. Increasing parameter count increases capabilities in all domains per the scaling laws. Models are larger than they were years ago, so capabilities in …
-
comment
Comment #49180316
Bribes are not speech though, and advertising is.
-
comment
Comment #49176445
It's interesting to claim that it would be difficult to explain why we punish successful crime more. A successful murder creates more suffering (victim's friends and family), and o…
-
comment
Comment #49167994
It does matter though. If you want to murder someone by hitting them with a plushie, you're not going to get charged with attempted murder because it's not possible that that would…
-
comment
Comment #49166693
The margins for groceries are objectively thin. The only way to provide food at lower prices is to provide a worse good or service, eg. less variety, less quality, less availabilit…
-
comment
Comment #49166530
Yes, but there is little evidence this had a meaningful effect on votes.
-
comment
Comment #49166195
They're important everywhere of course, but especially on mobile. If AI researchers figure out how to offload knowledge and expertise from reasoning weights, then a core reasoning …
-
comment
Comment #49142479
> I can't see how the model could include the actual subjective human experience. People who say LLMs have subjective experience aren't saying they have human-type subjective exper…
-
comment
Comment #49142437
Even mechanistic models generate interesting discussion. How many years have we discussed Turing machines and the lambda calculus? Almost a century of great work came out of those.…
-
comment
Comment #49136060
> Dunno about the parent commenter, but I personally interpret the concept as having a hidden representation of self that is continually tended to I don't see why an LLM could not …
-
comment
Comment #49135984
Making definitive claims about whether LLMs do or do not have specific properties absolutely does require precise definitions of those properties that can be used to evaluate those…
-
comment
Comment #49135944
> The qualia themselves, even those that are quite abstract, are rooted in our physical presence and evolution. There is no objective evidence of qualia. All evidence of qualia are…