Viewing profile — causal
causal
HN member- Joined
- Mon, Dec 11, 2023, 5:11 PM UTC
- HN karma
- 4,115
- Public activity
- 1,133 items
- HN profile
- View on Hacker News ↗
About causal
No profile information was provided.
Recent public activity
-
comment
Comment #49213914
Yeah if you look back at earlier posts on the same blog, definitely not the same style. I'm guessing the author is either lying to save face or has spent so much time with Claude t…
-
comment
Comment #49212937
It really reads like an OpenClaw agent instructed to consider itself a real human. The number of upvotes on this post is also REALLY high considering the number of comments calling…
-
comment
Comment #49212804
Suspect this is an OpenClaw bot that really believes itself not to be Claude.
-
comment
Comment #49202158
It's also just patently false. Taste is not all that is left. Are software engineers really so full of hubris that they thought coding is all that there is to making a product?
-
comment
Comment #49201352
I guess I should give it another shot because I had pretty bad experience with Luna when it first came out. Stuff Sonnet knew better.
-
comment
Comment #49184911
Old news, no?
-
comment
Comment #49168571
It's a turn off both because you don't know if it's accurate at all, and because LLMs have a way of turning 1 sentence ideas into pages of diluted slop.
-
comment
Comment #49130410
I'm usually pretty sensitive to AI written content but nothing in this article made me think it was
-
comment
Comment #49118268
Well they certainly don't intend to rent you one that works offline.
-
comment
Comment #49118234
Yeah. Though I wonder how capable AI is of evolving past? Especially as the content it outputs fortifies the training data that gets scraped.
-
comment
Comment #49069945
Right I was going to say, no way of knowing whether these issues are unique to Chinese models.
-
comment
Comment #49023927
I also suspect the questions asked matter a lot, and the system prompts matter a lot, because "the map is built from nothing but the words they choose" - so this is more a measure …
-
comment
Comment #49023861
So this shows distance relative to other models, but I don't have a good sense for what these numbers say in absolute terms. K3-to-Fable is blue at 0.42. Is 0.42 meaningful, or did…
-
comment
Comment #49012204
Yeah if anything it makes Anthropic look incompetent
-
comment
Comment #49008751
This is true, models themselves can be dangerous, but my point is that dangerous providers beget dangerous models.
-
comment
Comment #49007472
My takeaway is that closed model providers are dangerous. OpenAI and Anthropic are more motivated than anyone to prove that models can be dangerous, and so they will make dangerous…
-
comment
Comment #48996979
I think you misunderstand. Code is also representing something. It may be what gets executed, but that does not make it "correct".
-
comment
Comment #48993780
Not to mention the "distilled" models almost certainly have other training inputs as well, again watering down the meaning of the word. And if just partial output is all that it ta…
-
comment
Comment #48993312
Are you talking about my username? Yeah I liked the word before LLMs made it cool/uncool.
-
comment
Comment #48993251
1) Your own Wikipedia link goes on to describe using logits. Yes, language evolves to mean multiple things, and that is my point: Anthropic is pushing for a watered down definition…
-
comment
Comment #48991737
Open weights dude. You can literally run it on your own or rented hardware and give your data to exactly nobody, unlike closed models.
-
comment
Comment #48991729
"distillation attack" is such a loaded term that really pisses me off. Distillation is a technical term with real meaning, and historically requires logits which Anthropic does not…
-
comment
Comment #48970309
Am I dumb or does this chart make no sense? Or why does the line only go up even with compaction? Or maybe "overall trajectory size" is hiding some meaning I don't understand?
-
comment
Comment #48952886
"threatens" - Dawg there aint a lead anymore. I've been testing K3 and it is outperforming Sol and Fable on a project of mine, fixing stuff I couldn't get either to.
-
comment
Comment #48938685
We need to see private set results, but if this holds then it might represent a breakthrough in other domains as well.