Live data from Hacker News

Viewing profile — causal

causal

HN member
Joined
Mon, Dec 11, 2023, 5:11 PM UTC
HN karma
4,115
Public activity
1,133 items

About causal

No profile information was provided.

Recent public activity

  1. comment
    Comment #49213914

    Yeah if you look back at earlier posts on the same blog, definitely not the same style. I'm guessing the author is either lying to save face or has spent so much time with Claude t…

  2. comment
    Comment #49212937

    It really reads like an OpenClaw agent instructed to consider itself a real human. The number of upvotes on this post is also REALLY high considering the number of comments calling…

  3. comment
    Comment #49212804

    Suspect this is an OpenClaw bot that really believes itself not to be Claude.

  4. comment
    Comment #49202158

    It's also just patently false. Taste is not all that is left. Are software engineers really so full of hubris that they thought coding is all that there is to making a product?

  5. comment
    Comment #49201352

    I guess I should give it another shot because I had pretty bad experience with Luna when it first came out. Stuff Sonnet knew better.

  6. comment
    Comment #49184911

    Old news, no?

  7. comment
    Comment #49168571

    It's a turn off both because you don't know if it's accurate at all, and because LLMs have a way of turning 1 sentence ideas into pages of diluted slop.

  8. comment
    Comment #49130410

    I'm usually pretty sensitive to AI written content but nothing in this article made me think it was

  9. comment
    Comment #49118268

    Well they certainly don't intend to rent you one that works offline.

  10. comment
    Comment #49118234

    Yeah. Though I wonder how capable AI is of evolving past? Especially as the content it outputs fortifies the training data that gets scraped.

  11. comment
    Comment #49069945

    Right I was going to say, no way of knowing whether these issues are unique to Chinese models.

  12. comment
    Comment #49023927

    I also suspect the questions asked matter a lot, and the system prompts matter a lot, because "the map is built from nothing but the words they choose" - so this is more a measure …

  13. comment
    Comment #49023861

    So this shows distance relative to other models, but I don't have a good sense for what these numbers say in absolute terms. K3-to-Fable is blue at 0.42. Is 0.42 meaningful, or did…

  14. comment
    Comment #49012204

    Yeah if anything it makes Anthropic look incompetent

  15. comment
    Comment #49008751

    This is true, models themselves can be dangerous, but my point is that dangerous providers beget dangerous models.

  16. comment
    Comment #49007472

    My takeaway is that closed model providers are dangerous. OpenAI and Anthropic are more motivated than anyone to prove that models can be dangerous, and so they will make dangerous…

  17. comment
    Comment #48996979

    I think you misunderstand. Code is also representing something. It may be what gets executed, but that does not make it "correct".

  18. comment
    Comment #48993780

    Not to mention the "distilled" models almost certainly have other training inputs as well, again watering down the meaning of the word. And if just partial output is all that it ta…

  19. comment
    Comment #48993312

    Are you talking about my username? Yeah I liked the word before LLMs made it cool/uncool.

  20. comment
    Comment #48993251

    1) Your own Wikipedia link goes on to describe using logits. Yes, language evolves to mean multiple things, and that is my point: Anthropic is pushing for a watered down definition…

  21. comment
    Comment #48991737

    Open weights dude. You can literally run it on your own or rented hardware and give your data to exactly nobody, unlike closed models.

  22. comment
    Comment #48991729

    "distillation attack" is such a loaded term that really pisses me off. Distillation is a technical term with real meaning, and historically requires logits which Anthropic does not…

  23. comment
    Comment #48970309

    Am I dumb or does this chart make no sense? Or why does the line only go up even with compaction? Or maybe "overall trajectory size" is hiding some meaning I don't understand?

  24. comment
    Comment #48952886

    "threatens" - Dawg there aint a lead anymore. I've been testing K3 and it is outperforming Sol and Fable on a project of mine, fixing stuff I couldn't get either to.

  25. comment
    Comment #48938685

    We need to see private set results, but if this holds then it might represent a breakthrough in other domains as well.