Live data from Hacker News

Viewing profile — fdsjgfklsfd

fdsjgfklsfd

HN member
Joined
Fri, Sep 04, 2020, 5:00 PM UTC
HN karma
42
Public activity
65 items

About fdsjgfklsfd

No profile information was provided.

Recent public activity

  1. comment
    Comment #49124648

    Was already released before your comment: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731

  2. comment
    Comment #49124625

    I use DeepSeek V4 Flash (before this update) for most things: - OpenCode for codebase editing: python scientific computing and LLM projects - Open Interpreter Classic (python versi…

  3. comment
    Comment #49124447

    DeepSeek-V4-Flash-0731 scores higher than Fable 5 on Terminal-Bench, and Fable 5 is Mythos, correct?

  4. comment
    Comment #49124375

    Yes, they trained on my open source code and Wikipedia edits and Stack Overflow answers that I shared freely, so it's only fair that they release their models as open source and sh…

  5. comment
    Comment #48585369

    Yeah the reasoning is formatted differently and the replies are often in Chinese.

  6. comment
    Comment #48585335

    Sometimes it's faster than swyping on a phone, but mostly I use it to learn about stuff and hash out ideas while driving.

  7. comment
    Comment #47680584

    The US government is vastly more likely to break down my door and arrest me for my speech than the Chinese government is. (Because I live in the US.)

  8. comment
    Comment #47680550

    Qwen3.5-plus is quite good, Qwen3.6-plus is not.

  9. comment
    Comment #47680540

    Qwen3.5-plus has been my go-to model for non-autonomous chatbot that runs arbitrary code and shell commands on my local machine for one-off tasks (unlike Claude Code). It has no pr…

  10. comment
    Comment #47680473

    > In particular, Qwen3.5-Plus is the hosted version corresponding to Qwen3.5-397B-A17B with more production features, e.g., 1M context length by default, official built-in tools, a…

  11. comment
    Comment #47508605

    Reporting spam on GitHub requires you to click a link, specify the type of ticket, write a description of the problem, solve multiple CAPTCHAs of spinning animals, and press Submit…

  12. comment
    Comment #47291167

    https://www.thestack.technology/backlash-over-anthropic-ai-c... https://www.anthropic.com/news/detecting-and-preventing-dist...

  13. comment
    Comment #47291059

    The current biggest problem in the US is that the President is violating the Constitution with impunity

  14. comment
    Comment #47291056

    The Constitution and Founding Fathers are pretty great compared to what we have now. "At this point, Elbridge Gerry objected to Butler’s earlier-raised proposition that the clause …

  15. comment
    Comment #47291024

    Aggregators: https://fiftyplusone.news/polls/approval/president https://www.natesilver.net/p/trump-approval-ratings-nate-sil...

  16. comment
    Comment #47290980

    "Overplayed"? Did you see the actual footage of the event? The event in which people attacked Capitol Police and broke into the Capitol and tried to take power by force?

  17. comment
    Comment #47102813

    They aren't actually trying to solve any real problem.

  18. comment
    Comment #44904739

    Do you mean "all variants of the same stacked transformer architecture converge in performance"? Or do you know of tests against some other architecture? The diffusion-based LLMs?

  19. comment
    Comment #44837155

    What's openai/gpt-5 vs openai/gpt-5-chat?

  20. comment
    Comment #44828706

    I think they're just reaching the limits of this architecture and when a new type is invented it will be a much bigger step.

  21. comment
    Comment #44525271

    When I've had Grok evaluate images and dug into how it perceives them, it seemed to just have an image labeling model slapped onto the text input layer. I'm not sure it can really …

  22. comment
    Comment #44525212

    I dunno. Talking with Grok 3 about political issues, it does seem to be pretty "truth-seeking" and not biased. I asked it to come up with matter-of-fact political issues and evalua…

  23. comment
    Comment #44525173

    Hello, LLM slop.

  24. comment
    Comment #44525161

    You misspelled "principles".

  25. comment
    Comment #44525095

    I feel like they should train a dumb model that does nothing but recognize when someone has finished talking, and use that to determine when to stop listening and start responding.…