Live data from Hacker News

Viewing profile — easygenes

easygenes

HN member
Joined
Wed, Feb 03, 2016, 5:14 AM UTC
HN karma
1,393
Public activity
472 items

About easygenes

No profile information was provided.

Recent public activity

  1. comment
    Comment #49205174

    Oof, even the post-mortem promising none of it is AI slop is obvious AI slop.

  2. comment
    Comment #48936338

    That’s not what this indicates. This is the biggest and most expensive to serve, and most capable open weights model yet. They’re just pricing it in line with capabilities. Kimi al…

  3. comment
    Comment #48783240

    Was fun to see their developers make nods to Le Chaton Fat in the announcements for this on Twitter. I suspect a true "big new general-purpose" model is around the corner from them…

  4. comment
    Comment #48740650

    I'm a heavy enough user that I have both the OAI and Anth $200 plans. I always use at least 50% of my weekly Opus quota at Extra setting (meaning I use double the limit of the $100…

  5. comment
    Comment #48715458

    There are. If the kernels are nondeterministic (e.g. timing issues) there are minor changes between runs, on a single system, even with eager decode enabled (typically what tempera…

  6. comment
    Comment #48695652

    This is a strange one. We know the hardware capabilities of Cerebras force them to do aggressive REAP pruning to serve Kimi K2.6. Meaning that about 750B parameters is the upper li…

  7. comment
  8. comment
    Comment #48638673

    M5 Ultra will ship before end of year, likely. Though with current RAM shortage, likely max spec will be 256GB and in short supply. In late 2027 or early 2028, Nvidia will release …

  9. comment
    Comment #48594256

    Article reads as though written by someone who doesn't have much experience with deployments like this. Underestimates the memory needed to run with a reasonable amount of context.…

  10. comment
    Comment #48590700

    The Wired headline reframes the issue in a way that’s misleading. SK Telecom was a previously resolved issue (as in prior to Fable launch). It may have been a contributing factor, …

  11. comment
    Comment #48578878

    This headline is not what I would read from this. The numbers are more favorable than the general tone of rumors, and point towards the expected shape of a fast-growing R&D heavy b…

  12. comment
    Comment #48521149

    Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrett…

  13. comment
    Comment #48520363

    This release was rushed to hang on the coattails of the Mythos drama (“hey, sorry you can’t use Fable, but try us while you wait this weekend!”) I think they planned to release nex…

  14. comment
    Comment #48433934

    That happened a year ago when these shipped as the DGX Spark with only Linux pre installed.

  15. comment
    Comment #48433923

    Mostly a strategy move to protect the CUDA moat… Apple would take over mobile inference in a clean sweep without competition.

  16. comment
    Comment #48433733

    This is the same chip and same memory. Only difference is it is going in a laptop, so will be more thermally limited.

  17. comment
    Comment #48391828

    If I were paying API rates this year, I would have already burned through $20k in tokens. Looking forward to the costs of this level of capability coming down.

  18. comment
    Comment #48391687

    I have now also tried it on this scatter plot: https://3215535692-files.gitbook.io/~/files/v0/b/gitbook-x-p... Similarly, the 26B A4B Gemma 4 and the 35B A3B Qwen 3.6 identify it c…

  19. comment
    Comment #48391593

    They haven't made one for this new model, but Unsloth has a comprehensive quant KLD map of Gemma 4 26B A4B here: https://3215535692-files.gitbook.io/~/files/v0/b/gitbook-x-p...

  20. comment
    Comment #48391488

    I want to like the vision capabilities of the model. However, when I gave it an image which Gemma 26B A4B and Qwen 3.6 35B A3B has no problem correctly describing in detail, includ…

  21. comment
    Comment #48381166

    Have you run it through DeepSWE? I understand that's probably a high ask for this class of model, but would be interesting to see regardless. Even if it can't fully pass much, ther…

  22. comment
    Comment #48379719

    While I agree directionally, I'll caveat that "cost per token" != "cost per task". In the case of Qwen3.6 it tends to think 1.6x more than Haiku, so the cost of Haiku on the same t…

  23. comment
    Comment #48366769

    Speaking as someone who has had a DGX Spark all year and been active developing at the driver and kernel level for it and other ARM64 Linux devices the last couple of years, it's n…

  24. comment
    Comment #48366736

    Looks like RTX Spark desktop is the DGX Spark desktop, minus the expensive 200GbE Connect-X NIC. Only since the DGX Spark released, memory and nand prices have jumped, so it will l…

  25. comment
    Comment #48264082

    Claude Opus 4.7 defaults to exactly this design language for a lot of "just make me a rich html presentation page" requests without further specification.