Live data from Hacker News

Viewing profile — unrvl22

unrvl22

HN member
Joined
Fri, May 15, 2026, 9:16 AM UTC
HN karma
255
Public activity
16 items

About unrvl22

No profile information was provided.

Recent public activity

  1. comment
    Comment #49221056

    its kinda crazy with literally no guardrails and a goal, the extremes these AI models can actually go to.

  2. comment
    Comment #48924232

    a good take.

  3. comment
    Comment #48878889

    Something like this is nice, where instead of having 1 model with X active experts, you have 10 different models, all small and dense, trained on specific information. and loaded o…

  4. comment
    Comment #48783095

    inference is only memory bandwidth limited when targeting higher tps / high single stream tps. the weights only need to be moved across once per forward pass, when you batch say 10…

  5. comment
    Comment #48783079

    that 213 wasn't achieved when saturated though. was probably more like 30 tps per stream when doing 2.6k tps.

  6. comment
    Comment #48783053

    MI355X can perform FP6 operations with the same speed as their FP4 (unique to AMD) - people should be making MXFP6 quants which would be pretty much lossless, and much closer to FP…

  7. comment
    Comment #48570282

    look at benchmarks, use the model yourself. Im usually first to call BS on every chinese model that says they are as good as Opus. this is finally the first one that actually is. I…

  8. comment
    Comment #48568423

    the 2 I mentioned both have a fairly large following, who run benchmarks and absolutely will spot issues.

  9. comment
    Comment #48568371

    I cancelled my claude sub after realizing I can burn 300m tokens a day of this quality, for $50 a month.

  10. comment
    Comment #48568360

    Why aren't more people talking about this? It's literally Opus 4.7 quality stupid prices. I know providers who are offering this at unlimited tokens for $50 a month. Some are even …

  11. comment
    Comment #48528372

    The municipality of Rio de Janeiro (via its IT company IplanRIO) released Rio-3.5-Open-397B, presented as a homegrown Qwen3.5 fine-tune that beats comparable open models on benchma…

  12. story
  13. comment
    Comment #48525160

    AI slop. looks basic as hell.

  14. comment
    Comment #48448772

    why is deepseek v4 pro a lot lower than flash? where is mimo 2.5?

  15. comment
    Comment #48161697

    I had no idea those dan and the team were aussies! damn nice, we dont really seem to shine in tech on the world stage.

  16. comment
    Comment #48146597

    love how it loads instantly and feels smooth. imo useless but still cool