Viewing profile — unrvl22
unrvl22
HN member- Joined
- Fri, May 15, 2026, 9:16 AM UTC
- HN karma
- 255
- Public activity
- 16 items
- HN profile
- View on Hacker News ↗
About unrvl22
No profile information was provided.
Recent public activity
-
comment
Comment #49221056
its kinda crazy with literally no guardrails and a goal, the extremes these AI models can actually go to.
-
comment
Comment #48924232
a good take.
-
comment
Comment #48878889
Something like this is nice, where instead of having 1 model with X active experts, you have 10 different models, all small and dense, trained on specific information. and loaded o…
-
comment
Comment #48783095
inference is only memory bandwidth limited when targeting higher tps / high single stream tps. the weights only need to be moved across once per forward pass, when you batch say 10…
-
comment
Comment #48783079
that 213 wasn't achieved when saturated though. was probably more like 30 tps per stream when doing 2.6k tps.
-
comment
Comment #48783053
MI355X can perform FP6 operations with the same speed as their FP4 (unique to AMD) - people should be making MXFP6 quants which would be pretty much lossless, and much closer to FP…
-
comment
Comment #48570282
look at benchmarks, use the model yourself. Im usually first to call BS on every chinese model that says they are as good as Opus. this is finally the first one that actually is. I…
-
comment
Comment #48568423
the 2 I mentioned both have a fairly large following, who run benchmarks and absolutely will spot issues.
-
comment
Comment #48568371
I cancelled my claude sub after realizing I can burn 300m tokens a day of this quality, for $50 a month.
-
comment
Comment #48568360
Why aren't more people talking about this? It's literally Opus 4.7 quality stupid prices. I know providers who are offering this at unlimited tokens for $50 a month. Some are even …
-
comment
Comment #48528372
The municipality of Rio de Janeiro (via its IT company IplanRIO) released Rio-3.5-Open-397B, presented as a homegrown Qwen3.5 fine-tune that beats comparable open models on benchma…
- story
-
comment
Comment #48525160
AI slop. looks basic as hell.
-
comment
Comment #48448772
why is deepseek v4 pro a lot lower than flash? where is mimo 2.5?
-
comment
Comment #48161697
I had no idea those dan and the team were aussies! damn nice, we dont really seem to shine in tech on the world stage.
-
comment
Comment #48146597
love how it loads instantly and feels smooth. imo useless but still cool