Viewing profile — easygenes
easygenes
HN member- Joined
- Wed, Feb 03, 2016, 5:14 AM UTC
- HN karma
- 1,393
- Public activity
- 472 items
- HN profile
- View on Hacker News ↗
About easygenes
No profile information was provided.
Recent public activity
-
comment
Comment #49205174
Oof, even the post-mortem promising none of it is AI slop is obvious AI slop.
-
comment
Comment #48936338
That’s not what this indicates. This is the biggest and most expensive to serve, and most capable open weights model yet. They’re just pricing it in line with capabilities. Kimi al…
-
comment
Comment #48783240
Was fun to see their developers make nods to Le Chaton Fat in the announcements for this on Twitter. I suspect a true "big new general-purpose" model is around the corner from them…
-
comment
Comment #48740650
I'm a heavy enough user that I have both the OAI and Anth $200 plans. I always use at least 50% of my weekly Opus quota at Extra setting (meaning I use double the limit of the $100…
-
comment
Comment #48715458
There are. If the kernels are nondeterministic (e.g. timing issues) there are minor changes between runs, on a single system, even with eager decode enabled (typically what tempera…
-
comment
Comment #48695652
This is a strange one. We know the hardware capabilities of Cerebras force them to do aggressive REAP pruning to serve Kimi K2.6. Meaning that about 750B parameters is the upper li…
-
comment
Comment #48638862
[dead]
-
comment
Comment #48638673
M5 Ultra will ship before end of year, likely. Though with current RAM shortage, likely max spec will be 256GB and in short supply. In late 2027 or early 2028, Nvidia will release …
-
comment
Comment #48594256
Article reads as though written by someone who doesn't have much experience with deployments like this. Underestimates the memory needed to run with a reasonable amount of context.…
-
comment
Comment #48590700
The Wired headline reframes the issue in a way that’s misleading. SK Telecom was a previously resolved issue (as in prior to Fable launch). It may have been a contributing factor, …
-
comment
Comment #48578878
This headline is not what I would read from this. The numbers are more favorable than the general tone of rumors, and point towards the expected shape of a fast-growing R&D heavy b…
-
comment
Comment #48521149
Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrett…
-
comment
Comment #48520363
This release was rushed to hang on the coattails of the Mythos drama (“hey, sorry you can’t use Fable, but try us while you wait this weekend!”) I think they planned to release nex…
-
comment
Comment #48433934
That happened a year ago when these shipped as the DGX Spark with only Linux pre installed.
-
comment
Comment #48433923
Mostly a strategy move to protect the CUDA moat… Apple would take over mobile inference in a clean sweep without competition.
-
comment
Comment #48433733
This is the same chip and same memory. Only difference is it is going in a laptop, so will be more thermally limited.
-
comment
Comment #48391828
If I were paying API rates this year, I would have already burned through $20k in tokens. Looking forward to the costs of this level of capability coming down.
-
comment
Comment #48391687
I have now also tried it on this scatter plot: https://3215535692-files.gitbook.io/~/files/v0/b/gitbook-x-p... Similarly, the 26B A4B Gemma 4 and the 35B A3B Qwen 3.6 identify it c…
-
comment
Comment #48391593
They haven't made one for this new model, but Unsloth has a comprehensive quant KLD map of Gemma 4 26B A4B here: https://3215535692-files.gitbook.io/~/files/v0/b/gitbook-x-p...
-
comment
Comment #48391488
I want to like the vision capabilities of the model. However, when I gave it an image which Gemma 26B A4B and Qwen 3.6 35B A3B has no problem correctly describing in detail, includ…
-
comment
Comment #48381166
Have you run it through DeepSWE? I understand that's probably a high ask for this class of model, but would be interesting to see regardless. Even if it can't fully pass much, ther…
-
comment
Comment #48379719
While I agree directionally, I'll caveat that "cost per token" != "cost per task". In the case of Qwen3.6 it tends to think 1.6x more than Haiku, so the cost of Haiku on the same t…
-
comment
Comment #48366769
Speaking as someone who has had a DGX Spark all year and been active developing at the driver and kernel level for it and other ARM64 Linux devices the last couple of years, it's n…
-
comment
Comment #48366736
Looks like RTX Spark desktop is the DGX Spark desktop, minus the expensive 200GbE Connect-X NIC. Only since the DGX Spark released, memory and nand prices have jumped, so it will l…
-
comment
Comment #48264082
Claude Opus 4.7 defaults to exactly this design language for a lot of "just make me a rich html presentation page" requests without further specification.