Viewing profile — bertili
bertili
HN member- Joined
- Thu, Aug 15, 2024, 6:03 PM UTC
- HN karma
- 264
- Public activity
- 54 items
- HN profile
- View on Hacker News ↗
About bertili
No profile information was provided.
Recent public activity
-
comment
Comment #49153404
The 27B have many more active parameters than much bigger models such as DS4Flash, MiniMax etc, which makes it punch above its tiny weight. A great fit for a 5090 in a closet for m…
-
comment
Comment #49067721
That looks promising! As models become a commodity, this may turn out to be the real AI gold rush.
-
comment
Comment #49067628
Is there any (near future) technology that would permit burning this terrabyte into some kind of ROM chip?
-
comment
Comment #48971421
Wait.. the Qwen Max models have never been open-weight. But it sure sound like that's what they intend now? "Qwen3.8 is launching and going open-weight soon! With a massive 2.4T pa…
-
comment
Comment #48923780
AGI is almost here, but first, one more thing... a keyboard controller!
-
comment
Comment #48903259
Legislators, please require 10 seconds of load screen with a picture of a tree, for every online video. It worked for cigarette packs.
-
comment
Comment #48876493
[dead]
-
comment
Comment #48568378
This is GLM 5.2 Max. GLM 5.2 High which use less than half[1] the tokens. [1] https://z.ai/blog/glm-5.2
-
comment
Comment #48518980
Qwen 27b is a compute heavy dense model.
-
comment
Comment #48157858
Does this translate into a similar reduction in compute? What's the catch?
-
comment
Comment #48052594
equals 2 or 3 human brains in power usage. Amazing work!
-
comment
Comment #47797059
It's fascinating that a $999 Mac Mini (M4 32GB) with almost similar wattage as a human brain gets us this far.
-
comment
Comment #47793586
Is there any source for these claims?
-
comment
Comment #47793082
A relief to see the Qwen team still publishing open weights, after the kneecapping [1] and departures of Junyang Lin and others [2]! [1] https://news.ycombinator.com/item?id=472467…
-
comment
Comment #47617588
The timing is interesting as Apple supposedly will distill google models in the upcoming Siri update [1]. So maybe Gemma is a lower bound on what we can expect baked into iPhones. …
-
comment
Comment #47616892
Qwen: Hold my beer https://news.ycombinator.com/item?id=47615002
-
comment
Comment #47476848
Very impressive! I wonder if there is a similar path for Linux using system memory instead of SSD? Hell, maybe even a case for the return of some kind of ROMs of weights?
-
comment
Comment #47035295
Better than frontier pelicans as of 2025
-
comment
Comment #47034495
Most certainly not, but the Unsloth MLX fits 256GB.
-
comment
Comment #47034257
Last Chinese new year we would not have predicted a Sonnet 4.5 level model that runs local and fast on a 2026 M5 Max MacBook Pro, but it's now a real possibility.
-
comment
Comment #46980995
Exactly. The emperor has no clothes. The largest investments in US tech in history and yet there less than a year of moat. OpenAI or Anthropic will not be able to compete with Chin…
-
comment
Comment #46898512
Surely this is the elephant in the room, but the point here is that Apple as control over its ecosystem, so it may be able to sandbox and make entitlements and transparency good en…
-
comment
Comment #46778028
People are running the previous Kimi K2 on 2 Mac Studios at 21tokens/s or 4 Macs at 30tokens/s. Its still premature, but not a completely crazy proposition for the near future, giv…
-
comment
Comment #46777833
The other realistic setup is $20k, for a small company that needs a private AI for coding or other internal agentic use with two Mac Studios connected over thunderbolt 5 RMDA.
-
comment
Comment #46777339
The "Deepseek moment" is just one year ago today! Coincidence or not, let's just marvel for a second over this amount of magic/technology that's being given away for free... and ho…