For those who didn't read, this is the identity of the mysterious "Ox Alpha" model
GLM-5.3-Flash
41–50 of 605 posts
Re: GLM-5.3-Flash
#42Re: GLM-5.3-Flash
#43Earlier quoted context omitted.
> get myself four sparks at a decent price Wow, if you don't mind me asking. How and where?
I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.
$4,000 isn't priced insanely? ye gads
Re: GLM-5.3-Flash
#44> 320B total parameters and just 18B active parameters This is pretty hefty for a "flash" model, even a 256 GB setup is insufficient at q4 - and q4 is already the worst-but-still-acceptable quant in my experience. The benchmarks look great, especially since GLM tends to be more honest than the average Chinese lab, but you’ll need to splurge to run it at home. @edit: so many releases that I forgot to math. This fits j…
Looks like the M5 Ultra Studio wait times are going to increase again. Already at 10-12 weeks, I wonder how long it'll go?
Re: GLM-5.3-Flash
#45from the article, pareto frontier for open source models is completely dominated by GLM now.
Re: GLM-5.3-Flash
#46Earlier quoted context omitted.
I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.
> it's the only model in the whole lineup that isn't priced insanely $4,000 isn't priced insanely? ye gads
Re: GLM-5.3-Flash
#47e.g. "Agent Coding Performance by Effort Level" cuts Y-axis from 0~20.
- This makes it as if GLM-5.3-Flash made a bigger jump than it claimed as the Y-axis does not increase much (stupid trick used in biz reports)
I did mention that ox was working ok for me, and having an open-weight comparable to close to SOTA makes it very compelling for me to try it out locally (well, only if I got more VRAM)
Re: GLM-5.3-Flash
#48Earlier quoted context omitted.
> get myself four sparks at a decent price Wow, if you don't mind me asking. How and where?
I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.
Re: GLM-5.3-Flash
#49Re: GLM-5.3-Flash
#50> with all of this traffic served on Chinese AI chips RIP Nivida shareholders
Another self-inflicted own courtesy of US government policy. While I think China would always get to hardware self-sufficiency eventually, all export controls have done is (1) accelerate China's development, and (2) divert revenue that would've otherwise gone to NVIDIA/AMD/etc instead.