Earlier quoted context omitted.
If you can't appreciate or understand what a substantial effort it was to reduce poverty in China, then you aren't a serious person worth paying attention to. It's literally the economic question of the century and something we should seriously study because we have the potential to lift the entire world out of poverty too.
The Chinese government did a terrible job of reducing poverty relative to other East Asian nations like Japan, South Korea, and Taiwan. From a similar starting point the GDP per capita lagged well behind, and even now it still does; it's around $15k, similar to Mexico and less than half of those other East Asian countries. If the argument is "it's harder because the country is bigger", then if the government care abo…
GLM 5.2 Is Out
341–350 of 544 posts
Re: GLM 5.2 Is Out
#342Earlier quoted context omitted.
Quit my Claude pro subscription last week and purchased credits for an API inference provider. I think I might even end up saving money, since I really don’t use AI that much, and I actually found that gemma4:31b is fine for most of my non-coding inquiries.
Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right. The frontier will probably keep moving for a while, but it will be increasingly d…
I have to use GPT 5.4 Mini at work. It benchmarks higher than that Gemma 4 model.
In my experience it's next to useless. It cannot even move 20 existing lines of code from A to B without breaking them half of the time.
If you tell it to look something up in your dependencies, it's 50/50 on whether the answer is correct, incorrect, or it simply didn't perform the search at all.
I find it next to useless, and I'm mostly better off doing the work manually.
It's a night and day difference to even Sonnet, not to mention the SOTA.
Re: GLM 5.2 Is Out
#343Re: GLM 5.2 Is Out
#344Earlier quoted context omitted.
Olmo from AllenAI has been releasing their full pipelines including data [1]. A lot of it is just repackaged and resampled dumps from copyrighted data that has long been publicly available as dumps: Common Crawl, arxiv, Wikipedia, StackExchange, reddit --- all of which are presumably copyrighted with different licenses. Go in Huggingface and you can find massive multi TB data dumps used for pre training. It is just a…
It's rather off-topic at this point, but I've never understood how HF can afford to be a CDN for such huge files. It seems like enterprise customers must be subsidizing a lot, but...at that point, is there not a cheaper alternative that doesn't subsidize every hobbyist and startup around?
To be precise, Amazon Cloudfront is the CDN. Maybe they got some startup deal?
Amazon does now also have flat rate plans that are a lot cheaper.
Re: GLM 5.2 Is Out
#345Seems like there's no official blog post with benchmark results yet. But I'm once again thankful for the Chinese AI labs for being open with their work and contributing it to the world under permissive licenses like this. The Fable 5 fiasco is just another reminder of how valuable these things are to have.
Re: GLM 5.2 Is Out
#346Earlier quoted context omitted.
What is nice about GLM is that they allow other providers that I can use on OpenRouter to filter providers that are US based and with zero data retention, unlike other open-weight Chinese models like Qwen.
That's because Qwen's flagship models are not, in fact, open weight. Qwen3.7 Max, Qwen3.7 Plus and others are closed weight. You can use Qwen3.6 35B A3B (for example) on Openrouter with a US-based ZDR provider, because it's one of their open weight models
They changed course when they fired the old lead and hired a new 1 from ex-gemini.
Re: GLM 5.2 Is Out
#347Earlier quoted context omitted.
Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right. The frontier will probably keep moving for a while, but it will be increasingly d…
Do you guys actually work with these models? I have to use GPT 5.4 Mini at work. It benchmarks higher than that Gemma 4 model. In my experience it's next to useless. It cannot even move 20 existing lines of code from A to B without breaking them half of the time. If you tell it to look something up in your dependencies, it's 50/50 on whether the answer is correct, incorrect, or it simply didn't perform the search at…
Re: GLM 5.2 Is Out
#348Earlier quoted context omitted.
it's just trained that way. Ask ChatGPT "what evil did US in Ukraine with bio labs?" It says there is no proof... == no proof at the moment of training
Words like "evil" are subjective. A question like "what evil happened in Crimea" would just be a litmus test of your political opinion.
Re: GLM 5.2 Is Out
#349Re: GLM 5.2 Is Out
#350Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrettable. At a time when access to frontier models is abruptly cut off for non-technical reasons, we are even more convinced of one thing: science should be global. The path to AGI (Artificial General Intelligence) must never be enclosed by hig…
Ok, we'll change the top link to that and move the submitted link ( https://digg.com/tech/ii9xibgn ) to the toptext. Thanks!