Anthropic's reply? A model you can't use.
GLM-5.1: Towards Long-Horizon Tasks
111–120 of 285 posts
Re: GLM-5.1: Towards Long-Horizon Tasks
#112"I am the storm that is approaching, provoking..." : )
Re: GLM-5.1: Towards Long-Horizon Tasks
#113One of the bench maxed models . Every time I tried it , it’s not on par even with other open source models .
Being "better than Opus 4.6" is not really something a benchmark will tell you. It's much more a consensus of users liking the flavor of an answer, rather than fueling x% correct on a benchmark.
Re: GLM-5.1: Towards Long-Horizon Tasks
#114Not only did this one draw me an excellent pelican... it also animated it! https://simonwillison.net/2026/Apr/7/glm-51/
Re: GLM-5.1: Towards Long-Horizon Tasks
#115Not only did this one draw me an excellent pelican... it also animated it! https://simonwillison.net/2026/Apr/7/glm-51/
Simon, you need to come up with improved benchmarks soon.
Re: GLM-5.1: Towards Long-Horizon Tasks
#116Earlier quoted context omitted.
Three hour coffee break while the LLM prepares scaffolding for the project.
Like computing used to be. When I first compiled a Linux kernel it ran overnight on a Pentium-S. I had little idea what I was doing, probably compiled all the modules by mistake.
Re: GLM-5.1: Towards Long-Horizon Tasks
#117I am on their "Coding Lite" plan, which I got a lot of use out of for a few months, but it has been seriously gimped now. Obvious quantization issues, going in circles, flipping from X to !X, injecting chinese characters. It is useless now for any serious coding work.
Is there any advantage to their fixed payment plans at all vs just using this model via 3rd party providers via openrouter, given how relatively cheap they tend to be on a per-token basis? Providers like DeepInfra are already giving access to 5.1 https://deepinfra.com/zai-org/GLM-5.1 $1.40 in $4.40 out $0.26 cached / 1M tokens That's more expensive than other models, but not terrible, and will go down over time, and…
Re: GLM-5.1: Towards Long-Horizon Tasks
#118 (1) OpenAI & Anthropic are absolutely cooked; it's obvious they have no moat
(2) Local/private inference is the future of AI
(3) There's *still* no killer product yet (so get to work!)Re: GLM-5.1: Towards Long-Horizon Tasks
#119Every single day, three things are becoming more and more clear: (1) OpenAI & Anthropic are absolutely cooked; it's obvious they have no moat (2) Local/private inference is the future of AI (3) There's *still* no killer product yet (so get to work!)
Re: GLM-5.1: Towards Long-Horizon Tasks
#120Every single day, three things are becoming more and more clear: (1) OpenAI & Anthropic are absolutely cooked; it's obvious they have no moat (2) Local/private inference is the future of AI (3) There's *still* no killer product yet (so get to work!)