GLM-5.1: Towards Long-Horizon Tasks
1–10 of 285 posts
Re: GLM-5.1: Towards Long-Horizon Tasks
#2[stub for offtopicness] [[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Re: GLM-5.1: Towards Long-Horizon Tasks
#3[stub for offtopicness] [[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Re: GLM-5.1: Towards Long-Horizon Tasks
#4[stub for offtopicness] [[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Re: GLM-5.1: Towards Long-Horizon Tasks
#5[[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Re: GLM-5.1: Towards Long-Horizon Tasks
#6[stub for offtopicness] [[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Interesting.
Hopefully these aren't bots created by Z.AI because GLM doesn't need fake engagement.
Re: GLM-5.1: Towards Long-Horizon Tasks
#7[stub for offtopicness] [[you guys, please don't post like this to HN - it will just irritate the community and get you flamed]]
Hmm, three spam comments posted within 9 minutes of each other. The accounts were created 15 minutes ago, 51 days ago, and 3 months ago. Interesting. Hopefully these aren't bots created by Z.AI because GLM doesn't need fake engagement.
Thanks for watching out for the quality of HN...
Re: GLM-5.1: Towards Long-Horizon Tasks
#8For short-term bugfixing and tweaks though, it does about what I'd expect from Sonnet for a pretty low price.
Re: GLM-5.1: Towards Long-Horizon Tasks
#9It's an okay model. My biggest issue using GLM 5.1 in OpenCode is that it loses coherency over longer contexts. When you crest 128k tokens, there's a high chance that the model will start spouting gibberish until you compact the history. For short-term bugfixing and tweaks though, it does about what I'd expect from Sonnet for a pretty low price.
Since the entire purpose, focus and motivation of this model seems to have been "coherency over longer contexts", doesn't that issue makes it not an OK model? It's bad at the thing it's supposed to be good at, no?