Excuse my ignorance. Could one just say, "One expert is all I can handle" and strip the others from the model?
Show HN: Getting GLM 5.2 running on my slow computer
121–130 of 269 posts
Re: Show HN: Getting GLM 5.2 running on my slow computer
#122Earlier quoted context omitted.
The funny thing is Claude Cowork has taught me to be patient with response timelines. I’m now figuring I’ll be running locally no later than 2028. (I want to spend no more than $10k. And I want to run a model comparable to today’s SOTA.)
Today's SOTA also sounds totally sufficient to me, but I wonder how much our standards will inflate by 2028. Maybe a lot, maybe not at all...very hard to say.
Expectations seem to be rising at a faster rate than models can improve.
Re: Show HN: Getting GLM 5.2 running on my slow computer
#123Re: Show HN: Getting GLM 5.2 running on my slow computer
#124Earlier quoted context omitted.
The funny thing is Claude Cowork has taught me to be patient with response timelines. I’m now figuring I’ll be running locally no later than 2028. (I want to spend no more than $10k. And I want to run a model comparable to today’s SOTA.)
For 10k you can buy a used dual socket Intel or amd based rackmount server with a terabyte of ram, and run models on cpu only at a reasonable speed. Same server would have been 4-5k a couple years ago before ram price rise. Or buy one on eBay with 512GB that has half its slots populated and then buy the matching 512GB kit to add.
Cost/Value when compared to cloud services is just not there, but I see the merit for those who value privacy over quality of output and want a backup of huge condensed corpus of data within their control.
Kudos to OP though, They had clear goals and they achieved it.
Re: Show HN: Getting GLM 5.2 running on my slow computer
#125Earlier quoted context omitted.
> on hardware that ordinary people can afford These days, can "ordinary people" afford 24GB of ram and half a TB of NVME ssd? sigh
The very boring pair of two 16GB ddr5 6000 I had in my newegg shopping cart went from $399 to $475, so increasingly the answer will be "no".
Re: Show HN: Getting GLM 5.2 running on my slow computer
#126I just learned about Gemma4.pas at the beginning of this week. Now this. This make me wonder how can inference engines could be built that easy. I'm not knowledgeable in this, but I thought it would take very deep Mathematic and system level knowledge, ... and a lot of patience.
Re: Show HN: Getting GLM 5.2 running on my slow computer
#127Re: Show HN: Getting GLM 5.2 running on my slow computer
#128Earlier quoted context omitted.
For 10k you can buy a used dual socket Intel or amd based rackmount server with a terabyte of ram, and run models on cpu only at a reasonable speed. Same server would have been 4-5k a couple years ago before ram price rise. Or buy one on eBay with 512GB that has half its slots populated and then buy the matching 512GB kit to add.
Which CPU gen are you suggesting, is there any writeup on such setup where In my experience with rig half that cost, entire exercise of running coding models locally has been a huge disappointment. Cost/Value when compared to cloud services is just not there, but I see the merit for those who value privacy over quality of output and want a backup of huge condensed corpus of data within their control. Kudos to OP thou…
Or people who want or need to run an uncensored (abliterated) gguf file to deal with controversial topics that a paid LLM service will refuse to work with or ban you for.
Re: Show HN: Getting GLM 5.2 running on my slow computer
#129Earlier quoted context omitted.
The very boring pair of two 16GB ddr5 6000 I had in my newegg shopping cart went from $399 to $475, so increasingly the answer will be "no".
Does it have to be DDR5? Is the limit RAM speed, or SSD speed?
Re: Show HN: Getting GLM 5.2 running on my slow computer
#130Excuse my ignorance. Could one just say, "One expert is all I can handle" and strip the others from the model?