Re-quantizing a local LLM 14x faster by skipping the tensors that didn't change
andreaborio.substack.com
Re-quantizing a local LLM 14x faster by skipping the tensors that didn't change
1–2 of 2 posts
Re: Re-quantizing a local LLM 14x faster by skipping the tensors that didn't change
#2[dead]