Intel reduces latencies of chat LLM app using quantisation
community.intel.com
Intel reduces latencies of chat LLM app using quantisation
1–6 of 6 posts
Re: Intel reduces latencies of chat LLM app using quantisation
#2Only me or does the link also refer to general blog index for others?
I can't see the post.
Re: Intel reduces latencies of chat LLM app using quantisation
#3Bad link? This just takes me to an article index listing
Re: Intel reduces latencies of chat LLM app using quantisation
#4Here's (probably) the correct link: https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/T...
Re: Intel reduces latencies of chat LLM app using quantisation
#5Sorry, this is the link: https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/T...
Re: Intel reduces latencies of chat LLM app using quantisation
#6Sorry, this is the link: https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/T...
Our software changes submitted links to canonical URLs when it finds them, and that page has the canonical URL https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/b....
I've fixed it above now.