3 ms·
Intel reduces latencies of chat LLM app using quantisation
- thenaturalist 2y agoOnly me or does the link also refer to general blog index for others? I can't see the post.
- Havoc 2y agoBad link? This just takes me to an article index listing
- rrr_oh_man 2y agoHere's (probably) the correct link: https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/Tuning-and-Inference-for-Generative-AI-with-4th-Generation-Intel/post/1554777 https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/T...
- mariarmestre 2y agoSorry, this is the link: https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/Transforming-Customer-Service-How-an-Intel-Customer-Built-a/post/1595710 https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/T...
- dang 2y agoOur software changes submitted links to canonical URLs when it finds them, and that page has the canonical URL https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/bg-p/blog-cloud/spr-xeon-gen-ai-ml-part3.html https://community.intel.com/t5/Blogs/Tech-Innovation/Cloud/b.... I've fixed it above now.