Add Vulkan Quantizer to Llama export lib (#6169)
Summary: Pull Request resolved: https://github.com/pytorch/executorch/pull/6169 TSIA. Note that only 8 bit weight only quantization is supported for now since `VulkanQuantizer` does not support 4 bit weight only quantization at the moment. ghstack-source-id: 247613963 exported-using-ghexport Reviewed By: jorgep31415 Differential Revision: D64249615 fbshipit-source-id: 33ef8d06e56838da5f7866832e50fe74d8878811
S
Stephen Jia committed
4b3ffc4ae5f7a86f34e26c005086889032b40d15
Parent: 236e60d
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com>
on 10/11/2024, 11:24:25 PM