SIGN IN SIGN UP

Add Vulkan Quantizer to Llama export lib (#6169)

Summary:
Pull Request resolved: https://github.com/pytorch/executorch/pull/6169

TSIA.

Note that only 8 bit weight only quantization is supported for now since `VulkanQuantizer` does not support 4 bit weight only quantization at the moment.
ghstack-source-id: 247613963
exported-using-ghexport

Reviewed By: jorgep31415

Differential Revision: D64249615

fbshipit-source-id: 33ef8d06e56838da5f7866832e50fe74d8878811
S
Stephen Jia committed
4b3ffc4ae5f7a86f34e26c005086889032b40d15
Parent: 236e60d
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com> on 10/11/2024, 11:24:25 PM