SIGN IN SIGN UP

Qualcomm AI Engine Direct - Enable 4 bits BW quantization (#2506)

Summary:
- Add QNN_QUANTIZATION_ENCODING_BW... confings for qnn wrapper
- Add 4 bits quant config
- Add 4 bits quant single op tests
- Add per channel weight setting for quantizer
- Fix convert_to_linear error
- Refine quantizer

Pull Request resolved: https://github.com/pytorch/executorch/pull/2506

Reviewed By: kirklandsign

Differential Revision: D55142752

Pulled By: cccclai

fbshipit-source-id: f6ff6c753551450531e1389c50320eca42d69008
C
chunit-quic committed
3270d22e44ff89022774a91137aea4910d025c2b
Parent: 5a93353
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com> on 3/21/2024, 1:39:59 AM