dynamic quantize linear and lower to xnnpack (#1849)
Summary: Pull Request resolved: https://github.com/pytorch/executorch/pull/1849 Add a path to dynamic quantize linear layers along with embedding. Will factor our embedding quant in a separate diff. ghstack-source-id: 214288401 exported-using-ghexport Reviewed By: digantdesai Differential Revision: D53316360 fbshipit-source-id: 39d6a25d9da44ba84995c5d54a4a132da262b35d
K
Kimish Patel committed
b026e39afc356ca81d0e31f39889ca6ac381f0c7
Parent: 8e531f8
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com>
on 2/6/2024, 10:21:02 PM