Qualcomm AI Engine Direct - support embedding op (#2057)
Summary: - support embedding op with int32 index input - make mobilebert / llama2 be fully delegated - add requantize passes for mixed precision - bug fixes Pull Request resolved: https://github.com/pytorch/executorch/pull/2057 Reviewed By: dbort Differential Revision: D54348816 Pulled By: cccclai fbshipit-source-id: ec3c8e87cc879d6f642859231255d5094d78349f
H
haowhsu-quic committed
57e192b0391746a26e8598d30f02c96a8a34bb4c
Parent: 75352ad
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com>
on 3/3/2024, 11:37:16 PM