Qualcomm AI Engine Direct - Refine max spill fill buffer setting (#5989)
Summary: - Get required spillFillBufferSize from context binary and set to compiler_spec - Quantize embedding op in qnn. - If enable multi-contexts, maxSpillFillBuffer could not set to zero. Pull Request resolved: https://github.com/pytorch/executorch/pull/5989 Reviewed By: kirklandsign Differential Revision: D64056107 Pulled By: cccclai fbshipit-source-id: 9f9846e6ac7b4a27d734d2812ac3bbad32fb194f
S
Sheng Feng Wu committed
01fcdf420fef23b4ee0348c37abcab74bcea1449
Parent: 0d1250a
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com>
on 10/9/2024, 1:13:36 AM