SIGN IN SIGN UP

Qualcomm AI Engine Direct - Refine max spill fill buffer setting (#5989)

Summary:
- Get required spillFillBufferSize from context binary and set to compiler_spec
- Quantize embedding op in qnn.
- If enable multi-contexts, maxSpillFillBuffer could not set to zero.

Pull Request resolved: https://github.com/pytorch/executorch/pull/5989

Reviewed By: kirklandsign

Differential Revision: D64056107

Pulled By: cccclai

fbshipit-source-id: 9f9846e6ac7b4a27d734d2812ac3bbad32fb194f
S
Sheng Feng Wu committed
01fcdf420fef23b4ee0348c37abcab74bcea1449
Parent: 0d1250a
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com> on 10/9/2024, 1:13:36 AM