add xnnpack backend (#2585)
Summary: Adding support for OSS Dynamically Quantized Linear via XNNPACK for llama models. This now works for cmake on non-mac devices. Additionally, we modified the test-llama-runner-linux work flow and script to test exporting and running via XNNPACK DQ Linear Pull Request resolved: https://github.com/pytorch/executorch/pull/2585 Reviewed By: kirklandsign Differential Revision: D55229489 Pulled By: mcr229 fbshipit-source-id: 26ffbc6d19567ebc0df3c16389451150cc988d08
M
Max Ren committed
725c59055a04c335dc07114b578ae543aed83c7a
Parent: 14e31f0
Committed by Facebook GitHub Bot <facebook-github-bot@users.noreply.github.com>
on 3/23/2024, 1:30:14 AM