[Executorch][llama] Enable quantized sdpa (#10062)
Enable leveraging quantized sdpa op when quantized kv cache is used. Instead of adding yet another arg, at the moment I have chosen to leverage quantize_kv_cache option. Differential Revision: [D71833064](https://our.internmc.facebook.com/intern/diff/D71833064/)
P
pytorchbot committed
40beadec800dd618ca7c9ebb248fb01b096f5ca6
Parent: 8d6aa35
Committed by GitHub <noreply@github.com>
on 4/10/2025, 5:33:41 PM