[ExecuTorch][Llama] Split custom sdpa op and kv cache (#7412)
* [ExecuTorch][Llama] Split custom sdpa op and kv cache Summary: This enables us to do more easier module swap with model definitions from torchtune Test Plan: CI Reviewers: Subscribers: Tasks: Tags: [ghstack-poisoned] * Update on "[ExecuTorch][Llama] Split custom sdpa op and kv cache" Summary: This enables us to do more easier module swap with model definitions from torchtune Test Plan: CI Reviewers: Subscribers: Tasks: Tags: [ghstack-poisoned]
K
Kimish Patel committed
af7613c7a5dd39e480aafc1146cd78f55d40bbbb
Parent: 6d78026
Committed by GitHub <noreply@github.com>
on 1/16/2025, 5:59:32 PM