[Executorch][llm] Add ring buffer based kv cache and mask calculation to MHA (#10833)
This PR was created by the merge bot to help merge the original PR into the main branch. ghstack PR number: https://github.com/pytorch/executorch/pull/10609 by @kimishpatel ^ Please use this as the source of truth for the PR details, comments, and reviews ghstack PR base: https://github.com/pytorch/executorch/tree/gh/kimishpatel/186/base ghstack PR head: https://github.com/pytorch/executorch/tree/gh/kimishpatel/186/head Merge bot PR base: https://github.com/pytorch/executorch/tree/gh/kimishpatel/185/orig Merge bot PR head: https://github.com/pytorch/executorch/tree/gh/kimishpatel/186/orig @diff-train-skip-merge Co-authored-by: Kimish Patel <kimishpatel@fb.com>
P
pytorchbot committed
0bb059fa617508c25cabfff6b73ce3a451c27ee8
Parent: f8e7264
Committed by GitHub <noreply@github.com>
on 5/13/2025, 2:29:51 PM