COMMITS
/ examples/models/llama/source_transformation/sdpa.py May 13, 2025
P
April 21, 2025
April 10, 2025
P
[Executorch][llama] Enable quantized sdpa (#10062)
pytorchbot committed
February 1, 2025
M
Add abstract base class for attention mechanisms with unified interface
Mengtao Yuan committed
January 30, 2025
P
January 23, 2025
K
[ExecuTorch][BE] Split kv cache and SDPA for better code sharing
Kimish Patel committed
January 16, 2025
K
[ExecuTorch][Llama] Split custom sdpa op and kv cache (#7412)
Kimish Patel committed
December 7, 2024
P
[Executorch] Add quantized kv cache to oss ci (#7212)
pytorchbot committed
December 6, 2024
P
[Executorch][BE] Rename sdpa_with_kv_cache.py to custom_ops.py (#7210)
pytorchbot committed
October 16, 2024
M
Codemod examples/models/llama2 to examples/models/llama (#6302)
Mergen Nachin committed