COMMITS
/ examples/models/llama/TARGETS April 10, 2025
P
[Executorch][llama] Enable quantized sdpa (#10062)
pytorchbot committed
P
[Executorch][llama] Renamed quantized_kv_cache to custom_kv_cache (#10061)
pytorchbot committed
March 29, 2025
M
Remove old tokenizer/ directory in ExecuTorch
Mengwei Liu committed
March 26, 2025
J
Add buck target for hf_download
Jack committed
March 6, 2025
M
Add qk norm optionally before attention calculation
madhu-fb committed
February 23, 2025
Y
Update visibility of target examples/models/llama:source_transformation
YIWENX14 committed
February 21, 2025
C
fix export llama to qnn
cccclai committed
February 20, 2025
C
Refactor source_transformation to a seperate target
cccclai committed
February 7, 2025
S
Static attention implementation
Shen Chen Xu committed
February 4, 2025
J
Buckify Llama multimodal export (#7604)
Jack committed
February 1, 2025
M
Add abstract base class for attention mechanisms with unified interface
Mengtao Yuan committed
January 27, 2025
P
[ez] Add `delegation_info` dep to buck target
pytorchbot committed
January 14, 2025
J
Fix kv cache pyre and build
Jack Zhang committed
December 3, 2024
P
Update eager runner to support AttentionSink (#7149)
pytorchbot committed
November 27, 2024
P
implement position encoding for shifted tokens
pytorchbot committed
November 11, 2024
C
Qualcomm AI Engine Direct - Add llama sha transforming pass
Chun-I Tsai committed
October 28, 2024
L
add the ability to run eager runner via buck
Lunwen He committed
October 22, 2024
N
llama export with input vocab pruning
Naveen Suda committed
October 21, 2024
H
[ET-VK] Enable custom rotary embedding module replacement (#6424)
Hansong committed
October 16, 2024
M
Codemod examples/models/llama2 to examples/models/llama (#6302)
Mergen Nachin committed