COMMITS
/ examples/models/llama/runner/generation.py March 29, 2025
M
Remove old tokenizer/ directory in ExecuTorch
Mengwei Liu committed
March 3, 2025
J
Basic performance logging for Llama python runner (#8862)
Jack committed
February 13, 2025
J
Python hugging face tokenizer (#8354)
Jack committed
December 3, 2024
P
Update eager runner to support AttentionSink (#7149)
pytorchbot committed
November 18, 2024
P
Fix Cuda out of memory issue for eager runner (#6935)
pytorchbot committed
November 15, 2024
J
Make TorchTune Llama model KV cache compatible in eager (#6643)
Jack Zhang committed
November 14, 2024
J
Runner changes for TorchTune Llama3.2 vision text decoder (#6610)
Jack Zhang committed
November 12, 2024
P
Print the number of tokens generated (#6773)
pytorchbot committed
November 11, 2024
P
add the ability to have multi-round conversation with llama (#6769)
pytorchbot committed
P
update llama runner to decode single token (#6768)
pytorchbot committed
October 28, 2024
L
add the ability to run eager runner via buck
Lunwen He committed
October 22, 2024
H
fix eager run for cuda (#6429)
Hansong committed
October 18, 2024
L
fix llama eager runner and add ci (#6344)
Lunwen He committed
October 16, 2024
M
Codemod examples/models/llama2 to examples/models/llama (#6302)
Mergen Nachin committed