COMMITS
/ examples/models/llama2/README.md April 9, 2025
M
[doc] Fix tokenizer related documentation (#10000)
Mengwei Liu committed
October 18, 2024
M
Refactor out llama2 specific content out of Llama readme (#6359)
Mergen Nachin committed
October 16, 2024
M
Codemod examples/models/llama2 to examples/models/llama (#6302)
Mergen Nachin committed
October 15, 2024
H
Android NDK use r27b (#6092)
Hansong Zhang committed
October 12, 2024
L
add instructions about getting mmlu score for instruct models (#6175)
Lunwen He committed
October 9, 2024
H
Remove logic for appending or prepending tokens (#4920)
helunwencser committed
L
Add clarification about perpelexity discrepancy (#6053)
Lunwen He committed
October 8, 2024
J
Update Llama README.md for Stories110M tokenizer (#5960)
Jack Zhang committed
October 7, 2024
L
use --use_sdpa_with_kv_cache for 1B/3B bf16 (#5861)
Lunwen He committed
October 2, 2024
L
September 30, 2024
Y
Polish CoreML Llama Doc (#5745)
yifan_shen3 committed
M
Add animated gif for 3B SpinQuant (#5763)
Mergen Nachin committed
September 27, 2024
M
Improve llama README with SPinQuant
Mergen Nachin committed
September 26, 2024
L
add performance number for 1B/3B (#5704)
Lunwen He committed
M
IMprove README page
Mergen Nachin committed
L
add instruction for quantizing with SpinQuant (#5672)
Lunwen He committed
September 25, 2024
M
Add animated gif for Llama3.2 1B bf16 (#5671)
Mergen Nachin committed
M
Add llama3.2 1B and 3B instructions (#5647)
Mergen Nachin committed
M
Improve Llama page (#5639)
Mergen Nachin committed
September 18, 2024
M
Add llama animated gif to llama readme (#5474)
Mergen Nachin committed
L
Update stories cmd to use kv cache (#5460)
lucylq committed
September 17, 2024
M
Add SpinQuant into README (#5412)
Mergen Nachin committed
September 5, 2024
A
Switch to the new tensor API internally.
Anthony Shoumikhin committed
August 30, 2024
M
[llama] Build the runner with tiktoken by default
Mengwei Liu committed
August 28, 2024
C
update doc to disable dynamic shape (#4941)
cccclai committed
August 21, 2024
L
Allow multiple eos ids
Lunwen He committed
August 9, 2024
M
Update on Mac runner build
Mengtao Yuan committed
July 24, 2024
L
Add llama3.1 to readme (#4378)
lucylq committed
July 23, 2024
L
Update llama docs on main (#4361)
lucylq committed
July 15, 2024
L
Move tokenizer.py into extension/llm/tokenizer (#4255)
Lunwen He committed
July 2, 2024
K
Update read me with llama3 numbers
Kimish Patel committed
June 6, 2024
S
update <tokenizer.model> in llama2 readme (#3879)
Songhao Jia committed
June 5, 2024
L
Update llama readme, use main branch for llama3 (#3861)
Lucy Qiu committed
May 31, 2024
M
Update README.md to ask user to double check python env (#3782)
Mengwei Liu committed
May 21, 2024
J
update Llama 2 README to include more Llama 3 info (#3582)
Jeff Tang committed
May 15, 2024
S
add missing cmake flags to build llama runner for android (#3611)
salykova committed
May 13, 2024
M
Update README.md with correct link (#3591)
Mergen Nachin committed
May 8, 2024
M
Add instructions to convert Hugging Face models to PyTorch (#3523)
Mengtao Yuan committed
May 3, 2024
A
Add suffixes to cmake flags related to building kernels. (#3499)
Anthony Shoumikhin committed
April 29, 2024
K
Update llama3 8b enablement to include s24
Kimish Patel committed
April 25, 2024
M
Update llama2 readme file - main branch (#3340)
Mergen Nachin committed
April 24, 2024
L
llama2 readme (#3315)
Lucy Qiu committed
April 19, 2024
C
Docs for lower smaller models to mps/coreml/qnn (#3146)
Chen Lai committed
M
Instructions for Llama3 (#3154)
Mergen Nachin committed
M
Update Llama3 perplexity numbers in README.md (#3145)
Mengtao Yuan committed
M
Update README.md on the evaluation parameters (#3139)
Mengtao Yuan committed
April 18, 2024
M
Update README.md for llama3 (#3141)
Mengtao Yuan committed
M
Adding Gotchas in README.md (#3138)
Mergen Nachin committed
D
Update README.md (#3094)
Digant Desai committed