COMMITS
/ README.md August 4, 2025
A
New multi-step QAT API (#2629)
andrewor14 committed
August 1, 2025
A
[bc-breaking] Generalize FakeQuantizeConfig beyond intx (#2628)
andrewor14 committed
July 30, 2025
Z
[Easy] Fix git repo url in citation (#2599)
Zesheng Zong committed
July 17, 2025
A
Update paper link readme (#2563)
andrewor14 committed
A
Update README with PEFT integration + installation (#2559)
andrewor14 committed
June 28, 2025
S
Update README.md to include Flux-Fast (#2457)
Sayak Paul committed
June 25, 2025
A
Call out axolotl + QAT integration on README (#2442)
andrewor14 committed
June 16, 2025
A
Revamp README (#2374)
andrewor14 committed
June 11, 2025
S
Update README.md to include seamless v2 (#2355)
Sayak Paul committed
June 10, 2025
S
Update QAT docs, highlight axolotl integration (#2266)
salman committed
May 22, 2025
D
Update Readme (#1526)
Driss Guessous committed
April 12, 2025
D
[BE] Remove hf_eval.py and add documentation on using lm-eval (#2045)
Daniel Vega-Myhre committed
April 1, 2025
V
refresh float8 training section of main readme (#1985)
Vasiliy Kuznetsov committed
March 10, 2025
M
Promote Low Bit Optim out of prototype (#1864)
Mark Saroufim committed
March 5, 2025
A
Revert "Move torchao/_models to benchmarks/_models" (#1844)
Apurva Jain committed
March 4, 2025
A
Move torchao/_models to benchmarks/_models (#1784)
Apurva Jain committed
February 28, 2025
H
Updating Cuda 12.1/12.4 to 12.4/12.6 to reflect current state (#1794)
HDCharles committed
February 14, 2025
V
update torchao READMEs with new configuration APIs (#1711)
Vasiliy Kuznetsov committed
January 13, 2025
A
Update QAT READMEs using new APIs (#1541)
andrewor14 committed
November 30, 2024
L
Update README.md: Fix bibtex and sglang links (#1361)
Lianmin Zheng committed
November 26, 2024
V
Fixed invalid url in citation section (#1348)
Vladimir Malinovskii committed
November 20, 2024
D
Update README.md (#1319)
Driss Guessous committed
October 26, 2024
M
Add bibtex (#1177)
Mark Saroufim committed
October 24, 2024
A
[reland] Move QAT out of prototype (#1152)
andrewor14 committed
October 4, 2024
M
Add more videos (#1011)
Mark Saroufim committed
September 24, 2024
M
rename cuda mode -> gpu mode (#925)
Mark Saroufim committed
September 12, 2024
J
Renaming fpx to floatx (#877)
Jerry Zhang committed
September 9, 2024
J
Update README.md (#849)
Jerry Zhang committed
September 8, 2024
M
Update README (#823)
Mark Saroufim committed
September 6, 2024
J
Update docs + add deprecation warning (#825)
Jesse Cai committed
September 5, 2024
V
Update main README.md with more current float8 speedup (#816)
Vasiliy Kuznetsov committed
H
adding kv cache quantization to READMEs (#813)
HDCharles committed
H
int4 fixes and improvements (#804)
HDCharles committed
August 28, 2024
M
Revert "more empathy fixes" (#768)
Mark Saroufim committed
M
more empathy fixes (#759)
Mark Saroufim committed
August 27, 2024
M
empathy day fixes (#757)
Mark Saroufim committed
August 25, 2024
M
README typos (#747)
Mark Saroufim committed
M
1 more doc revamp (#745)
Mark Saroufim committed
August 23, 2024
A
update README example with correct import of `sparsify_` (#741)
Aryan committed
August 21, 2024
M
torchao-nightly -> --pre torchao in README (#723)
Mark Saroufim committed
August 13, 2024
J
Move Uintx out of prototype for future extension (#635)
Jerry Zhang committed
August 2, 2024
J
move sam eval from `scripts` to `torchao/_models` (#591)
Jesse Cai committed
July 31, 2024
J
Update README.md (#583)
Jerry Zhang committed
July 30, 2024
V
move float8_experimental to torchao/float8
Vasiliy Kuznetsov committed
July 26, 2024
J
Implement sparsity as a AQT Layout (#498)
Jesse Cai committed
July 9, 2024
M
Upgrade pytorch version to 3.9 (#488)
Mark Saroufim committed
July 5, 2024
J
Add sparsify API to torchao (#473)
Jesse Cai committed
July 4, 2024
J
Renaming `quantize` to `quantize_` (#467)
Jerry Zhang committed
July 2, 2024
J
Add segment-anything-fast perf/acc benchmarks to torchao (#457)
Jesse Cai committed
June 28, 2024
M
Add Blogs and Videos section (#460)
Mark Saroufim committed