COMMITS
/ benchmarks/float8/bench_matmul.py August 12, 2025
V
fix float8 training benchmarks on AMD (#2737)
Vasiliy Kuznetsov committed
July 3, 2025
D
[moe training] add benchmarking script for grouped mm (#2490)
Daniel Vega-Myhre committed
June 24, 2025
D
add-to-benchmarks (#2427)
Driss Guessous committed
V
rename `torchao.testing.float8` to `torchao.testing.training` (#2415)
Vasiliy Kuznetsov committed
J
Revert "Build mxfp4 kernel for sm120a" (#2428)
Jerry Zhang committed
June 21, 2025
T
Build mxfp4 kernel for sm120a (#2285)
Thien Tran committed
March 5, 2025
V
float8 matmul benchmark: hook up cublas mxfp8 gemm (#1830)
Vasiliy Kuznetsov committed
V
float8 matmul bench: make peak tops lookup dynamic (#1829)
Vasiliy Kuznetsov committed
January 8, 2025
A
Lint benchmark folder (#1519)
Apurva Jain committed
October 7, 2024
V
add axiswise scaling to Float8Linear (#920)
Vasiliy Kuznetsov committed
September 10, 2024
V
add CI to disallow syntax errors and undefined vars in all Python files (#861)
Vasiliy Kuznetsov committed
September 4, 2024
V
make f8 roofline script calculate observed overhead (#734)
Vasiliy Kuznetsov committed
August 13, 2024
V
float8 gemm benchmarks: add option for gpu time (#666)
Vasiliy Kuznetsov committed
August 7, 2024
V
update float8 benchmarks to be more useful for smaller shapes (#615)
Vasiliy Kuznetsov committed
August 5, 2024
V
QOL improvements to float8 gemm benchmark (#596)
Vasiliy Kuznetsov committed
July 30, 2024
V
move float8_experimental to torchao/float8
Vasiliy Kuznetsov committed