SIGN IN SIGN UP

[executorch][cuda] Optimize short-query INT6 matvec kernels (#21504)

This PR was created by the merge bot to help merge the original PR into
the main branch.
ghstack PR number: https://github.com/pytorch/executorch/pull/21474 by
@Gasoonjia
^ Please use this as the source of truth for the PR details, comments,
and reviews
ghstack PR base:
https://github.com/pytorch/executorch/tree/gh/gasoonjia/179/base
ghstack PR head:
https://github.com/pytorch/executorch/tree/gh/gasoonjia/179/head
Merge bot PR base:
https://github.com/pytorch/executorch/tree/gh/gasoonjia/178/orig
Merge bot PR head:
https://github.com/pytorch/executorch/tree/gh/gasoonjia/179/orig
Differential Revision:
[D114032330](https://our.internmc.facebook.com/intern/diff/D114032330/)
@diff-train-skip-merge

---------

Co-authored-by: gasoonjia <gasoonjia@icloud.com>
Co-authored-by: Gasoonjia <gasoonjia@meta.com>
P
pytorchbot committed
6898c99ac615d639d3d0164b6df8ec79eaa27b44
Parent: f1ab0f8
Committed by GitHub <noreply@github.com> on 7/30/2026, 10:47:05 PM