[CUDA] MatMulBlockQuantizedFp8Weight: fold the W8A8 activation QDQ into the decode GEMV - #31481
Open
tianleiwu wants to merge 2 commits into
Open
[CUDA] MatMulBlockQuantizedFp8Weight: fold the W8A8 activation QDQ into the decode GEMV#31481tianleiwu wants to merge 2 commits into
tianleiwu wants to merge 2 commits into
GitHub Advanced Security / lintrunner
succeeded
Aug 2, 2026 in 2s
No new alerts in code changed by this pull request
Loading