This website requires JavaScript.
Explore
Help
Sign In
obscura
/
vllm
Watch
2
Star
0
Fork
0
mirror of
https://github.com/vllm-project/vllm.git
synced
2026-08-24 14:40:09 +00:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
main
vllm
/
csrc
/
libtorch_stable
/
quantization
/
w8a8
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Gabriel Wu
and
GitHub
2b7fcbf527
[Kernel] SM120: stop routing misaligned-M blockwise FP8 GEMMs to the small-M swapAB config (
#52775
)
...
Signed-off-by: Zihua Wu <
[email protected]
>
2026-08-19 21:25:29 +08:00
..
cutlass
[Kernel] SM120: stop routing misaligned-M blockwise FP8 GEMMs to the small-M swapAB config (
#52775
)
2026-08-19 21:25:29 +08:00
fp8
New stable abi cleanup (
#46656
)
2026-07-03 14:02:26 +08:00
int8
New stable abi cleanup (
#46656
)
2026-07-03 14:02:26 +08:00
per_token_group_quant_8bit.h
Faster per-token fp8 group quant packed kernel for blackwell (
#41326
)
2026-04-30 18:09:55 -07:00