TensorRT-LLMs

mirror of https://github.com/NVIDIA/TensorRT-LLM.git synced 2026-01-14 06:27:45 +08:00

History

Gabriel Wu 2e0cd7922e fix: add SM90 guard for FP8 Blockscale GEMM (#3575 ) * fix: add SM90 guard for FP8 Blockscale GEMM Signed-off-by: Zihua Wu <13583761+lucifer1004@users.noreply.github.com> * fix: add SM90 guard for FP8 Blockscale GEMM Signed-off-by: Zihua Wu <13583761+lucifer1004@users.noreply.github.com> --------- Signed-off-by: Zihua Wu <13583761+lucifer1004@users.noreply.github.com> Co-authored-by: Tao Li @ NVIDIA <tali@nvidia.com>		2025-04-16 14:44:37 +08:00
..
batch_manager	chore: Clean up cpp runtime (#3505 )	2025-04-14 18:00:03 +08:00
common	chore: Stabilize ABI boundary for internal kernel library (#3117 )	2025-04-11 15:07:50 +08:00
deep_gemm	fix: add SM90 guard for FP8 Blockscale GEMM (#3575 )	2025-04-16 14:44:37 +08:00
executor	chore: Clean up cpp runtime (#3537 )	2025-04-15 16:06:14 +08:00
kernels	Update TensorRT-LLM (#2873 )	2025-03-11 21:13:42 +08:00
layers	v1.2 (#3082 )	2025-03-26 23:31:29 +08:00
plugins/api	Update TensorRT-LLM (#2532 )	2024-12-04 21:16:56 +08:00
runtime	chore: Clean up cpp runtime (#3537 )	2025-04-15 16:06:14 +08:00