TensorRT-LLMs

mirror of https://github.com/NVIDIA/TensorRT-LLM.git synced 2026-01-14 06:27:45 +08:00

History

Pamela Peng 6cdfc54883 feat: Add FP8 support for SM 120 (#3248 ) * Allow FP8 on SM120 Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> * fix sm121 Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> * fix Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> * fix pre-commit Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> * review update Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> --------- Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com> Co-authored-by: Sharan Chetlur <116769508+schetlur-nv@users.noreply.github.com>		2025-04-14 16:05:41 -07:00
..
__init__.py	Update TensorRT-LLM (#2936 )	2025-03-18 21:25:19 +08:00
cpp_paths.py	Update TensorRT-LLM (#2936 )	2025-03-18 21:25:19 +08:00
llm_data.py	Update TensorRT-LLM (#2936 )	2025-03-18 21:25:19 +08:00
runtime_defaults.py	Update TensorRT-LLM (#2936 )	2025-03-18 21:25:19 +08:00
test_medusa_utils.py	Update TensorRT-LLM (#2936 )	2025-03-18 21:25:19 +08:00
torch_ref.py	test: reorganize tests folder hierarchy (#2996 )	2025-03-27 12:07:53 +08:00
util.py	feat: Add FP8 support for SM 120 (#3248 )	2025-04-14 16:05:41 -07:00