TensorRT-LLMs/cpp/tensorrt_llm/plugins/mixtureOfExperts
Pamela Peng 52465216f4
[https://nvbugs/5295389][fix]fix moe fp4 on sm120 (#4624)
Signed-off-by: Pamela Peng <179191831+pamelap-nvidia@users.noreply.github.com>
2025-05-29 09:50:47 -07:00
..
CMakeLists.txt Update TensorRT-LLM (#524) 2023-12-01 22:27:51 +08:00
mixtureOfExpertsPlugin.cpp [https://nvbugs/5295389][fix]fix moe fp4 on sm120 (#4624) 2025-05-29 09:50:47 -07:00
mixtureOfExpertsPlugin.h feat: support add internal cutlass kernels as subproject (#3658) 2025-05-06 11:35:07 +08:00