TensorRT-LLMs/cpp/tensorrt_llm/plugins/mixtureOfExperts
Daniel Stokes 3a4851b7c3
feat: Add Mixture of Experts FP8xMXFP4 support (#4750)
Signed-off-by: Daniel Stokes <40156487+djns99@users.noreply.github.com>
2025-06-09 13:25:04 +08:00
..
CMakeLists.txt Update TensorRT-LLM (#524) 2023-12-01 22:27:51 +08:00
mixtureOfExpertsPlugin.cpp feat: Add Mixture of Experts FP8xMXFP4 support (#4750) 2025-06-09 13:25:04 +08:00
mixtureOfExpertsPlugin.h feat: support add internal cutlass kernels as subproject (#3658) 2025-05-06 11:35:07 +08:00