TensorRT-LLMs

mirror of https://github.com/NVIDIA/TensorRT-LLM.git synced 2026-02-17 00:04:57 +08:00

History

Wanli Jiang 421eb9e39c [None][feat] Optimize NemotronH model with elementwise and nvfp4 fusion (#11273 ) Signed-off-by: Wanli Jiang <35160485+Wanli-Jiang@users.noreply.github.com>		2026-02-12 09:25:31 -05:00
..
mamba	[None][feat] Optimize NemotronH model with elementwise and nvfp4 fusion (#11273 )	2026-02-12 09:25:31 -05:00
moe	[TRTLLM-9111][feat] provide the uniform test framework to test all MoE backends (#11128 )	2026-02-04 15:57:56 +08:00
tests_lora_modules	[None][chore] update torch_dtype -> dtype in 'transformers' (#8263 )	2025-10-15 17:09:30 +09:00
test_awq_quantization.py	[TRTLLM-9872][fix] clear the failed test at CI when enalbe_configurab… (#10067 )	2025-12-21 08:14:50 -05:00
test_fused_activation_quant.py	[None][feat] Optimize NemotronH model with elementwise and nvfp4 fusion (#11273 )	2026-02-12 09:25:31 -05:00
test_fused_add_rms_norm_quant.py	[None][feat] Optimize NemotronH model with elementwise and nvfp4 fusion (#11273 )	2026-02-12 09:25:31 -05:00
test_fused_moe.py	[TRTLLM-9111][feat] provide the uniform test framework to test all MoE backends (#11128 )	2026-02-04 15:57:56 +08:00
test_group_rmn_norm.py	[None][ci] move unittests to sub-directories (#6635 )	2025-08-20 05:42:22 -04:00
test_mla_helix.py	[None][fix] Remove unused params in attn (#10652 )	2026-01-20 03:08:59 -05:00
test_moe_host_sharer.py	feat: large-scale EP(part 6: Online EP load balancer integration for GB200 nvfp4) (#4818 )	2025-06-08 10:25:18 +08:00
test_moe_load_balancer.py	[None][refactor] Refactor Torch Compile Backend, MoeLoadBalancer and warmup Logic (#6615 )	2025-08-19 09:58:44 +08:00
test_moe_routing.py	[None] [feat] Add model gpt-oss (#6645 )	2025-08-07 03:04:18 -04:00
test_rotary_embedding.py	[TRTLLM-7385][feat] Optimize Qwen2/2.5-VL performance (#7250 )	2025-09-22 03:40:02 -07:00
test_triton_linear.py	[https://nvbugs/5761391 ][fix] Include triton-kernels as a packaged dependency (#10471 )	2026-01-28 19:56:32 -08:00