TensorRT-LLMs/tests/unittest/_torch/modules
2025-12-12 00:22:13 +08:00
..
tests_lora_modules [None][chore] update torch_dtype -> dtype in 'transformers' (#8263) 2025-10-15 17:09:30 +09:00
conftest.py [TRTLLM-9603][feat] Enable ConfigurableMoE test in the CI (#9645) 2025-12-08 10:19:40 +08:00
test_awq_quantization.py [OMNIML-2932] [feat] nvfp4 awq support (#8698) 2025-12-03 19:47:13 +02:00
test_fused_moe.py [TRTLLM-8959][feat] ConfigurableMoE support CUTLASS (#9772) 2025-12-12 00:22:13 +08:00
test_group_rmn_norm.py [None][ci] move unittests to sub-directories (#6635) 2025-08-20 05:42:22 -04:00
test_mla_helix.py [https://nvbugs/5708475][fix] Fix e2e eval accuracy for helix parallelism (#9647) 2025-12-03 15:13:59 +08:00
test_moe_host_sharer.py feat: large-scale EP(part 6: Online EP load balancer integration for GB200 nvfp4) (#4818) 2025-06-08 10:25:18 +08:00
test_moe_load_balancer.py [None][refactor] Refactor Torch Compile Backend, MoeLoadBalancer and warmup Logic (#6615) 2025-08-19 09:58:44 +08:00
test_moe_routing.py [None] [feat] Add model gpt-oss (#6645) 2025-08-07 03:04:18 -04:00
test_rotary_embedding.py [TRTLLM-7385][feat] Optimize Qwen2/2.5-VL performance (#7250) 2025-09-22 03:40:02 -07:00
test_triton_linear.py [None] [feat] Add model gpt-oss (#6645) 2025-08-07 03:04:18 -04:00