Yiqing Yan
|
05dd437084
|
[https://nvbugs/5565541][fix] Add timeout threshold for H100 FHMA test (#8354)
Signed-off-by: Yiqing Yan <yiqingy@nvidia.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
bhsueh_NV
|
69325e1aa3
|
[https://nvbugs/5574556][fix] fix bug of Qwen3_235B_A22B::test_fp8 CI (#8351)
Signed-off-by: bhsueh <11360707+byshiue@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Lizhi Zhou
|
982d4b65e8
|
[https://nvbugs/5550671][fix] fix disagg-serving multinodes test failure (#8307)
Signed-off-by: Lizhi Zhou <1432185+reasonsolo@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Chuang Zhu
|
18a534d2b4
|
[https://nvbugs/5465642][fix] Increase server timeout to wait weight loading (#8297)
Signed-off-by: Chuang Zhu <111838961+chuangz0@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Enwei Zhu
|
526cad37d7
|
[https://nvbugs/5568951][fix] Fix guided decoding disagg tests (#8311)
Signed-off-by: Enwei Zhu <21126786+syuoni@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Ivy Zhang
|
1b559ba91d
|
[None][chore] Update test configs for release (#8224)
Signed-off-by: Ivy Zhang <25222398+crazydemo@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Ivy Zhang
|
4789c1e588
|
[TRTLLM-8246][test] add multimodal kvcache+chunked_prefil cases in to QA test list (#8212)
Signed-off-by: Ivy Zhang <25222398+crazydemo@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Ivy Zhang
|
be2ab98233
|
[None][chore] Update constaintfor release (#8211)
Signed-off-by: Ivy Zhang <25222398+crazydemo@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
Yukun He
|
179c7dc501
|
[https://nvbugs/5536131][fix] Fix illegal access issue when scale is not provided in Llama3/4. (#7960)
Signed-off-by: Yukun He <23156053+hyukn@users.noreply.github.com>
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-16 22:46:19 +08:00 |
|
xinhe-nv
|
f70eff30b3
|
[TRTLLM-8638][fix] waive llam4 tests on H20 (#8416)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-16 03:14:56 -07:00 |
|
HuiGao-NV
|
4e6a492aa3
|
[None][chore] Isolate several intermittent cases (#8408)
Signed-off-by: Hui Gao <huig@nvidia.com>
|
2025-10-15 23:48:31 -07:00 |
|
xiweny
|
4143887370
|
[https://nvbugs/5541494] [fix] Remove waivers (#8353)
Signed-off-by: xiweny <13230610+VALLIS-NERIA@users.noreply.github.com>
|
2025-10-15 19:10:35 -07:00 |
|
Chuang Zhu
|
40d129a415
|
[None][fix] Fix cache buffer size for window (#8320)
Signed-off-by: Chuang Zhu <111838961+chuangz0@users.noreply.github.com>
|
2025-10-16 09:01:11 +08:00 |
|
dongfengy
|
7a0aa64973
|
[None][fix] Refactor triton paddings (#6980)
Signed-off-by: Dongfeng Yu <dongfengy@nvidia.com>
Signed-off-by: dongfengy <99041270+dongfengy@users.noreply.github.com>
Co-authored-by: hlu1 <14827759+hlu1@users.noreply.github.com>
|
2025-10-15 12:59:01 -07:00 |
|
mpikulski
|
0510b34588
|
[TRTLLM-8551][feat] add cache_salt in LLM.generate and refactor test_return_logits.py (#8317)
Signed-off-by: ixlmar <206748156+ixlmar@users.noreply.github.com>
|
2025-10-15 02:53:57 -07:00 |
|
QI JUN
|
1a1c9a29ab
|
[None][ci] move all llama4 test cases to post merge (#8387)
Signed-off-by: junq <22017000+QiJune@users.noreply.github.com>
|
2025-10-15 16:36:37 +08:00 |
|
mpikulski
|
93a4b7f1b6
|
[None][chore] update torch_dtype -> dtype in 'transformers' (#8263)
Signed-off-by: ixlmar <206748156+ixlmar@users.noreply.github.com>
|
2025-10-15 17:09:30 +09:00 |
|
Jin Li
|
206a9930df
|
[https://nvbugs/5547435][fix] Fix a merge conflict (#8365)
Signed-off-by: Jin Li <59594262+liji-nv@users.noreply.github.com>
|
2025-10-15 10:43:10 +08:00 |
|
Emma Qiao
|
493da020c1
|
[TRTLLM-7351][infra] Add isolate marker for L0 (#7497)
Signed-off-by: qqiao <qqiao@nvidia.com>
Signed-off-by: Emma Qiao <qqiao@nvidia.com>
Co-authored-by: Yanchao Lu <yanchaol@nvidia.com>
|
2025-10-14 16:58:14 -07:00 |
|
dongfengy
|
9d855f47ad
|
[None][fix] Remove outdated test waives for GPTOSS (#8183)
Signed-off-by: Dongfeng Yu <dongfengy@nvidia.com>
|
2025-10-14 16:20:38 -07:00 |
|
Michal Guzek
|
1cdb0b62c3
|
[https://nvbugs/5563469][fix] Temporarily disable test_nemotron_nano_8b_lora_torch in L0 due to Torch non-determinism (#8206)
Signed-off-by: Michal Guzek <mguzek@nvidia.com>
|
2025-10-14 17:55:28 +02:00 |
|
William Zhang
|
72d65d079a
|
[https://nvbugs/5542878][fix] Unwaive test (#8027)
Signed-off-by: William Zhang <133824995+2ez4bz@users.noreply.github.com>
|
2025-10-14 07:58:07 +02:00 |
|
xinhe-nv
|
371fcb0338
|
[TRTLLM-8366][feat] add kimi multi nodes case (#8025)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-13 21:36:03 -07:00 |
|
Yuxian Qiu
|
3450fe9944
|
[None][fix] Fix dummy load format for key models. (#7993)
Signed-off-by: Yuxian Qiu <142763828+yuxianq@users.noreply.github.com>
|
2025-10-14 11:18:39 +08:00 |
|
Robin Kobus
|
db8c63b9b1
|
[TRTLLM-4517] [feat] Additional model outputs (#7206)
Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>
|
2025-10-13 15:33:18 +02:00 |
|
xinhe-nv
|
9fe63dd8db
|
[None][chore] Add failed cases into waives.txt (#8290)
Signed-off-by: xinhe-nv <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-13 00:07:00 -07:00 |
|
xinhe-nv
|
72fcff1044
|
[None][fix] add timeout for llama4 (#8254)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-12 21:04:20 -07:00 |
|
Guoming Zhang
|
989c25fcba
|
[None][doc] Add qwen3-next doc into deployment guid and test case into L0. (#8288)
Signed-off-by: nv-guomingz <137257613+nv-guomingz@users.noreply.github.com>
Co-authored-by: Faradawn Yang <faradawny@gmail.com>
Co-authored-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>
|
2025-10-13 10:25:45 +08:00 |
|
Emma Qiao
|
fdbeea51d3
|
[None][infra] Skip failed cases for main branch (#8293)
Signed-off-by: qqiao <qqiao@nvidia.com>
|
2025-10-12 08:04:09 -07:00 |
|
brb-nv
|
56a539cd37
|
[None][chore] Waive failing pre-merge test on main (#8282)
Signed-off-by: Balaram Buddharaju <169953907+brb-nv@users.noreply.github.com>
|
2025-10-10 23:52:05 -07:00 |
|
Yilin Fan
|
2695d70d42
|
[None][feat] Add request timing breakdown option in benchmark_serving (#8128)
Signed-off-by: nv-yilinf <206948969+nv-yilinf@users.noreply.github.com>
|
2025-10-10 09:24:54 -07:00 |
|
xinhe-nv
|
2655995a09
|
[None][fix] add gc for test fixture (#8220)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-10 02:50:25 -07:00 |
|
bhsueh_NV
|
d3059dbd8a
|
[https://nvbugs/5547416][fix] unwaive no_cache test (#8213)
Signed-off-by: bhsueh <11360707+byshiue@users.noreply.github.com>
|
2025-10-10 01:50:13 -07:00 |
|
xinhe-nv
|
b555f1ff98
|
[None][chore] Add failed cases into waives.txt (#8229)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-09 23:45:28 -07:00 |
|
xinhe-nv
|
e8c9bae37e
|
[None][chore] Remove closed bugs (#8151)
Signed-off-by: Xin He (SW-GPU) <200704525+xinhe-nv@users.noreply.github.com>
|
2025-10-10 16:39:40 +11:00 |
|
Emma Qiao
|
ccd949ea5b
|
[None][infra] Waive failed tests on main 10/09 (#8230)
Signed-off-by: qqiao <qqiao@nvidia.com>
|
2025-10-09 22:46:07 +08:00 |
|
bhsueh_NV
|
27677a36f5
|
[https://nvbugs/5516666][fix] unwaive some Qwen3 CI tests (#8130)
Signed-off-by: bhsueh <11360707+byshiue@users.noreply.github.com>
|
2025-10-09 09:44:58 +08:00 |
|
Lizhi Zhou
|
fdf29ab8fa
|
[TRTLLM-7846][feat] Http disagg-cluster management implemention (#7869)
Signed-off-by: Lizhi Zhou <1432185+reasonsolo@users.noreply.github.com>
|
2025-10-09 09:44:01 +08:00 |
|
QI JUN
|
6884d06aed
|
[None][ci] move some llama4 test cases to pre merge (#8189)
Signed-off-by: junq <22017000+QiJune@users.noreply.github.com>
|
2025-10-08 18:34:08 -07:00 |
|
Liao Lanyu
|
ed8e00ad4a
|
[https://nvbugs/5522746][fix] unwaive tests caused by node issues after rebooting (#8193)
Signed-off-by: Lanyu Liao <lancelly@users.noreply.github.com>
Co-authored-by: Lanyu Liao <lancelly@users.noreply.github.com>
|
2025-10-09 08:45:56 +08:00 |
|
Mike Iovine
|
c88913dc03
|
[https://nvbugs/5541545][fix] Remove test_llama4 (#8031)
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-08 15:20:15 -07:00 |
|
brb-nv
|
80517b7812
|
[None][chore] Waive some tests failing on main post merge (#8186)
Signed-off-by: Balaram Buddharaju <169953907+brb-nv@users.noreply.github.com>
|
2025-10-08 06:52:30 -07:00 |
|
mpikulski
|
8298e93bd8
|
[TRTLLM-8414][chore] BREAKING CHANGE: refine sampling strategy selection (#8132)
Signed-off-by: ixlmar <206748156+ixlmar@users.noreply.github.com>
|
2025-10-08 15:46:50 +02:00 |
|
Liao Lanyu
|
d57b8f0951
|
[https://nvbugs/5455140][fix] unwaive tests related to GB200 OOM (#8159)
Signed-off-by: Lanyu Liao <lancelly@users.noreply.github.com>
Co-authored-by: Lanyu Liao <lancelly@users.noreply.github.com>
|
2025-10-08 13:14:12 +08:00 |
|
ruodil
|
971610e3ff
|
[None][test] add test-model-suites option in integration conftest.py (#8016)
Signed-off-by: Ruodi Lu <ruodil@users.noreply.github.com>
Co-authored-by: Ruodi Lu <ruodil@users.noreply.github.com>
|
2025-10-08 10:38:31 +08:00 |
|
Mike Iovine
|
7facac077b
|
[None][fix] Fix MTP illegal memory access (#8161)
Signed-off-by: Mike Iovine <6158008+mikeiovine@users.noreply.github.com>
|
2025-10-07 14:02:55 -04:00 |
|
Emma Qiao
|
ca9da1f1c2
|
[None][infra] Skip failed cases for main (#8176)
Signed-off-by: qqiao <qqiao@nvidia.com>
|
2025-10-07 06:37:51 -07:00 |
|
xiweny
|
9298f1bdcc
|
[None] [test] Add B300 cases to CI (#8056)
Signed-off-by: Xiwen Yu <13230610+VALLIS-NERIA@users.noreply.github.com>
|
2025-10-06 19:23:31 -07:00 |
|
Faraz
|
27a5091fcb
|
[None][feat] GPT-OSS Sm120/Sm121 Support (#7937)
Signed-off-by: Perkz Zheng <67892460+PerkzZheng@users.noreply.github.com>
Signed-off-by: list <58580514+farazkh80@users.noreply.github.com>
Signed-off-by: Vincent Huang <vincenth@nvidia.com>
Co-authored-by: Perkz Zheng <67892460+PerkzZheng@users.noreply.github.com>
Co-authored-by: Vincent Huang <vincenth@nvidia.com>
|
2025-10-06 16:59:06 -04:00 |
|
Lucas Liebenwein
|
3492391feb
|
[None][chore] AutoDeploy: clean up accuracy test configs (#8134)
Signed-off-by: Lucas Liebenwein <11156568+lucaslie@users.noreply.github.com>
|
2025-10-06 12:51:01 -07:00 |
|