..
apps
[None][chroe] Rename TensorRT-LLM to TensorRT LLM for source code. ( #7851 )
2025-09-25 21:02:35 +08:00
auto_deploy
[ #8921 ][chore] AutoDeploy NanoV3 to use SYMM_MEM allreduce strategy ( #9797 )
2025-12-09 13:05:38 -08:00
bindings /executor
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
configs
[TRTC-43] [feat] Add config db and docs ( #9420 )
2025-12-12 04:00:03 +08:00
cpp /executor
[TRTLLM-9197][infra] Move thirdparty stuff to it's own listfile ( #8986 )
2025-11-20 16:44:23 -08:00
cpp_library
[None][chroe] Rename TensorRT-LLM to TensorRT LLM for source code. ( #7851 )
2025-09-25 21:02:35 +08:00
disaggregated
[None][chore] Add GB300 support since it does not support segment ( #9731 )
2025-12-10 18:36:55 -08:00
dora
Update TensorRT-LLM ( #2755 )
2025-02-11 03:01:00 +00:00
draft_target_model
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
eagle
[None][chroe] Rename TensorRT-LLM to TensorRT LLM for source code. ( #7851 )
2025-09-25 21:02:35 +08:00
infinitebench
Update TensorRT-LLM ( #1725 )
2024-06-04 20:26:32 +08:00
language_adapter
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
layer_wise_benchmarks
[None][feat] Add weights initialization and context phase parser to layer-wise benchmarks ( #9667 )
2025-12-04 13:41:15 +08:00
llm-api
[TRTLLM-9089][chore] Port prepare_dataset into trtllm-bench ( #9250 )
2025-12-08 10:37:40 -08:00
llm-eval /lm-eval-harness
[TRTLLM-9065][chore] remove PyTorchConfig completely ( #8856 )
2025-11-06 22:37:03 -08:00
longbench
[None] [feat] Optimize the algorithm part of RocketKV ( #9333 )
2025-12-01 09:04:09 +08:00
lookahead
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
medusa
[OMNIML-3036][doc] Re-branding TensorRT-Model-Optimizer as Nvidia Model-Optimizer ( #9679 )
2025-12-07 07:14:05 -08:00
models
[TRTC-43] [feat] Add config db and docs ( #9420 )
2025-12-12 04:00:03 +08:00
ngram
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
openai_triton
[None][chroe] Rename TensorRT-LLM to TensorRT LLM for source code. ( #7851 )
2025-09-25 21:02:35 +08:00
opentelemetry
[None][chore] Change trt-server to trtlllm-server in opentelemetry readme ( #9173 )
2025-11-17 22:02:24 -08:00
python_plugin
[None][doc] Rename TensorRT-LLM to TensorRT LLM for homepage and the … ( #7850 )
2025-09-25 21:02:35 +08:00
quantization
[OMNIML-3036][doc] Re-branding TensorRT-Model-Optimizer as Nvidia Model-Optimizer ( #9679 )
2025-12-07 07:14:05 -08:00
ray_orchestrator
[None][chore] Use cached model in all ray tests ( #8962 )
2025-11-06 15:14:15 +01:00
redrafter
[None][chore] update torch_dtype -> dtype in 'transformers' ( #8263 )
2025-10-15 17:09:30 +09:00
sample_weight_stripping
[None][chore] Weekly mass integration of release/1.1 -- rebase ( #9522 )
2025-11-29 21:48:48 +08:00
scaffolding
[None][feat] Deep Research Implemented with Scaffolding ( #8452 )
2025-11-06 10:33:28 +08:00
serve
[None][doc] VDR 1.0 trtllm-serve doc enhancement ( #9443 )
2025-12-05 17:50:12 -05:00
sparse_attention
[None][feat] Add RocketKV usage doc and e2e accuracy test on LongBenchV2 ( #9572 )
2025-12-03 11:33:46 +08:00
trtllm-eval
test: Add LLGuidance test and refine guided decoding ( #5348 )
2025-06-25 14:12:56 +08:00
wide_ep
[TRTLLM-9706] [doc] Update wide EP documents ( #9724 )
2025-12-08 11:21:11 +08:00
constraints.txt
[None][chore] bump version to 1.2.0rc6 ( #9874 )
2025-12-10 04:53:26 -08:00
eval_long_context.py
[None][feat] Support ignored prompt length for penalties via new sampling config parameter ( #8127 )
2025-10-27 13:12:31 -04:00
generate_checkpoint_config.py
[None][chroe] Rename TensorRT-LLM to TensorRT LLM for source code. ( #7851 )
2025-09-25 21:02:35 +08:00
generate_xgrammar_tokenizer_info.py
Update TensorRT-LLM ( #2783 )
2025-02-13 18:40:22 +08:00
hf_lora_convert.py
Update TensorRT-LLM ( #2755 )
2025-02-11 03:01:00 +00:00
mmlu.py
[None][chore] update torch_dtype -> dtype in 'transformers' ( #8263 )
2025-10-15 17:09:30 +09:00
run.py
[None][feat] Support ignored prompt length for penalties via new sampling config parameter ( #8127 )
2025-10-27 13:12:31 -04:00
summarize.py
[None][feat] Support ignored prompt length for penalties via new sampling config parameter ( #8127 )
2025-10-27 13:12:31 -04:00
utils.py
[None][feat] Support ignored prompt length for penalties via new sampling config parameter ( #8127 )
2025-10-27 13:12:31 -04:00