TensorRT-LLMs

mirror of https://github.com/NVIDIA/TensorRT-LLM.git synced 2026-01-14 06:27:45 +08:00

History

wili eba3623a54 Feat: Variable-Beam-Width-Search (VBWS) part4 (#3979 ) * feat/vbws-part4-v1.8: rebase Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * feat/vbws-part4-v1.9: fix incorrect output when using short output length Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * v1.9.1: remove useless variables Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * v1.9.2:fix incorrect output when using short output length Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * v1.9.3: rebase Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * v1.9.4: rebase Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> * v1.9.5: remove API change Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> --------- Signed-off-by: wili-65535 <wili-65535@users.noreply.github.com> Co-authored-by: wili-65535 <wili-65535@users.noreply.github.com>		2025-05-12 22:32:29 +02:00
..
_static	doc: update switcher.json config (#4220 )	2025-05-12 20:40:55 +08:00
_templates	Update TensorRT-LLM (#1725 )	2024-06-04 20:26:32 +08:00
advanced	Feat: Variable-Beam-Width-Search (VBWS) part4 (#3979 )	2025-05-12 22:32:29 +02:00
architecture	doc: fix path after examples migration (#3814 )	2025-04-24 02:36:45 +08:00
blogs	chore: bump version to 0.19.0 (#3598 ) (#3841 )	2025-04-29 16:57:22 +08:00
commands	feat: trtllm-serve multimodal support (#3590 )	2025-04-19 05:01:28 +08:00
dev-on-cloud	doc: add doc ahout developent on cloud or runpod (#3194 )	2025-04-02 18:10:56 +08:00
examples	doc: refactor trtllm-serve examples and doc (#3187 )	2025-04-04 11:40:43 +08:00
installation	relax the limitation of setuptools (#2992 )	2025-03-24 13:36:10 +08:00
llm-api	doc: fix path after examples migration (#3814 )	2025-04-24 02:36:45 +08:00
media	L4 added to readme (#3301 )	2025-04-06 19:09:28 +08:00
performance	doc: TRTLLM-4797 Update perf-analysis.md (#4100 )	2025-05-08 17:24:44 +08:00
python-api	Update TensorRT-LLM (#1492 )	2024-04-24 14:44:22 +08:00
reference	chore: fix some invalid paths of contrib models (#3818 )	2025-04-24 05:36:16 +08:00
torch	Remove dummy forward path (#3669 )	2025-04-18 16:17:50 +08:00
conf.py	doc: update switcher.json config (#4220 )	2025-05-12 20:40:55 +08:00
helper.py	doc: refactor trtllm-serve examples and doc (#3187 )	2025-04-04 11:40:43 +08:00
index.rst	doc: refactor trtllm-serve examples and doc (#3187 )	2025-04-04 11:40:43 +08:00
key-features.md	Update TensorRT-LLM (#2562 )	2024-12-11 00:31:05 -08:00
overview.md	chore: Mass integration of release/0.18 (#3421 )	2025-04-16 10:03:29 +08:00
quick-start-guide.md	doc: fix path after examples migration (#3814 )	2025-04-24 02:36:45 +08:00
release-notes.md	chore: Mass integration of release/0.18 (#3421 )	2025-04-16 10:03:29 +08:00
torch.md	chore: bump version to 0.19.0 (#3598 ) (#3841 )	2025-04-29 16:57:22 +08:00