TensorRT-LLMs/docs/source
Robin Kobus d31fefde2c
[TRTLLM-5171] chore: Remove GptSession/V1 from TRT workflow (#4092)
* chore: Remove GptSession/V1 from TRT workflow

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove stateful decoders

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove GptSession buffers

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove GptSession utils

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove GptSession kernels

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove V1 GPT models from tests

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove gptSessionBenchmark from scripts and docs

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove gptSession IO classes

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove GptSession from test lists

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove GptSession from docs

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove useless encoder test

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove mActualBatchSize from DecoderState

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

* chore: Remove static batching from ExecutorTest

- Updated `validateContextLogits` and `validateGenerationLogits` functions to remove the `batchingType` parameter.
- Adjusted related test functions to reflect the changes in parameter lists.
- Cleaned up the instantiation of test cases to eliminate unnecessary batchingType references.

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>

---------

Signed-off-by: Robin Kobus <19427718+Funatiq@users.noreply.github.com>
2025-05-14 23:10:04 +02:00
..
_static doc: update switcher.json config (#4220) 2025-05-12 20:40:55 +08:00
_templates Update TensorRT-LLM (#1725) 2024-06-04 20:26:32 +08:00
advanced [TRTLLM-5171] chore: Remove GptSession/V1 from TRT workflow (#4092) 2025-05-14 23:10:04 +02:00
architecture doc: fix path after examples migration (#3814) 2025-04-24 02:36:45 +08:00
blogs chore: bump version to 0.19.0 (#3598) (#3841) 2025-04-29 16:57:22 +08:00
commands feat: trtllm-serve multimodal support (#3590) 2025-04-19 05:01:28 +08:00
dev-on-cloud doc: add doc ahout developent on cloud or runpod (#3194) 2025-04-02 18:10:56 +08:00
examples doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
installation [TRTLLM-5171] chore: Remove GptSession/V1 from TRT workflow (#4092) 2025-05-14 23:10:04 +02:00
llm-api doc: fix path after examples migration (#3814) 2025-04-24 02:36:45 +08:00
media L4 added to readme (#3301) 2025-04-06 19:09:28 +08:00
performance doc: TRTLLM-4797 Update perf-analysis.md (#4100) 2025-05-08 17:24:44 +08:00
python-api Update TensorRT-LLM (#1492) 2024-04-24 14:44:22 +08:00
reference [TRTLLM-5171] chore: Remove GptSession/V1 from TRT workflow (#4092) 2025-05-14 23:10:04 +02:00
torch Remove dummy forward path (#3669) 2025-04-18 16:17:50 +08:00
conf.py doc: update switcher.json config (#4220) 2025-05-12 20:40:55 +08:00
helper.py doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
index.rst doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
key-features.md Update TensorRT-LLM (#2562) 2024-12-11 00:31:05 -08:00
overview.md chore: Mass integration of release/0.18 (#3421) 2025-04-16 10:03:29 +08:00
quick-start-guide.md doc: fix path after examples migration (#3814) 2025-04-24 02:36:45 +08:00
release-notes.md chore: Mass integration of release/0.18 (#3421) 2025-04-16 10:03:29 +08:00
torch.md chore: bump version to 0.19.0 (#3598) (#3841) 2025-04-29 16:57:22 +08:00