TensorRT-LLMs/docs/source
Kefeng-Duan 67949f7c39
Update README and add benchmarking blog for DeepSeek-R1 (#3232)
- Added a new entry in the README for the published benchmarking best practices for DeepSeek-R1.
- Introduced a new blog post detailing performance benchmarking configurations and procedures for DeepSeek-R1 in TensorRT-LLM, including installation, dataset preparation, and benchmarking steps for both B200 and H200 GPUs.

Signed-off-by: taoli <litaotju@users.noreply.github.com>
Co-authored-by: taoli <litaotju@users.noreply.github.com>
2025-04-10 17:00:49 +08:00
..
_templates Update TensorRT-LLM (#1725) 2024-06-04 20:26:32 +08:00
advanced Doc: update steps of using Draft-Target-Model (DTM) in the documents. (#3366) 2025-04-09 17:35:01 +08:00
architecture Update TensorRT-LLM (#2562) 2024-12-11 00:31:05 -08:00
blogs Update README and add benchmarking blog for DeepSeek-R1 (#3232) 2025-04-10 17:00:49 +08:00
commands doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
dev-on-cloud doc: add doc ahout developent on cloud or runpod (#3194) 2025-04-02 18:10:56 +08:00
examples doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
installation relax the limitation of setuptools (#2992) 2025-03-24 13:36:10 +08:00
llm-api Update TensorRT-LLM (#2755) 2025-02-11 03:01:00 +00:00
media L4 added to readme (#3301) 2025-04-06 19:09:28 +08:00
performance Update TensorRT-LLM (#2820) 2025-02-25 21:21:49 +08:00
python-api Update TensorRT-LLM (#1492) 2024-04-24 14:44:22 +08:00
reference Feat: Variable-Beam-Width-Search (VBWS) part3 (#3338) 2025-04-08 23:51:27 +08:00
torch feat: no-cache attention in PyTorch workflow (#3085) 2025-04-05 01:54:32 +08:00
conf.py Update (#2978) 2025-03-23 16:39:35 +08:00
helper.py doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
index.rst doc: refactor trtllm-serve examples and doc (#3187) 2025-04-04 11:40:43 +08:00
key-features.md Update TensorRT-LLM (#2562) 2024-12-11 00:31:05 -08:00
overview.md Update (#2978) 2025-03-23 16:39:35 +08:00
quick-start-guide.md Update (#2978) 2025-03-23 16:39:35 +08:00
release-notes.md Update TensorRT-LLM (#2820) 2025-02-25 21:21:49 +08:00
torch.md Update TensorRT-LLM (#2820) 2025-02-25 21:21:49 +08:00