TensorRT-LLMs/docs/source/blogs
Kefeng-Duan 67949f7c39
Update README and add benchmarking blog for DeepSeek-R1 (#3232)
- Added a new entry in the README for the published benchmarking best practices for DeepSeek-R1.
- Introduced a new blog post detailing performance benchmarking configurations and procedures for DeepSeek-R1 in TensorRT-LLM, including installation, dataset preparation, and benchmarking steps for both B200 and H200 GPUs.

Signed-off-by: taoli <litaotju@users.noreply.github.com>
Co-authored-by: taoli <litaotju@users.noreply.github.com>
2025-04-10 17:00:49 +08:00
..
media Update TensorRT-LLM (#1358) 2024-03-26 20:47:14 +08:00
.gitkeep Add Latest News section (#315) 2023-11-08 15:04:33 +08:00
Best_perf_practice_on_DeepSeek-R1_in_TensorRT-LLM.md Update README and add benchmarking blog for DeepSeek-R1 (#3232) 2025-04-10 17:00:49 +08:00
Falcon180B-H200.md Update TensorRT-LLM (#1492) 2024-04-24 14:44:22 +08:00
H100vsA100.md Update TensorRT-LLM (#1492) 2024-04-24 14:44:22 +08:00
H200launch.md Update TensorRT-LLM (#1492) 2024-04-24 14:44:22 +08:00
quantization-in-TRT-LLM.md Update TensorRT-LLM (#2792) 2025-02-18 21:27:39 +08:00
XQA-kernel.md Update TensorRT-LLM (#2008) 2024-07-23 23:05:09 +08:00