mirror of
https://github.com/NVIDIA/TensorRT-LLM.git
synced 2026-01-14 06:27:45 +08:00
- Added a new entry in the README for the published benchmarking best practices for DeepSeek-R1. - Introduced a new blog post detailing performance benchmarking configurations and procedures for DeepSeek-R1 in TensorRT-LLM, including installation, dataset preparation, and benchmarking steps for both B200 and H200 GPUs. Signed-off-by: taoli <litaotju@users.noreply.github.com> Co-authored-by: taoli <litaotju@users.noreply.github.com> |
||
|---|---|---|
| .. | ||
| media | ||
| .gitkeep | ||
| Best_perf_practice_on_DeepSeek-R1_in_TensorRT-LLM.md | ||
| Falcon180B-H200.md | ||
| H100vsA100.md | ||
| H200launch.md | ||
| quantization-in-TRT-LLM.md | ||
| XQA-kernel.md | ||