mirror of
https://github.com/NVIDIA/TensorRT-LLM.git
synced 2026-02-08 04:01:51 +08:00
Signed-off-by: Fred Wei <20514172+WeiHaocheng@users.noreply.github.com> Signed-off-by: zheyuf <zheyuf@NVIDIA.com> Co-authored-by: zheyuf <zheyuf@NVIDIA.com> |
||
|---|---|---|
| .. | ||
| media | ||
| tech_blog | ||
| Best_perf_practice_on_DeepSeek-R1_in_TensorRT-LLM.md | ||
| Falcon180B-H200.md | ||
| H100vsA100.md | ||
| H200launch.md | ||
| quantization-in-TRT-LLM.md | ||
| XQA-kernel.md | ||