mirror of
https://github.com/NVIDIA/TensorRT-LLM.git
synced 2026-02-14 15:03:48 +08:00
* Update TensorRT-LLM --------- Co-authored-by: Bhuvanesh Sridharan <bhuvanesh.sridharan@sprinklr.com> Co-authored-by: Qingquan Song <ustcsqq@gmail.com> |
||
|---|---|---|
| .. | ||
| media | ||
| .gitkeep | ||
| Falcon180B-H200.md | ||
| H100vsA100.md | ||
| H200launch.md | ||
| quantization-in-TRT-LLM.md | ||
| XQA-kernel.md | ||