TensorRT-LLMs/cpp/include/tensorrt_llm
Zheng Duan ce7f5fae5a
sort llm request state (#4607)
Signed-off-by: Zheng Duan <200704041+zhengd-nv@users.noreply.github.com>
2025-05-26 13:47:01 +08:00
..
batch_manager sort llm request state (#4607) 2025-05-26 13:47:01 +08:00
common feat: NIXL interface integration (#3934) 2025-05-19 18:18:22 +08:00
deep_gemm Feat: add deep_gemm swapab Kernel (#4430) 2025-05-21 10:48:43 +08:00
executor Agent interface impl for NIXL (#4125) 2025-05-22 09:09:41 +08:00
kernels Update TensorRT-LLM (#2873) 2025-03-11 21:13:42 +08:00
layers v1.2 (#3082) 2025-03-26 23:31:29 +08:00
plugins/api Update TensorRT-LLM (#2532) 2024-12-04 21:16:56 +08:00
runtime fix: [nvbugs/5287097] Align PP layer distribution between pytorch and TRT flow. (#4399) 2025-05-19 14:25:36 -07:00