TensorRT-LLMs/cpp/include/tensorrt_llm/executor
brb-nv 9a2b44d0f2
[None][chore] No-op changes to support context parallelism in disaggregated serving later (#7063)
Signed-off-by: Balaram Buddharaju <169953907+brb-nv@users.noreply.github.com>
2025-08-21 08:21:27 -07:00
..
cacheCommunicator.h Agent interface impl for NIXL (#4125) 2025-05-22 09:09:41 +08:00
dataTransceiverState.h [None][chore] No-op changes to support context parallelism in disaggregated serving later (#7063) 2025-08-21 08:21:27 -07:00
disaggServerUtil.h Update TensorRT-LLM (#2792) 2025-02-18 21:27:39 +08:00
executor.h [None][fix] acceptance rate calculation fix in benchmark_serving (#6746) 2025-08-19 17:29:36 +08:00
serialization.h [TRTLLM-6881][feat] Include attention dp rank info with KV cache events (#6563) 2025-08-07 14:17:07 +02:00
tensor.h Update TensorRT-LLM (#1918) 2024-07-09 14:42:22 +08:00
transferAgent.h Agent interface impl for NIXL (#4125) 2025-05-22 09:09:41 +08:00
types.h [TRTLLM-5000][feat] NGrams V2 (#4569) 2025-06-27 23:00:17 +08:00