TensorRT-LLMs/cpp
Yao Yao 3545d59635
Support speculative decoding with Hopper XQA (#3269)
Signed-off-by: Yao Yao <lowsfer@users.noreply.github.com>
2025-04-07 17:14:34 +08:00
..
cmake fix #3109: early exit cmake if find_library() does not find any lib (#3113) 2025-03-29 19:59:03 +08:00
include/tensorrt_llm chore: remove usernames from comments (#3291) 2025-04-05 13:44:28 +08:00
micro_benchmarks perf: Add optimizations for deepseek in min latency mode (#3093) 2025-04-02 09:05:24 +08:00
tensorrt_llm Support speculative decoding with Hopper XQA (#3269) 2025-04-07 17:14:34 +08:00
tests chore: remove usernames from comments (#3291) 2025-04-05 13:44:28 +08:00
CMakeLists.txt fix: upgrade cmake minimum from 3.18 to 3.27 (#3208) 2025-04-02 15:14:36 +08:00