TensorRT-LLMs/cpp
Yibin Li 32ae1564bd
update FP4 quantize layout (#3045)
Signed-off-by: Yibin Li <109242046+yibinl-nvidia@users.noreply.github.com>
2025-04-03 13:13:54 -04:00
..
cmake fix #3109: early exit cmake if find_library() does not find any lib (#3113) 2025-03-29 19:59:03 +08:00
include/tensorrt_llm chore: Add output of first token to additional generation outputs (#3205) 2025-04-02 20:14:16 +08:00
micro_benchmarks perf: Add optimizations for deepseek in min latency mode (#3093) 2025-04-02 09:05:24 +08:00
tensorrt_llm update FP4 quantize layout (#3045) 2025-04-03 13:13:54 -04:00
tests update FP4 quantize layout (#3045) 2025-04-03 13:13:54 -04:00
CMakeLists.txt fix: upgrade cmake minimum from 3.18 to 3.27 (#3208) 2025-04-02 15:14:36 +08:00