TensorRT-LLMs/tensorrt_llm/_torch/compilation
Yi Zhang a69bd2a6fa
[https://nvbugs/5550409][fix] Disable torch compile in piecewise attention part to Avoid host overhead (#8708)
Signed-off-by: yizhang-nv <187001205+yizhang-nv@users.noreply.github.com>
2025-10-29 18:12:58 +08:00
..
multi_stream
patterns
__init__.py
backend.py
piecewise_optimizer.py [https://nvbugs/5550409][fix] Disable torch compile in piecewise attention part to Avoid host overhead (#8708) 2025-10-29 18:12:58 +08:00
recover_pass.py
remove_copy_pass.py
utils.py