This website requires JavaScript.
Explore
Help
Sign In
obscura
/
vllm
Watch
2
Star
0
Fork
0
mirror of
https://github.com/vllm-project/vllm.git
synced
2026-08-20 20:50:15 +00:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
optimize-prefix-caching-scheduling
vllm
/
csrc
/
cpu
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Yuan
and
GitHub
cafb8e06c5
[CI/BUILD] enable intel queue for longer CPU tests (
#4113
)
2024-06-03 10:39:50 -07:00
..
activation.cpp
[CI/Build] Enforce style for C++ and CUDA code with
clang-format
(
#4722
)
2024-05-22 07:18:41 +00:00
attention.cpp
[Model] Support MAP-NEO model (
#5081
)
2024-05-30 19:24:41 -07:00
cache.cpp
[CI/Build] Enforce style for C++ and CUDA code with
clang-format
(
#4722
)
2024-05-22 07:18:41 +00:00
cpu_types.hpp
[Hardware][Intel] Add CPU inference backend (
#3634
)
2024-04-01 22:07:30 -07:00
layernorm.cpp
[CI/Build] Enforce style for C++ and CUDA code with
clang-format
(
#4722
)
2024-05-22 07:18:41 +00:00
pos_encoding.cpp
[CI/BUILD] enable intel queue for longer CPU tests (
#4113
)
2024-06-03 10:39:50 -07:00
pybind.cpp
[CI/Build] Enforce style for C++ and CUDA code with
clang-format
(
#4722
)
2024-05-22 07:18:41 +00:00