This website requires JavaScript.
Explore
Help
Sign In
obscura
/
vllm
Watch
2
Star
0
Fork
0
mirror of
https://github.com/vllm-project/vllm.git
synced
2026-08-21 21:20:15 +00:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
running-deque
vllm
/
tests
/
neuron
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Liangfu Chen
and
GitHub
c91b64f749
[neuron] add reshape_and_cache (
#14391
)
2025-03-10 18:37:29 -07:00
..
test_activation.py
[Neuron] Add custom_ops for neuron backend (
#13246
)
2025-02-25 11:47:49 -08:00
test_block_table.py
[Neuron][Kernel] Vectorize KV cache load in FlashPagedAttention to maximize DMA bandwidth (
#13245
)
2025-02-20 17:45:45 -08:00
test_cache.py
[neuron] add reshape_and_cache (
#14391
)
2025-03-10 18:37:29 -07:00
test_comm_ops.py
[Neuron] Add Neuron device communicator for vLLM v1 (
#14085
)
2025-03-10 18:37:04 -07:00
test_layernorm.py
[Neuron] Add custom_ops for neuron backend (
#13246
)
2025-02-25 11:47:49 -08:00
test_logits_processor.py
Update deprecated Python 3.8 typing (
#13971
)
2025-03-02 17:34:51 -08:00
test_prefix_prefill.py
[Neuron] Add custom_ops for neuron backend (
#13246
)
2025-02-25 11:47:49 -08:00
test_rotary_embedding.py
[Neuron] Add custom_ops for neuron backend (
#13246
)
2025-02-25 11:47:49 -08:00