This website requires JavaScript.
Explore
Help
Sign In
obscura
/
vllm
Watch
2
Star
0
Fork
0
mirror of
https://github.com/vllm-project/vllm.git
synced
2026-08-18 03:30:20 +00:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
main
vllm
/
tests
/
models
/
transformers
/
fusers
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Harry Mellor
and
GitHub
70b84f0bcb
Fix Gemma 4 for upcoming Transformers version (
#49797
)
...
Signed-off-by: Harry Mellor <
[email protected]
>
2026-08-10 07:48:26 +00:00
..
__init__.py
Make the Transformers modeling backend as fast as native vLLM (
#47187
)
2026-07-06 16:59:14 +01:00
test_linear.py
Fix Gemma 4 for upcoming Transformers version (
#49797
)
2026-08-10 07:48:26 +00:00
test_mla.py
Support MLA properly in the Transformers modeling backend (
#48250
)
2026-08-04 13:24:11 +01:00
test_moe.py
Fix MLA padding and grouped topk routing in the Transformers modelling backend (
#49982
)
2026-07-27 18:32:32 +00:00
test_rms_norm.py
fix: fuse weightless RMSNorms at their declared width (
#50867
)
2026-08-04 16:03:09 +00:00