This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
v0.14.1
vllm
/
tests
/
models
/
quantization
History
Matthew Bonanni
2612ba9285
[1/N][Attention] Restructure attention: move files (
#31916
)
...
Signed-off-by: Matthew Bonanni <
mbonanni@redhat.com
>
2026-01-09 13:10:24 -08:00
..
__init__.py
[CI/Build] Reorganize models tests (
#17459
)
2025-04-30 23:03:08 -07:00
test_awq.py
…
test_bitblas.py
…
test_bitsandbytes.py
Default model load/config/tokenizer to
mistral
format if relevant files exist (
#28659
)
2025-11-21 13:58:59 -08:00
test_fp8.py
[1/N][Attention] Restructure attention: move files (
#31916
)
2026-01-09 13:10:24 -08:00
test_gguf.py
[Bugfix][Quantization] Support BF16 tensors on GGUF (
#29948
)
2025-12-03 10:33:46 +00:00
test_gpt_oss_attn_quantization.py
[Quantization] fix attention quantization of gpt_oss model (
#27334
)
2025-11-11 12:06:00 -05:00
test_gptq_bitblas.py
Convert formatting to use
ruff
instead of
yapf
+
isort
(
#26247
)
2025-10-05 07:06:22 -07:00
test_gptq_marlin_24.py
[Quantization] Deprecate Long Tail of Schemes (
#31688
)
2026-01-08 19:07:45 -05:00
test_gptq_marlin.py
…
test_modelopt.py
…
test_mxfp4.py
…
test_nvfp4.py
…