This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
a4511e38db375a85b4dd784c2c38528747288f46
vllm
/
vllm
/
platforms
History
Matthew Bonanni
4c23690f43
[Attention] FlashAttention ViT support, make default backend (
#28763
)
...
Signed-off-by: Matthew Bonanni <
mbonanni@redhat.com
>
2025-11-18 20:06:21 -08:00
..
__init__.py
[TPU] Rename path to tpu platform (
#28452
)
2025-11-11 19:16:47 +00:00
cpu.py
[Misc] Make
SchedulerConfig.max_model_len
init-only (
#28733
)
2025-11-15 01:59:31 -08:00
cuda.py
[Attention] FlashAttention ViT support, make default backend (
#28763
)
2025-11-18 20:06:21 -08:00
interface.py
[CI Failure] Fix backend selection for encoder-only models (
#28534
)
2025-11-13 10:11:27 -05:00
rocm.py
Enable bitsandbytes quantization on AMD GPUs that use warp size 32 (
#27307
)
2025-11-19 03:12:31 +00:00
tpu.py
[Misc] Make
SchedulerConfig.max_model_len
init-only (
#28733
)
2025-11-15 01:59:31 -08:00
xpu.py
[Misc] Make
SchedulerConfig.max_model_len
init-only (
#28733
)
2025-11-15 01:59:31 -08:00