This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
bb51d5b40db6076fc477df27a57e70b7421d87c1
vllm
/
csrc
/
quantization
History
mikaylagawarecki
ab1a6a43fa
[3/n] Migrate cutlass/scaled_mm_entry.cu torch stable ABI (
#37221
)
...
Signed-off-by: Mikayla Gawarecki <
mikaylagawarecki@gmail.com
>
2026-03-30 11:20:13 -07:00
..
awq
…
cutlass_w4a8
…
fp4
[NVIDIA] Bugfix NVFP4 DGX Spark and RTX50 (
#38423
)
2026-03-30 09:36:18 -07:00
fused_kernels
[2/n] Migrate per_token_group_quant to torch stable ABI (
#36058
)
2026-03-25 10:15:13 -07:00
gguf
…
gptq
…
gptq_allspark
…
hadamard
/hadacore
…
machete
[NVIDIA] Bugfix NVFP4 DGX Spark and RTX50 (
#38423
)
2026-03-30 09:36:18 -07:00
marlin
[Bugfix]fix output Nan/Inf in marlin if dtype=float16 (
#33972
)
2026-03-27 16:36:08 -07:00
w8a8
[3/n] Migrate cutlass/scaled_mm_entry.cu torch stable ABI (
#37221
)
2026-03-30 11:20:13 -07:00
activation_kernels.cu
[Bugfix]fix output Nan/Inf in marlin if dtype=float16 (
#33972
)
2026-03-27 16:36:08 -07:00
utils.cuh
[Feature][ROCm]Enable fusion pass for torch.compile on ROCm (
#15050
)
2025-03-31 04:42:18 -07:00