This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
1363e3d6d5659b58376fa5284afc2c8be548cc9d
vllm
/
csrc
/
quantization
/
fp4
History
Roberto L. Castro
fcb9df99bd
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
...
Signed-off-by: LopezCastroRoberto <
rocastro@redhat.com
>
2026-01-24 18:45:27 -07:00
..
activation_nvfp4_quant_fusion_kernels.cu
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
2026-01-24 18:45:27 -07:00
nvfp4_blockwise_moe_kernel.cu
[Perf] Fuse stride preparation for NVFP4 cutlass_moe (
#31837
)
2026-01-07 13:31:26 -05:00
nvfp4_experts_quant.cu
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
2026-01-24 18:45:27 -07:00
nvfp4_quant_entry.cu
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
2026-01-24 18:45:27 -07:00
nvfp4_quant_kernels.cu
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
2026-01-24 18:45:27 -07:00
nvfp4_scaled_mm_entry.cu
SM120 / NVFP4: add device guard and runtime SM dispatch to cutlass_scaled_fp4_mm (
#29711
)
2025-12-01 17:24:18 -08:00
nvfp4_scaled_mm_kernels.cu
…
nvfp4_scaled_mm_sm120_kernels.cu
Support CUTLASS NVFP4 (w4a4) for Blackwell Geforce GPUs (SM120) (
#21309
)
2025-08-03 00:54:22 -07:00
nvfp4_utils.cuh
[Perf][Kernel] Optimize FP4 quantization kernels (SM100F) (
#32520
)
2026-01-24 18:45:27 -07:00