This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
eeabf740bb4645f97b0db093c29745039e6b1891
vllm
/
tests
/
v1
/
attention
History
haosdent
116ed130f4
[Bugfix] Fix GDN attention crash with mixed decode/spec-decode batches (
#34871
)
...
Signed-off-by: haosdent <
haosdent@gmail.com
>
2026-03-16 10:30:23 +01:00
..
test_attention_backends_selection.py
…
test_attention_backends.py
[ROCm] AITER fused RoPE+KVCache (
#33443
)
2026-02-23 19:06:00 -08:00
test_attention_splitting.py
…
test_batch_reordering.py
…
test_chunked_local_attention.py
…
test_gdn_metadata_builder.py
[Bugfix] Fix GDN attention crash with mixed decode/spec-decode batches (
#34871
)
2026-03-16 10:30:23 +01:00
test_mamba_update_block_table.py
[Model][Spec Decode] Nemotron-H MTP and Mamba Speculative Decoding Support (
#33726
)
2026-02-24 09:49:56 -08:00
test_mla_backends.py
[Perf] Support FP8 KV cache for Flashinfer MLA Sparse (
#35891
)
2026-03-07 13:51:54 -08:00
test_rocm_attention_backends_selection.py
[ROCm][CI] Fix ROCm attention backend validation for head sizes, block sizes, and compute capability checks (
#36292
)
2026-03-09 12:02:41 -05:00
test_sparse_mla_backends.py
[Perf] Support FP8 KV cache for Flashinfer MLA Sparse (
#35891
)
2026-03-07 13:51:54 -08:00
test_trtllm_attention_integration.py
TRTLLM gen-full attn Test Coverage (
#34986
)
2026-03-03 11:35:34 -05:00
utils.py
[V0 Deprecation] Remove unused swap_space parameter (
#36216
)
2026-03-07 22:09:55 +08:00