Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
ad0012a0ac507973656cfcf5750d603af4a5fdcc
vllm/tests/kernels/attention
History
Thomas Parnell e6b8e65d2d [Bugfix] Fix fp8 tests for triton_unified_attention for Triton 3.3 (#18013)
Signed-off-by: Thomas Parnell <tpa@zurich.ibm.com>
Co-authored-by: Lucas Wilkinson <lwilkinson@neuralmagic.com>
2025-05-15 13:26:34 +08:00
..
conftest.py
…
test_attention_selector.py
fix broken test vllm:test_kernels - test_attention_selector.py::test_flash_attn (#17873)
2025-05-10 10:46:54 +08:00
test_attention.py
…
test_blocksparse_attention.py
…
test_cache.py
…
test_cascade_flash_attn.py
…
test_encoder_decoder_attn.py
…
test_flash_attn.py
Update test_flash_attn.py (#17102)
2025-04-26 22:17:35 +00:00
test_flashinfer.py
…
test_flashmla.py
[Bugfix] Fix triton import with local TritonPlaceholder (#17446)
2025-05-06 17:53:09 +08:00
test_lightning_attn.py
…
test_merge_attn_states.py
…
test_mha_attn.py
…
test_mla_decode_cpu.py
…
test_prefix_prefill.py
…
test_rocm_attention_selector.py
[FEAT][ROCm]: Support AITER MLA on V1 Engine (#17523)
2025-05-09 10:42:05 +08:00
test_triton_decode_attention.py
…
test_triton_unified_attention.py
[Bugfix] Fix fp8 tests for triton_unified_attention for Triton 3.3 (#18013)
2025-05-15 13:26:34 +08:00
Powered by Gitea Version: 1.25.2 Page: 1158ms Template: 11ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API