Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
9b0187003e62bdb7311b23b5b5026ea8e4e207d3
vllm/tests/models
History
Chen Zhang 2b4fc9bd9b Support FlashAttention Backend for Hybrid SSM Models (#23299)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
2025-08-26 12:41:52 +00:00
..
fixtures
[Mistral-Small 3.1] Update docs and tests (#14977)
2025-03-18 03:29:42 -07:00
language
Support FlashAttention Backend for Hybrid SSM Models (#23299)
2025-08-26 12:41:52 +00:00
multimodal
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
quantization
[V0 Deprecation] Remove V0 FlashInfer attention backend (#22776)
2025-08-18 19:54:16 -07:00
__init__.py
[CI/Build] Move test_utils.py to tests/utils.py (#4425)
2024-05-13 23:50:09 +09:00
registry.py
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
test_initialization.py
[Model] Add LFM2 architecture (#22845)
2025-08-21 09:35:07 +02:00
test_oot_registration.py
[CI/Build] Fix plugin tests (#21758)
2025-07-28 15:08:05 +00:00
test_registry.py
[Deprecation][2/N] Replace --task with --runner and --convert (#21470)
2025-07-27 19:42:40 -07:00
test_transformers.py
Enable headless models for pooling in the Transformers backend (#21767)
2025-08-01 10:31:29 -07:00
test_utils.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
test_vision.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
utils.py
[Model] Pooling models default to using chunked prefill & prefix caching if supported. (#20930)
2025-08-11 09:41:37 -07:00
Powered by Gitea Version: 1.25.2 Page: 303ms Template: 3ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API