Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
513298f1b44157f7ae2f7007ef7b17c2929d11d4
vllm/tests/models
History
Chen Zhang 2b4fc9bd9b Support FlashAttention Backend for Hybrid SSM Models (#23299)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
2025-08-26 12:41:52 +00:00
..
fixtures
[Mistral-Small 3.1] Update docs and tests (#14977)
2025-03-18 03:29:42 -07:00
language
Support FlashAttention Backend for Hybrid SSM Models (#23299)
2025-08-26 12:41:52 +00:00
multimodal
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
quantization
[V0 Deprecation] Remove V0 FlashInfer attention backend (#22776)
2025-08-18 19:54:16 -07:00
__init__.py
[CI/Build] Move test_utils.py to tests/utils.py (#4425)
2024-05-13 23:50:09 +09:00
registry.py
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
test_initialization.py
[Model] Add LFM2 architecture (#22845)
2025-08-21 09:35:07 +02:00
test_oot_registration.py
[CI/Build] Fix plugin tests (#21758)
2025-07-28 15:08:05 +00:00
test_registry.py
[Deprecation][2/N] Replace --task with --runner and --convert (#21470)
2025-07-27 19:42:40 -07:00
test_transformers.py
Enable headless models for pooling in the Transformers backend (#21767)
2025-08-01 10:31:29 -07:00
test_utils.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
test_vision.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
utils.py
[Model] Pooling models default to using chunked prefill & prefix caching if supported. (#20930)
2025-08-11 09:41:37 -07:00
Powered by Gitea Version: 1.25.2 Page: 41ms Template: 11ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API