Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
7ea22e42d5f666a26b3ce4117724dadfdb4d3887
vllm/tests/models
History
Chen Zhang 2b4fc9bd9b Support FlashAttention Backend for Hybrid SSM Models (#23299)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
2025-08-26 12:41:52 +00:00
..
fixtures
[Mistral-Small 3.1] Update docs and tests (#14977)
2025-03-18 03:29:42 -07:00
language
Support FlashAttention Backend for Hybrid SSM Models (#23299)
2025-08-26 12:41:52 +00:00
multimodal
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
quantization
[V0 Deprecation] Remove V0 FlashInfer attention backend (#22776)
2025-08-18 19:54:16 -07:00
__init__.py
[CI/Build] Move test_utils.py to tests/utils.py (#4425)
2024-05-13 23:50:09 +09:00
registry.py
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
test_initialization.py
[Model] Add LFM2 architecture (#22845)
2025-08-21 09:35:07 +02:00
test_oot_registration.py
[CI/Build] Fix plugin tests (#21758)
2025-07-28 15:08:05 +00:00
test_registry.py
[Deprecation][2/N] Replace --task with --runner and --convert (#21470)
2025-07-27 19:42:40 -07:00
test_transformers.py
Enable headless models for pooling in the Transformers backend (#21767)
2025-08-01 10:31:29 -07:00
test_utils.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
test_vision.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
utils.py
[Model] Pooling models default to using chunked prefill & prefix caching if supported. (#20930)
2025-08-11 09:41:37 -07:00
Powered by Gitea Version: 1.25.2 Page: 382ms Template: 11ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API