Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
79a1d25bbd09420ad7d8631671a70cbcaf496cfb
vllm/tests/v1/core
History
Chen Zhang f0d610a8ae [v1][KVCacheManager] Avoid full cache hit by controlling max_length (#17999)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
Co-authored-by: Woosuk Kwon <woosuk.kwon@berkeley.edu>
2025-05-13 06:50:38 +00:00
..
test_kv_cache_utils.py
[v1] Move block management logic from KVCacheManager to SpecializedManager (#17474)
2025-05-09 15:25:34 +00:00
test_prefix_caching.py
[v1] Move block management logic from KVCacheManager to SpecializedManager (#17474)
2025-05-09 15:25:34 +00:00
test_scheduler_e2e.py
[V1] Support long_prefill_token_threshold in v1 scheduler (#15419)
2025-03-25 14:22:26 -07:00
test_scheduler.py
[P/D] NIXL Integration (#17751)
2025-05-12 09:46:16 -07:00
test_specialized_manager.py
[v1][KVCacheManager] Avoid full cache hit by controlling max_length (#17999)
2025-05-13 06:50:38 +00:00
Powered by Gitea Version: 1.25.2 Page: 96ms Template: 7ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API