Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
d34f5fe939eafa1eb4c6ecbe84020f09ebe09284
vllm/vllm/v1/core
History
Chauncey 61fbfe5274 [Bugfix] fixed inconsistent finish_reason handling between V0 and V1 engines (#27555)
Signed-off-by: chaunceyjiang <chaunceyjiang@gmail.com>
2025-10-28 02:18:08 +00:00
..
sched
[Bugfix] fixed inconsistent finish_reason handling between V0 and V1 engines (#27555)
2025-10-28 02:18:08 +00:00
__init__.py
[V1] Implement vLLM V1 [1/N] (#9289)
2024-10-22 01:24:07 -07:00
block_pool.py
[Core] Reuse empty block lists whenever possible in KVCacheBlocks to mitigate GC costs (#24964)
2025-10-14 12:58:43 -07:00
encoder_cache_manager.py
[Misc] Simplify max tokens in multimodal registry (#27500)
2025-10-24 23:56:01 -07:00
kv_cache_coordinator.py
[Core] Reuse empty block lists whenever possible in KVCacheBlocks to mitigate GC costs (#24964)
2025-10-14 12:58:43 -07:00
kv_cache_manager.py
[Metrics] [KVConnector] Add connector prefix cache hit rate stats (#26245)
2025-10-23 12:21:08 +02:00
kv_cache_utils.py
[Chore]:Extract math and argparse utilities to separate modules (#27188)
2025-10-26 04:03:32 -07:00
single_type_kv_cache_manager.py
[Chore]:Extract math and argparse utilities to separate modules (#27188)
2025-10-26 04:03:32 -07:00
Powered by Gitea Version: 1.25.2 Page: 407ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API