This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
d34f5fe939eafa1eb4c6ecbe84020f09ebe09284
vllm
/
vllm
/
v1
/
core
History
Chauncey
61fbfe5274
[Bugfix] fixed inconsistent finish_reason handling between V0 and V1 engines (
#27555
)
...
Signed-off-by: chaunceyjiang <
chaunceyjiang@gmail.com
>
2025-10-28 02:18:08 +00:00
..
sched
[Bugfix] fixed inconsistent finish_reason handling between V0 and V1 engines (
#27555
)
2025-10-28 02:18:08 +00:00
__init__.py
[V1] Implement vLLM V1 [1/N] (
#9289
)
2024-10-22 01:24:07 -07:00
block_pool.py
[Core] Reuse empty block lists whenever possible in KVCacheBlocks to mitigate GC costs (
#24964
)
2025-10-14 12:58:43 -07:00
encoder_cache_manager.py
[Misc] Simplify max tokens in multimodal registry (
#27500
)
2025-10-24 23:56:01 -07:00
kv_cache_coordinator.py
[Core] Reuse empty block lists whenever possible in KVCacheBlocks to mitigate GC costs (
#24964
)
2025-10-14 12:58:43 -07:00
kv_cache_manager.py
[Metrics] [KVConnector] Add connector prefix cache hit rate stats (
#26245
)
2025-10-23 12:21:08 +02:00
kv_cache_utils.py
[Chore]:Extract math and argparse utilities to separate modules (
#27188
)
2025-10-26 04:03:32 -07:00
single_type_kv_cache_manager.py
[Chore]:Extract math and argparse utilities to separate modules (
#27188
)
2025-10-26 04:03:32 -07:00