This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
f5d17400303149bbb480f6abfb6f7bb646c1d895
vllm
/
vllm
/
v1
/
spec_decode
History
Ofir Zafrir
e9ec2a72d8
[Bugfix] Fix stale
common_attn_metadata.max_seq_len
in speculative decoding with Eagle (
#32312
)
...
Signed-off-by: Ofir Zafrir <
ofir.zafrir@intel.com
>
2026-01-15 06:39:37 +00:00
..
__init__.py
[V1][BugFix] Add __init__.py to v1/spec_decode/ (
#13359
)
2025-02-16 09:39:08 -08:00
eagle.py
[Bugfix] Fix stale
common_attn_metadata.max_seq_len
in speculative decoding with Eagle (
#32312
)
2026-01-15 06:39:37 +00:00
medusa.py
[V1][Spec Decode] Optimize Medusa proposer to avoid GPU-CPU sync (
#29723
)
2025-12-10 00:15:01 +00:00
metadata.py
[V1][spec decode] return logprobs for spec decoding (
#26060
)
2025-10-22 22:59:59 -07:00
metrics.py
[mypy] Enable type checking for more directories (
#29674
)
2025-11-28 08:39:27 -08:00
ngram_proposer.py
[Cleanup] Remove obsolete spec decoding compatibility logic (
#32003
)
2026-01-09 05:44:18 +00:00
suffix_decoding.py
[Cleanup] Remove obsolete spec decoding compatibility logic (
#32003
)
2026-01-09 05:44:18 +00:00
utils.py
[Cleanup] Remove obsolete spec decoding compatibility logic (
#32003
)
2026-01-09 05:44:18 +00:00