Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
a1946570d80c1bef78063e84b097951d8e8d4e6a
vllm/vllm/model_executor
History
tc-mb e042d7e685 Add flagos in MiniCPM-o (#34126)
Signed-off-by: tc-mb <caitianchi@modelbest.cn>
Signed-off-by: Vincent-Xiao <vincent.xiao.me@gmail.com>
Co-authored-by: Vincent-Xiao <vincent.xiao.me@gmail.com>
2026-02-10 02:51:48 -08:00
..
layers
[Bugfix][ROCm][GPT-OSS] Use old triton_kernels implementation on ROCm if the new API is not available (#34153)
2026-02-09 17:38:54 -06:00
model_loader
[Bugfix] Sort hf_weights_files in fastsafetensors_weights_iterator to match #33491 (#34190)
2026-02-09 23:06:30 -08:00
models
Add flagos in MiniCPM-o (#34126)
2026-02-10 02:51:48 -08:00
warmup
[Kernel] Add KernelConfig flag to enable/disable FlashInfer autotune (#34006)
2026-02-07 05:24:44 -08:00
__init__.py
[Platform] Deprecate seed_everything (#31659)
2026-01-04 18:34:04 -08:00
custom_op.py
[torch.compile] Compile CustomOp.forward_native for SiluAndMul and QuantFP8 to avoid raw torch ops inside opaque custom ops (#32806)
2026-01-22 19:52:26 -08:00
parameter.py
[QeRL] Layerwise Reloading (#32133)
2026-01-30 08:50:05 -07:00
utils.py
[BugFix] Fix EPLB fail for MoeFP4 model with Marlin backend (#33262)
2026-01-29 16:52:11 +08:00
Powered by Gitea Version: 1.25.2 Page: 174ms Template: 4ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API