This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
f9582fd8f4f0e03a1b0d5b959bbf15d297336775
vllm
/
vllm
/
model_executor
History
Eugene Khvedchenya
f9582fd8f4
[Model] Allow passing custom number of max tiles to Nano 2 VL (
#26403
)
...
Signed-off-by: Eugene Khvedchenia <
ekhvedchenia@nvidia.com
>
2025-10-08 11:19:39 +00:00
..
layers
Add SwigluOAI implementation for CPUFusedMOE (
#26347
)
2025-10-07 20:17:49 -06:00
model_loader
[TPU] Rename tpu_commons to tpu_inference (
#26279
)
2025-10-07 23:30:52 -07:00
models
[Model] Allow passing custom number of max tiles to Nano 2 VL (
#26403
)
2025-10-08 11:19:39 +00:00
warmup
Convert formatting to use
ruff
instead of
yapf
+
isort
(
#26247
)
2025-10-05 07:06:22 -07:00
__init__.py
Convert formatting to use
ruff
instead of
yapf
+
isort
(
#26247
)
2025-10-05 07:06:22 -07:00
custom_op.py
[Frontend] CompilationConfig overhaul (
#20283
): deprecate use_inductor in favor of backend, simplify custom_ops (
#26113
)
2025-10-07 12:53:43 -07:00
parameter.py
Convert formatting to use
ruff
instead of
yapf
+
isort
(
#26247
)
2025-10-05 07:06:22 -07:00
utils.py
Convert formatting to use
ruff
instead of
yapf
+
isort
(
#26247
)
2025-10-05 07:06:22 -07:00