This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
e6680f9e25a433bcd754181705e72034ce6c470c
vllm
/
vllm
/
utils
History
wuhang
e6680f9e25
[Bugfix] Add log prefix in non-dp mode engine core (
#21889
)
...
Signed-off-by: wuhang <
wuhang6@huawei.com
>
2025-08-01 09:04:16 +00:00
..
__init__.py
[Bugfix] Add log prefix in non-dp mode engine core (
#21889
)
2025-08-01 09:04:16 +00:00
deep_gemm.py
[Refactor] Remove Duplicate
per_block_cast_to_fp8
, Remove Dependencies of DeepGEMM (
#21787
)
2025-08-01 01:13:27 +00:00
flashinfer.py
[NVIDIA] Add SM100 Flashinfer MoE per tensor scale fp8 backend (
#21458
)
2025-07-31 06:00:01 -07:00
tensor_schema.py
Migrate FuyuImagePatchInputs to TensorSchema (
#21662
)
2025-07-26 19:34:14 -07:00