This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
a3c226e7eb19b976a937e745f3867eb05f809278
vllm
/
vllm
/
executor
History
bigPYJ1151
0e3f06fe9c
[Hardware][Intel] Add CPU inference backend (
#3634
)
...
Co-authored-by: Kunshang Ji <
kunshang.ji@intel.com
> Co-authored-by: Yuan Zhou <
yuan.zhou@intel.com
>
2024-04-01 22:07:30 -07:00
..
__init__.py
Add distributed model executor abstraction (
#3191
)
2024-03-11 11:03:45 -07:00
cpu_executor.py
[Hardware][Intel] Add CPU inference backend (
#3634
)
2024-04-01 22:07:30 -07:00
executor_base.py
[Feature] Add vision language model support. (
#3042
)
2024-03-25 14:16:30 -07:00
gpu_executor.py
[Core][Bugfix]Refactor block manager for better testability (
#3492
)
2024-03-27 23:59:28 -07:00
neuron_executor.py
[Bugfix] Update neuron_executor.py to add optional vision_language_config (
#3695
)
2024-03-28 10:43:34 -07:00
ray_gpu_executor.py
[Core] Support multi-node inference(eager and cuda graph) (
#3686
)
2024-03-28 15:01:55 -07:00
utils.py
Add distributed model executor abstraction (
#3191
)
2024-03-11 11:03:45 -07:00