Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
a3c226e7eb19b976a937e745f3867eb05f809278
vllm/vllm/executor
History
bigPYJ1151 0e3f06fe9c [Hardware][Intel] Add CPU inference backend (#3634)
Co-authored-by: Kunshang Ji <kunshang.ji@intel.com>
Co-authored-by: Yuan Zhou <yuan.zhou@intel.com>
2024-04-01 22:07:30 -07:00
..
__init__.py
Add distributed model executor abstraction (#3191)
2024-03-11 11:03:45 -07:00
cpu_executor.py
[Hardware][Intel] Add CPU inference backend (#3634)
2024-04-01 22:07:30 -07:00
executor_base.py
[Feature] Add vision language model support. (#3042)
2024-03-25 14:16:30 -07:00
gpu_executor.py
[Core][Bugfix]Refactor block manager for better testability (#3492)
2024-03-27 23:59:28 -07:00
neuron_executor.py
[Bugfix] Update neuron_executor.py to add optional vision_language_config (#3695)
2024-03-28 10:43:34 -07:00
ray_gpu_executor.py
[Core] Support multi-node inference(eager and cuda graph) (#3686)
2024-03-28 15:01:55 -07:00
utils.py
Add distributed model executor abstraction (#3191)
2024-03-11 11:03:45 -07:00
Powered by Gitea Version: 1.25.2 Page: 57ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API