Logo
Explore Help
Register Sign In
biondizzle/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions 2 Packages Projects Releases Wiki Activity
Files
f990bab2a4198c4de6b5b349d35fc74bf0f36f3e
vllm/csrc/core
History
Lucas Wilkinson aeb37c2a72 [CI/Build] Per file CUDA Archs (improve wheel size and dev build times) (#8845)
2024-10-03 22:55:25 -04:00
..
exception.hpp
[Bugfix] Fix Marlin MoE act order when is_k_full == False (#8741)
2024-09-28 18:19:40 -07:00
registration.h
[CI/Build] Per file CUDA Archs (improve wheel size and dev build times) (#8845)
2024-10-03 22:55:25 -04:00
scalar_type.hpp
[Bugfix] Allow ScalarType to be compiled with pytorch 2.3 and add checks for registering FakeScalarType and dynamo support. (#7886)
2024-08-27 23:13:45 -04:00
torch_bindings.cpp
[Misc] Disambiguate quantized types via a new ScalarType (#6396)
2024-08-02 13:51:58 -07:00
Powered by Gitea Version: 1.25.2 Page: 5670ms Template: 1ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API