This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
1,312
Commits
2
Branches
140
Tags
51d4094fda63b1d738f55ae9dd75d354b9c1143c
Commit Graph
3 Commits
Author
SHA1
Message
Date
alexm-nm
e288df0632
[Bugfix] Fine-tune gptq_marlin configs to be more similar to marlin (
#4626
)
2024-05-08 17:14:31 -07:00
alexm-nm
7038e8b803
[Kernel] Support running GPTQ 8-bit models in Marlin (
#4533
)
2024-05-02 12:56:22 -04:00
Robert Shaw
73c8d677e5
[Kernel] Marlin Expansion: Support AutoGPTQ Models with Marlin (
#3922
)
...
Co-authored-by: alexm <
alexm@neuralmagic.com
> Co-authored-by: mgoin <
michael@neuralmagic.com
>
2024-04-29 09:35:34 -07:00