Commit Graph

5 Commits

Author SHA1 Message Date
Woosuk Kwon
e070829ae8 Support bfloat16 data type (#54) 2023-05-03 14:09:44 -07:00
Zhuohan Li
27f1410d06 New weight loader without np copy (#52) 2023-05-03 15:32:04 +08:00
Zhuohan Li
721fa3df15 FastAPI-based working frontend (#10) 2023-03-29 14:48:56 +08:00
Zhuohan Li
2f49f15585 Support tensor parallel (#2) 2023-03-21 13:45:42 -07:00
Woosuk Kwon
e9d3f2ff77 Add memory analyzer & utomatically configure KV cache size (#6) 2023-03-11 23:23:14 -08:00