Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 11 Packages Projects Releases Wiki Activity
9,885 Commits 1 Branch 0 Tags
cc63c99f112f781605017db018fced686cdc94de
Commit Graph
9 Commits
This Branch
This Branch
All Branches
Author SHA1 Message Date
Jerry Zhang feb2b768ba Add integration with gemlite weight only quant (#2528) 2024-12-21 00:25:25 +08:00
Jerry Zhang 82699474fd Small fixes for torchao quant (#2476) 2024-12-16 14:08:12 -08:00
Jerry Zhang 9cc733b38c move apply_torchao_config_ to model_runner (#2342) 2024-12-04 17:26:42 -08:00
Jerry Zhang 7f8fcd39cd Turn off autotune for scaled mm for fp8 dynamic quant in torchao (#2116) 2024-11-21 12:19:49 -08:00
Jerry Zhang 5c6a41facf Error out when torchao-config option is not recognized (#2107) 2024-11-20 17:37:28 -08:00
Jerry Zhang 9b0926ceeb Add llama implementation with no tensor parallel linears (#1561) 2024-10-05 11:22:27 -07:00
Jerry Zhang 63e845d0bb Add float8 dynamic quant to torchao_utils (#1528) 2024-09-28 12:27:54 -07:00
Jerry Zhang 30b404ce72 Add torchao quant for mixtral and qwen_moe (#1418) 2024-09-14 06:46:55 +00:00
Jerry ZhangandLianmin Zheng a7c47e0f02 Add torchao quant (int4/int8/fp8) to llama models (#1341)
Co-authored-by: Lianmin Zheng <lianminzheng@gmail.com>
2024-09-09 05:32:41 -07:00
Powered by Gitea Version: 1.27.2 Page: 778ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API