Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
Files
a68cb201dd5f4ae6155b324d22054bbb0de15fba
sglang/python/sglang/srt/layers
History
Ke Bao a68cb201dd Fix triton head num (#1482)
2024-09-21 10:25:20 +08:00
..
fused_moe
[Feature, Hardware] Enable SGLang on AMD GPUs via PyTorch for ROCm (#1420)
2024-09-17 07:43:52 +00:00
quantization
feat: update linear deps 1/N (#1305)
2024-09-19 20:53:11 +08:00
triton_attention
Enable torch.compile for triton backend (#1422)
2024-09-14 15:38:37 -07:00
activation.py
feat: update linear deps 1/N (#1305)
2024-09-19 20:53:11 +08:00
attention_backend.py
Fix triton head num (#1482)
2024-09-21 10:25:20 +08:00
flashinfer_utils.py
Support cuda graph in the triton attention backend (#1401)
2024-09-12 00:36:55 -07:00
layernorm.py
[Bugfix] Enable SGLang on AMD GPUs via PyTorch for ROCm (#1419) (#1453)
2024-09-18 02:01:35 -07:00
linear.py
feat: update linear deps 1/N (#1305)
2024-09-19 20:53:11 +08:00
logits_processor.py
[Fix] Fix logprob and normalized_logprob (#1428)
2024-09-15 06:36:06 -07:00
pooler.py
Add e5-mistral modules [unreachable code] - step 1/3 (#983)
2024-08-08 00:04:15 -07:00
radix_attention.py
Refactor attention backend (#1381)
2024-09-11 11:44:26 -07:00
sampler.py
Fuse top_k and top_k in the sampler (#1457)
2024-09-18 04:35:35 -07:00
torchao_utils.py
Add torchao quant for mixtral and qwen_moe (#1418)
2024-09-14 06:46:55 +00:00
Powered by Gitea Version: 1.25.5 Page: 70ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API