Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 14 Packages Projects Releases Wiki Activity
Files
0ca3e56802046c1e8c7f2af64d2d64a484a89ccd
sglang/sgl-kernel/python/sgl_kernel
T
History
Yineng Zhang d71f3f0a2a chore: bump sgl-kernel v0.1.4 (#6522)
2025-05-22 09:47:42 -07:00
..
__init__.py
[1/2] Support Qserve (#6457)
2025-05-21 19:48:59 -07:00
allreduce.py
sgl-kernel transfer custom allreduce from trt kernel to vllm kernel (#5079)
2025-04-05 14:23:20 -07:00
attention.py
Add Cutlass MLA attention backend (#5390)
2025-04-27 20:58:53 -07:00
elementwise.py
Add typo checker in pre-commit (#6179)
2025-05-11 12:55:00 +08:00
flash_attn.py
Revert "fix some typos" (#6244)
2025-05-12 12:53:26 -07:00
gemm.py
[1/2] Support Qserve (#6457)
2025-05-21 19:48:59 -07:00
grammar.py
fix sgl-kernel unit tests (#5666)
2025-04-23 01:18:30 -07:00
moe.py
[2/2] Add python wrapper for CUTLASS FP8 Blockscale MoE Kernel. (#5694)
2025-05-16 13:14:07 -07:00
sampling.py
Fix sampler nan check when calling top_k_top_p_sampling_from_probs (#5546)
2025-04-19 21:47:23 -07:00
sparse_flash_attn.py
[Feat] QWen-1M context support[1/2]: Update block sparse attention backend utils kernel (#5847)
2025-04-28 11:03:17 -07:00
speculative.py
use default for torch.ops (#4835)
2025-03-27 19:09:58 -07:00
utils.py
…
version.py
chore: bump sgl-kernel v0.1.4 (#6522)
2025-05-22 09:47:42 -07:00
Powered by Gitea Version: 1.27.2 Page: 1310ms Template: 3ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API