Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
Files
f8ca2368b20d2f7eb378dce7f2e0056beb144c4b
sglang/sgl-kernel/python/sgl_kernel
History
Hubert Lu af4b9bae95 [AMD] Add silu_and_mul, gelu_and_mul, gelu_tanh_and_mul, and gelu_quick kernels for AMD GPUs (#7135)
Co-authored-by: yiakwy-xpu-ml-framework-team <961186938@qq.com>
Co-authored-by: HAI <hixiao@gmail.com>
2025-07-24 23:44:28 -07:00
..
__init__.py
[AMD] Add silu_and_mul, gelu_and_mul, gelu_tanh_and_mul, and gelu_quick kernels for AMD GPUs (#7135)
2025-07-24 23:44:28 -07:00
allreduce.py
[Feature] Integrate quick allreduce and select the best allreduce implementation (#6619)
2025-07-24 20:48:42 -07:00
attention.py
…
cutlass_moe.py
…
elementwise.py
[AMD] Add silu_and_mul, gelu_and_mul, gelu_tanh_and_mul, and gelu_quick kernels for AMD GPUs (#7135)
2025-07-24 23:44:28 -07:00
flash_attn.py
…
fused_moe.py
[1/n] chore: decouple quantization implementation from vLLM dependency (#7992)
2025-07-16 15:56:26 -07:00
gemm.py
Add bf16 output option for dsv3_router_gemm kernel (#7999)
2025-07-20 09:49:37 +08:00
grammar.py
…
kvcacheio.py
breakdown kernel update (#8334)
2025-07-25 08:33:17 +08:00
marlin.py
…
moe.py
…
sampling.py
…
sparse_flash_attn.py
…
spatial.py
[Feature] CUDA Green Context Support (#7649)
2025-07-15 02:49:16 +08:00
speculative.py
Add treemask mode to build_eagle_tree & release sgl-kernel 0.2.3 (#7756)
2025-07-05 12:17:05 -07:00
top_k.py
…
utils.py
…
version.py
fix: workaround for deepgemm warmup issue (#8302)
2025-07-23 12:01:51 -07:00
Powered by Gitea Version: 1.25.5 Page: 2031ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API