Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 14 Packages Projects Releases Wiki Activity
Files
bd95944cf6f02e8cbf6db373a64b62ca5eed3384
sglang/sgl-kernel/csrc/moe
T
History
Yuan Luoluoyuan.luoXiaoyu Zhang
616a3e20df [sgl-kernel] Support moe_sum_reduce cuda kernel (#10321)
Co-authored-by: luoyuan.luo <luoyuan.luo@antgroup.com>
Co-authored-by: Xiaoyu Zhang <35585791+BBuf@users.noreply.github.com>
2025-09-19 14:12:09 +08:00
..
cutlass_moe/w4a8
pass a_scale from fp8 quant result instead of hard code to 1.0f (#10241)
2025-09-10 12:56:05 -07:00
marlin_moe_wna16
Support compile sgl-kernel on cuda 13.0 (#9721)
2025-08-28 10:18:03 -07:00
cutlass_moe_helper.cu
…
fp8_blockwise_moe_kernel.cu
Update CUTLASS. Refine KernelSchedule for fp8 (grouped) gemm. (#10491)
2025-09-16 02:47:37 -07:00
moe_align_kernel.cu
…
moe_fused_gate.cu
Fix correction bias undefined behavior for nvfp4 models (#10426)
2025-09-14 18:41:09 -07:00
moe_sum_reduce.cu
[sgl-kernel] Support moe_sum_reduce cuda kernel (#10321)
2025-09-19 14:12:09 +08:00
moe_topk_softmax_kernels.cu
Support compile sgl-kernel on cuda 13.0 (#9721)
2025-08-28 10:18:03 -07:00
nvfp4_blockwise_moe.cu
…
prepare_moe_input.cu
fix: fix apply_shuffle_mul_sum (#7444)
2025-07-04 23:23:30 -07:00
Powered by Gitea Version: 1.27.2 Page: 1404ms Template: 42ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API