Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 13 Packages Projects Releases Wiki Activity
Files
9c0b1eb5adb5b37f6bc99658c042b7499fd5510d
sglang/sgl-kernel/csrc/moe
T
History
PGFLMG 8fdcd98efe [7/n] decouple quantization impl from vllm dependency - gguf kernel (#11019)
2025-10-11 14:04:57 -07:00
..
cutlass_moe/w4a8
pass a_scale from fp8 quant result instead of hard code to 1.0f (#10241)
2025-09-10 12:56:05 -07:00
marlin_moe_wna16
Support compile sgl-kernel on cuda 13.0 (#9721)
2025-08-28 10:18:03 -07:00
cutlass_moe_helper.cu
[Fix]Fix index oob in get_group_gemm_starts kernel. (#8564)
2025-07-30 19:49:35 -07:00
fp8_blockwise_moe_kernel.cu
Update CUTLASS. Refine KernelSchedule for fp8 (grouped) gemm. (#10491)
2025-09-16 02:47:37 -07:00
moe_align_kernel.cu
[AMD] Reorganize hip-related header files in sgl-kernel (#9320)
2025-08-18 16:53:44 -07:00
moe_fused_gate.cu
Fix correction bias undefined behavior for nvfp4 models (#10426)
2025-09-14 18:41:09 -07:00
moe_sum_reduce.cu
[sgl-kernel] Support float64 moe_sum_reduce cuda kernel (#11068)
2025-10-07 14:31:11 +00:00
moe_sum.cu
[7/n] decouple quantization impl from vllm dependency - gguf kernel (#11019)
2025-10-11 14:04:57 -07:00
moe_topk_softmax_kernels.cu
Support compile sgl-kernel on cuda 13.0 (#9721)
2025-08-28 10:18:03 -07:00
nvfp4_blockwise_moe.cu
…
prepare_moe_input.cu
…
Powered by Gitea Version: 1.27.2 Page: 1404ms Template: 4ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API