Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 11 Packages Projects Releases Wiki Activity
Files
993ec178ef2c9efa5b4e87a9ebeee179ac1a51a3
sglang/sgl-kernel/csrc/elementwise
History
Yifan Cui 45fe51a28e Reduce topk kernel shared memory from 128KB to 32KB for better occupancy (#17747)
Co-authored-by: Claude <noreply@anthropic.com>
2026-01-30 21:42:21 -08:00
..
activation.cu
…
cast.cu
move all get_stream in sgl_kernel to c++ to reduce the launch overhead (#12521)
2025-11-02 13:15:05 -08:00
concat_mla.cu
[Fix] concat_mla_absorb_q_kernel fails for long inputs (#12453)
2025-11-02 11:52:06 -08:00
copy.cu
Support copying tensor from cpu to gpu without using copy engines (#10007)
2025-09-05 20:07:19 +08:00
fused_add_rms_norm_kernel.cu
…
pos_enc.cu
[1/2] Add rope kernel in sgl-kernel (#14334)
2025-12-04 16:45:44 +08:00
pos_enc.cuh
Add PDL support for quant kernel and rope kernel (#9106)
2025-08-20 01:56:29 -07:00
rope.cu
Add FP32 dtype support for RoPE - Part1 (#13181)
2025-11-15 11:37:18 -08:00
topk.cu
Reduce topk kernel shared memory from 128KB to 32KB for better occupancy (#17747)
2026-01-30 21:42:21 -08:00
utils.cuh
[sgl-kernel] Optimize concat_mla_k kernel (#10543)
2025-09-28 23:04:22 +08:00
Powered by Gitea Version: 1.25.5 Page: 2060ms Template: 10ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API