This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
11
Packages
Projects
Releases
Wiki
Activity
Files
5d299c25c04a9e28170659b7588afcc0b6f84599
sglang
/
sgl-kernel
/
csrc
T
History
Serge Panev
and
Fan Yin
e95668abc7
[NVIDIA] Fix CUDA arch requirement in nvfp4 cast (
#12581
)
...
Signed-off-by: Serge Panev <
spanev@nvidia.com
> Co-authored-by: Fan Yin <
1106310035@qq.com
>
2026-01-21 20:21:11 -08:00
..
allreduce
[amd] Add deterministic all-reduce kernel for AMD (ROCm) (
#15340
)
2025-12-18 23:36:03 -08:00
attention
[sgl-kernel Code Clean] Remove useless lightning_attention kernel (
#13819
)
2025-11-24 18:26:25 +08:00
cpu
[CPU] Add 4D input support for ROPE in sgl-kernel (
#9337
)
2025-12-16 17:27:39 +08:00
cutlass_extensions
…
elementwise
[AMD] Support fast_topk kernels in sgl-kernel (
#15172
)
2025-12-19 22:19:09 -08:00
expert_specialization
[sgl-kernel][Feat][B200][1/N] Support MXFP8 Grouped GEMM in Blackwell (
#13731
)
2025-12-04 10:09:09 +08:00
gemm
[NVIDIA] Fix CUDA arch requirement in nvfp4 cast (
#12581
)
2026-01-21 20:21:11 -08:00
grammar
[AMD] Expand test coverage for AMD CI and enable apply_token_bitmask_inplace_cuda in sgl-kernel (
#8268
)
2025-08-15 12:32:51 -07:00
kvcacheio
…
mamba
Update GDN causal conv1d cuda kernel - prepare for new changes (
#13188
)
2025-11-13 14:09:47 -08:00
memory
…
moe
[perf]optimize w4afp8 kernel on deepseek-v3-0324 (
#12921
)
2025-12-18 18:13:22 +08:00
quantization
/gguf
…
sgl_diffusion
/elementwise
Revert "[Diffusion] Move diffusion time embedding to jit kernel" (
#17257
)
2026-01-17 21:29:22 +08:00
spatial
…
speculative
…
common_extension_rocm.cc
[Piecewise] Support PCG weak_ref_tensor cuda kernel on AMD (
#17291
)
2026-01-20 14:05:32 -08:00
common_extension.cc
Revert "[Diffusion] Move diffusion time embedding to jit kernel" (
#17257
)
2026-01-17 21:29:22 +08:00
flash_extension.cc
…
flashmla_extension.cc
…
spatial_extension.cc
…