This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
11
Packages
Projects
Releases
Wiki
Activity
Files
1b427dae0269024ec7c7330bc7b5e181b557d342
sglang
/
sgl-kernel
/
python
/
sgl_kernel
T
History
Yineng Zhang
f98e88b9fb
chore: bump sgl-kernel v0.2.6 (
#8165
)
2025-07-19 00:56:18 -07:00
..
__init__.py
[Feature] CUDA Green Context Support (
#7649
)
2025-07-15 02:49:16 +08:00
allreduce.py
…
attention.py
…
cutlass_moe.py
[1/n]: add cutlass W4A8 moe kernel for hopper architecture (
#7772
)
2025-07-04 20:50:12 -07:00
elementwise.py
…
flash_attn.py
…
fused_moe.py
[1/n] chore: decouple quantization implementation from vLLM dependency (
#7992
)
2025-07-16 15:56:26 -07:00
gemm.py
…
grammar.py
…
kvcacheio.py
…
marlin.py
[1/n] apply wna16marlin kernel in moe weight only quantization (
#7683
)
2025-07-01 23:21:25 -07:00
moe.py
[optimize] fuse renormalize into moe_topk_softmax (
#7744
)
2025-07-03 12:42:44 -07:00
sampling.py
…
sparse_flash_attn.py
…
spatial.py
[Feature] CUDA Green Context Support (
#7649
)
2025-07-15 02:49:16 +08:00
speculative.py
Add treemask mode to build_eagle_tree & release sgl-kernel 0.2.3 (
#7756
)
2025-07-05 12:17:05 -07:00
top_k.py
…
utils.py
…
version.py
chore: bump sgl-kernel v0.2.6 (
#8165
)
2025-07-19 00:56:18 -07:00