This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
12
Packages
Projects
Releases
Wiki
Activity
Files
121f92c58309b9f57177eaefe32955e35a78c8bb
sglang
/
sgl-kernel
/
python
/
sgl_kernel
T
History
3 people
HandH1998
yych0745
sleepcoo
4d643f6c7a
[1/2] Support Qserve (
#6457
)
...
Co-authored-by: yych0745 <
1398089567@qq.com
> Co-authored-by: sleepcoo <
sleepcoo@gmail.com
>
2025-05-21 19:48:59 -07:00
..
__init__.py
[1/2] Support Qserve (
#6457
)
2025-05-21 19:48:59 -07:00
allreduce.py
sgl-kernel transfer custom allreduce from trt kernel to vllm kernel (
#5079
)
2025-04-05 14:23:20 -07:00
attention.py
Add Cutlass MLA attention backend (
#5390
)
2025-04-27 20:58:53 -07:00
elementwise.py
Add typo checker in pre-commit (
#6179
)
2025-05-11 12:55:00 +08:00
flash_attn.py
Revert "fix some typos" (
#6244
)
2025-05-12 12:53:26 -07:00
gemm.py
[1/2] Support Qserve (
#6457
)
2025-05-21 19:48:59 -07:00
grammar.py
fix sgl-kernel unit tests (
#5666
)
2025-04-23 01:18:30 -07:00
moe.py
[2/2] Add python wrapper for CUTLASS FP8 Blockscale MoE Kernel. (
#5694
)
2025-05-16 13:14:07 -07:00
sampling.py
Fix sampler nan check when calling top_k_top_p_sampling_from_probs (
#5546
)
2025-04-19 21:47:23 -07:00
sparse_flash_attn.py
[Feat] QWen-1M context support[1/2]: Update block sparse attention backend utils kernel (
#5847
)
2025-04-28 11:03:17 -07:00
speculative.py
use default for torch.ops (
#4835
)
2025-03-27 19:09:58 -07:00
utils.py
…
version.py
chore: bump sgl-kernel v0.1.3 (
#6368
)
2025-05-17 00:15:55 -07:00