This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
You've already forked sglang
Code
Issues
Pull Requests
Actions
12
Packages
Projects
Releases
Wiki
Activity
Files
90dfe3de4c1ce54c5ee11f4ba0c63b74c0bfeca4
sglang
/
benchmark
/
kernels
History
Xiaoyu Zhang
b1fb7e458c
[benchmark] add flashinfer_allreduce_fusion benchmark (
#9937
)
2025-09-03 16:31:01 +08:00
..
all_reduce
support 1 shot allreduce in 1-node and 2-node using mscclpp (
#6277
)
2025-06-04 22:11:24 -07:00
decoding_attention_triton
…
deepep
Support tuning DeepEP configs (
#6742
)
2025-05-29 08:12:22 -07:00
deepseek
…
fbgemm
[benchmark] add flashinfer_allreduce_fusion benchmark (
#9937
)
2025-09-03 16:31:01 +08:00
flashinfer_allreduce_fusion
[benchmark] add flashinfer_allreduce_fusion benchmark (
#9937
)
2025-09-03 16:31:01 +08:00
fused_moe_triton
Add A100 fused MoE kernel configs for Dpsk (
#9677
)
2025-08-26 20:49:48 -07:00
minmax-text-01-lightning_attention
…
quantization
[NVIDIA] [2/N] Optimize
silu_and_mul_scaled_fp4_grouped_quant
perf (
#9556
)
2025-08-29 17:17:03 -07:00
rmsnorm
…
scheduler_batch
[test] add ut and bm for get_last_loc (
#6746
)
2025-05-29 11:47:21 -07:00
sliding_window_attention_triton
Optimize triton swa kernel by skipping computation (
#8860
)
2025-08-06 21:37:50 +08:00