This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
You've already forked sglang
Code
Issues
Pull Requests
Actions
11
Packages
Projects
Releases
Wiki
Activity
Files
e7e89349c901129fd712f376cd44be0f0c402a92
sglang
/
benchmark
/
kernels
History
Shu Wang
6664083522
Replace [silu_and_mul_]scaled_fp4_group_quant by Flashinfer equivalent (
#12376
)
2025-11-13 00:26:00 -08:00
..
all_reduce
[AMD] Add AITER Custom All-Reduce (
#13102
)
2025-11-12 21:53:44 -08:00
decoding_attention_triton
…
deepep
fix(deepep): resolve benchmark failure on 4×IB-card setup by aligning tuning config with DeepEP commit bdd119f8 (
#11965
)
2025-10-22 21:20:54 -07:00
deepseek
Restruct sgl-kernel benchmark (
#10861
)
2025-09-25 07:45:25 +08:00
elementwise
[sgl-kernel] Optimize concat_mla_k kernel (
#10543
)
2025-09-28 23:04:22 +08:00
flashinfer_allreduce_fusion
[benchmark] add flashinfer_allreduce_fusion benchmark (
#9937
)
2025-09-03 16:31:01 +08:00
fused_moe_triton
fix tuning_fused_moe_triton_sep tool per_channel_quant bug (
#13027
)
2025-11-11 10:33:54 +08:00
minmax-text-01-lightning_attention
[lint] improve ruff check (
#11922
)
2025-10-22 11:32:50 +08:00
quantization
Replace [silu_and_mul_]scaled_fp4_group_quant by Flashinfer equivalent (
#12376
)
2025-11-13 00:26:00 -08:00
scheduler_batch
…
sliding_window_attention_triton
…