This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
You've already forked sglang
Code
Issues
Pull Requests
Actions
11
Packages
Projects
Releases
Wiki
Activity
Files
ca5f2e2ed13c03e5e6f76b4c4e5428a45839268d
sglang
/
benchmark
/
kernels
History
RoyWang
a1ef8e2cc0
[AMD] optimize Kimi K2.5 fused_moe_triton performance by tuning (
#19228
)
2026-02-26 11:50:13 -08:00
..
all_reduce
[AMD] Support --enable-aiter-allreduce-fusion on AMD GPUs (
#13747
)
2026-02-24 23:11:55 -08:00
decoding_attention_triton
Fix benchmark import for should_use_tensor_core (
#17232
)
2026-01-16 17:48:36 -05:00
deepep
fix(deepep): resolve benchmark failure on 4×IB-card setup by aligning tuning config with DeepEP commit bdd119f8 (
#11965
)
2025-10-22 21:20:54 -07:00
deepseek
[NVIDIA] Add fp8 gemm benchmark on blackwell (
#13528
)
2025-11-19 19:35:00 -08:00
elementwise
…
flashinfer_allreduce_fusion
…
fused_moe_triton
[AMD] optimize Kimi K2.5 fused_moe_triton performance by tuning (
#19228
)
2026-02-26 11:50:13 -08:00
quantization
Refactor tuning block wise kernel and opt Qwen/Qwen3-VL-32B-Instruct-FP8 (
#14141
)
2025-12-08 09:24:58 +08:00
scheduler_batch
…
sliding_window_attention_triton
…