This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
You've already forked sglang
Code
Issues
Pull Requests
Actions
12
Packages
Projects
Releases
Wiki
Activity
Files
65dd08153dc87d7494f8702d3421f47666c876f9
sglang
/
benchmark
/
kernels
History
Mook
abc672e717
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
..
all_reduce
[AMD] Support --enable-aiter-allreduce-fusion on AMD GPUs (
#13747
)
2026-02-24 23:11:55 -08:00
decoding_attention_triton
Fix benchmark import for should_use_tensor_core (
#17232
)
2026-01-16 17:48:36 -05:00
deepep
…
deepseek
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
elementwise
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
flashinfer_allreduce_fusion
…
fused_moe_triton
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
quantization
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
scheduler_batch
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00
sliding_window_attention_triton
[Benchmark] use flashinfer bench_gpu_time instead of triton do_bench (
#20305
)
2026-03-12 04:04:30 +00:00