Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 11 Packages Projects Releases Wiki Activity
Files
f55933e1cc50263e0bcde65c3b78969b56225c7f
sglang/sgl-kernel/benchmark
History
Zhaoyi Li 3c9740d200 update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
..
bench_awq_dequant.py
Add awq dequantize kernel to sgl with 1x to 3x speedup (#4104)
2025-03-12 00:10:02 -07:00
bench_fp8_blockwise_gemm.py
sgl-kernel use cutlass latest version for fp8 blockwise gemm (#5207)
2025-04-09 11:47:04 -07:00
bench_fp8_gemm.py
…
bench_int8_gemm.py
…
bench_lightning_attention_decode.py
[Fix] use torch.cat instead of torch.concat to prevent entering the Autograd backends. (#4466)
2025-03-16 00:02:47 -07:00
bench_moe_align_block_size.py
reduce moe_align_block_size_kernel small batch mode overhead (#5086)
2025-04-09 17:59:35 -07:00
bench_moe_fused_gate.py
Add deepseek style fused moe group gate selection kernel (#4530)
2025-03-29 11:51:45 -07:00
bench_moe_topk_softmax.py
Add moe topk softmax templated from vllm (#4302)
2025-03-14 12:03:33 -07:00
bench_per_tensor_quant_fp8.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
bench_per_token_group_quant_8bit.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
bench_per_token_quant_fp8.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
Powered by Gitea Version: 1.25.5 Page: 5944ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API