Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 15 Packages Projects Releases Wiki Activity
Files
a9499885e97e77b8b0a57642423c6e2c1a6fcaa8
sglang/sgl-kernel/benchmark
T
History
Zhaoyi Li 3c9740d200 update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
..
bench_awq_dequant.py
Add awq dequantize kernel to sgl with 1x to 3x speedup (#4104)
2025-03-12 00:10:02 -07:00
bench_fp8_blockwise_gemm.py
sgl-kernel use cutlass latest version for fp8 blockwise gemm (#5207)
2025-04-09 11:47:04 -07:00
bench_fp8_gemm.py
…
bench_int8_gemm.py
Add shapes for int8 gemm benchmark (#3093)
2025-01-24 12:27:30 +08:00
bench_lightning_attention_decode.py
[Fix] use torch.cat instead of torch.concat to prevent entering the Autograd backends. (#4466)
2025-03-16 00:02:47 -07:00
bench_moe_align_block_size.py
reduce moe_align_block_size_kernel small batch mode overhead (#5086)
2025-04-09 17:59:35 -07:00
bench_moe_fused_gate.py
Add deepseek style fused moe group gate selection kernel (#4530)
2025-03-29 11:51:45 -07:00
bench_moe_topk_softmax.py
Add moe topk softmax templated from vllm (#4302)
2025-03-14 12:03:33 -07:00
bench_per_tensor_quant_fp8.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
bench_per_token_group_quant_8bit.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
bench_per_token_quant_fp8.py
update variable naming and comments for rocm (#5299)
2025-04-11 23:15:05 -07:00
Powered by Gitea Version: 1.27.2 Page: 4118ms Template: 6ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API