This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
13
Packages
Projects
Releases
Wiki
Activity
Files
e6d59884426f381be9c91df35409f091868c4c34
sglang
/
sgl-kernel
/
tests
T
History
ykcombat
1ebec1a8b0
[Feature] CUDA Green Context Support (
#7649
)
2025-07-15 02:49:16 +08:00
..
spatial
…
speculative
…
test_activation.py
…
test_apply_token_bitmask_inplace.py
…
test_awq_dequant.py
Fix AWQ Dequant and Weight Loading of deepseek v2 (
#6842
)
2025-06-17 13:45:10 -07:00
test_bmm_fp8.py
…
test_custom_allreduce.py
sgl-kernel transfer custom allreduce from trt kernel to vllm kernel (
#5079
)
2025-04-05 14:23:20 -07:00
test_cutlass_mla.py
[fix] fix cutlass_mla_backend with cuda_graph and add sm_scale for sgl-kernel cutlass_mla (
#7184
)
2025-06-14 12:45:41 -07:00
test_cutlass_w4a8_moe_mm.py
…
test_dsv3_fused_a_gemm.py
…
test_dsv3_router_gemm.py
…
test_ep_moe_post_reorder_kernel.py
…
test_ep_moe_pre_reorder_kernel.py
…
test_ep_moe_silu_and_mul_kernel.py
…
test_flash_attention.py
…
test_fp4_gemm.py
…
test_fp4_quantize.py
…
test_fp8_blockwise_gemm.py
…
test_fp8_blockwise_moe.py
…
test_fp8_gemm.py
…
test_int8_gemm.py
…
test_kvcacheio.py
…
test_lightning_attention_decode.py
…
test_marlin_repack.py
[1/n] apply wna16marlin kernel in moe weight only quantization (
#7683
)
2025-07-01 23:21:25 -07:00
test_merge_state_v2.py
…
test_merge_state.py
…
test_moe_align.py
…
test_moe_fused_gate.py
…
test_moe_topk_softmax.py
…
test_mscclpp.py
…
test_norm.py
…
test_per_tensor_quant_fp8.py
…
test_per_token_group_quant_8bit.py
…
test_per_token_quant_fp8.py
…
test_qserve_w4a8_per_chn_gemm.py
…
test_qserve_w4a8_per_group_gemm.py
…
test_rotary_embedding.py
[Misc] Clean sgl-kernel test (
#5216
)
2025-04-10 11:28:41 -07:00
test_sampling.py
…
test_sparse_flash_attn.py
…