This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
13
Packages
Projects
Releases
Wiki
Activity
Files
8ad700f735c48bb9627f1d8ea21d5d44561777f5
sglang
/
test
/
srt
T
History
…
..
ascend
…
configs
…
cpu
…
entrypoints
/http_server
…
ep
…
hicache
…
lora
…
models
…
openai_server
…
quant
…
rl
…
double-sparsity-config-Llama-3.1-8B-Instruct.json
…
experiment_runner.py
…
kv_cache_scales_llama3_1_8b.json
…
kv_cache_scales_llama3_8b.json
…
kv_cache_scales_qwen2_1_5b.json
…
parse_results.py
…
run_suite.py
…
test_abort.py
Support updating weights at once by stopping all requests (
#6698
)
2025-07-02 22:26:06 -07:00
test_bench_one_batch.py
…
test_bench_serving.py
…
test_bnb.py
…
test_chunked_prefill.py
…
test_cpp_radix_cache.py
…
test_cpu_graph.py
…
test_create_kvindices.py
…
test_custom_allreduce.py
[AMD] Add unit-test-sgl-kernel-amd to AMD CI (
#7539
)
2025-06-29 15:50:09 -07:00
test_data_parallelism.py
…
test_disaggregation_different_tp.py
…
test_disaggregation_pp.py
…
test_disaggregation.py
…
test_double_sparsity.py
…
test_dp_attention.py
…
test_eagle_infer_a.py
…
test_eagle_infer_b.py
…
test_ebnf_constrained.py
…
test_eval_accuracy_large.py
…
test_eval_fp8_accuracy.py
…
test_expert_distribution.py
…
test_expert_location_updater.py
…
test_fa3.py
…
test_fim_completion.py
…
test_flashmla.py
…
test_forward_split_prefill.py
…
test_full_deepseek_v3.py
…
test_function_call_parser.py
…
test_fused_moe.py
…
test_get_weights_by_name.py
…
test_gguf.py
…
test_gpt_oss_1gpu.py
…
test_gpt_oss_4gpu.py
feat: add gpt oss b200 ci (
#9988
)
2025-09-03 17:26:38 -07:00
test_gpt_oss_common.py
…
test_gptqmodel_dynamic.py
…
test_harmony_parser.py
…
test_health_check.py
…
test_hidden_states.py
…
test_hybrid_attn_backend.py
…
test_input_embeddings.py
Add retry for flaky tests in CI (
#4755
)
2025-03-25 16:53:12 -07:00
test_intel_amx_attention_backend.py
…
test_io_struct.py
…
test_jinja_template_utils.py
…
test_kv_events.py
…
test_local_attn.py
…
test_metrics_utils.py
…
test_metrics.py
…
test_mla_deepseek_v3.py
…
test_mla_flashinfer.py
…
test_mla_fp8.py
…
test_mla_int8_deepseek_v3.py
…
test_mla_tp.py
…
test_mla.py
…
test_modelopt_fp8kvcache.py
…
test_modelopt.py
…
test_models_from_modelscope.py
…
test_moe_eval_accuracy_large.py
…
test_mscclpp.py
…
test_multi_instance_release_memory_occupation.py
…
test_multi_tokenizer.py
…
test_nightly_gsm8k_eval_amd.py
…
test_nightly_gsm8k_eval.py
…
test_no_chunked_prefill.py
…
test_no_overlap_scheduler.py
…
test_original_logprobs.py
…
test_page_size.py
…
test_patch_torch.py
…
test_penalty.py
…
test_pp_single_node.py
…
test_pytorch_sampling_backend.py
…
test_quick_allreduce.py
…
test_radix_attention.py
…
test_reasoning_parser.py
…
test_regex_constrained.py
…
test_release_memory_occupation.py
…
test_request_queue_validation.py
…
test_retract_decode.py
…
test_rope_rocm.py
…
test_sagemaker_server.py
…
test_schedule_policy.py
…
test_score_api.py
…
test_server_args.py
…
test_session_control.py
…
test_skip_tokenizer_init.py
…
test_srt_endpoint.py
[Feature] Add Logit Bias (
#6579
)
2025-06-10 15:39:25 -07:00
test_srt_engine_with_quant_args.py
…
test_srt_engine.py
chore: upgrade flashinfer v0.2.6.post1 jit (
#6958
)
2025-06-09 09:22:39 -07:00
test_standalone_speculative_decoding.py
…
test_start_profile.py
…
test_swa_unittest.py
…
test_tokenizer_batch_encode.py
…
test_torch_compile_moe.py
…
test_torch_compile.py
Improve profiler and integrate profiler in bench_one_batch_server (
#6787
)
2025-05-31 15:53:55 -07:00
test_torch_native_attention_backend.py
…
test_torch_tp.py
…
test_torchao.py
…
test_triton_attention_backend.py
…
test_triton_attention_kernels.py
…
test_triton_attention_rocm_mla.py
…
test_triton_fused_moe.py
…
test_triton_moe_channel_fp8_kernel.py
…
test_triton_moe_wna16.py
…
test_triton_sliding_window.py
…
test_two_batch_overlap.py
…
test_utils_update_weights.py
…
test_vertex_endpoint.py
…
test_vision_chunked_prefill.py
…
test_vision_openai_server_a.py
…
test_vision_openai_server_b.py
…
test_vision_openai_server_common.py
…
test_vllm_dependency.py
…
test_vlm_accuracy.py
…
test_vlm_input_format.py
…
test_wave_attention_backend.py
[AMD] Support Wave attention backend with AMD GPU optimizations (
#8660
)
2025-08-12 13:49:11 -07:00
test_wave_attention_kernels.py
…
test_weight_version.py
…