This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
11
Packages
Projects
Releases
Wiki
Activity
Files
95575aa76a0716aa018ac2359dc63b934f94ff51
sglang
/
test
/
srt
T
History
…
..
configs
…
models
…
double-sparsity-config-Llama-3.1-8B-Instruct.json
…
experiment_runner.py
…
kv_cache_scales_llama3_1_8b.json
…
kv_cache_scales_llama3_8b.json
…
kv_cache_scales_qwen2_1_5b.json
…
run_suite.py
Reasoning parser (
#4000
)
2025-03-03 21:16:36 -08:00
test_abort.py
…
test_bench_one_batch.py
…
test_bench_serving.py
…
test_block_int8.py
…
test_cache_report.py
…
test_chunked_prefill.py
…
test_create_kvindices.py
…
test_custom_allreduce.py
…
test_data_parallelism.py
…
test_double_sparsity.py
Crash the server correctly during error (
#2231
)
2024-11-28 00:22:39 -08:00
test_dp_attention.py
…
test_eagle_infer.py
…
test_ebnf_constrained.py
…
test_embedding_openai_server.py
…
test_eval_accuracy_large_chunked_prefill.py
…
test_eval_accuracy_large_mixed_chunked_prefill.py
…
test_eval_accuracy_large.py
…
test_eval_accuracy_mini.py
…
test_fp8_kernel.py
…
test_fp8_kvcache.py
support e4m3 kvcache in qwen2 & add kv scaling facotr json (
#2894
)
2025-01-18 11:43:22 +08:00
test_function_calling.py
…
test_fused_moe.py
…
test_get_weights_by_name.py
…
test_gguf.py
…
test_health_check.py
…
test_hidden_states.py
…
test_input_embeddings.py
…
test_json_constrained.py
…
test_large_max_new_tokens.py
…
test_matched_stop.py
Crash the server correctly during error (
#2231
)
2024-11-28 00:22:39 -08:00
test_metrics.py
…
test_mla_flashinfer.py
…
test_mla_fp8.py
…
test_mla_tp.py
…
test_mla.py
Support penalty in overlap mode; return logprob with chunked prefill; improve benchmark scripts (
#3988
)
2025-03-03 00:12:04 -08:00
test_modelopt_fp8kvcache.py
…
test_models_from_modelscope.py
…
test_moe_ep.py
…
test_moe_eval_accuracy_large.py
…
test_nightly_gsm8k_eval.py
…
test_nightly_human_eval.py
…
test_nightly_math_eval.py
…
test_no_chunked_prefill.py
…
test_no_overlap_scheduler.py
…
test_openai_server.py
…
test_penalty.py
…
test_pytorch_sampling_backend.py
…
test_radix_attention.py
…
test_reasoning_content.py
Reasoning parser (
#4000
)
2025-03-03 21:16:36 -08:00
test_regex_constrained.py
…
test_release_memory_occupation.py
…
test_request_length_validation.py
…
test_retract_decode.py
Crash the server correctly during error (
#2231
)
2024-11-28 00:22:39 -08:00
test_sagemaker_server.py
…
test_schedule_policy.py
…
test_server_args.py
…
test_session_control.py
…
test_skip_tokenizer_init.py
…
test_srt_endpoint.py
…
test_srt_engine_with_quant_args.py
…
test_srt_engine.py
…
test_torch_compile_moe.py
Improve torch compile for fused moe (
#2327
)
2024-12-03 01:58:25 -08:00
test_torch_compile.py
…
test_torch_native_attention_backend.py
Add a simple torch native attention backend (
#2241
)
2024-12-01 03:01:25 -08:00
test_torch_tp.py
…
test_torchao.py
…
test_triton_attention_backend.py
…
test_triton_attention_kernels.py
…
test_triton_attention_rocm_mla.py
…
test_update_weights_from_disk.py
…
test_update_weights_from_distributed.py
…
test_update_weights_from_tensor.py
…
test_verl_engine.py
…
test_vertex_endpoint.py
…
test_vision_chunked_prefill.py
Fix CI and install docs (
#3821
)
2025-02-24 16:17:38 -08:00
test_vision_llm.py
…
test_vision_openai_server.py
…
test_w8a8_quantization.py
…