Co-authored-by: AI-bot-easy <litchys0123@outlook.com> Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
18 lines
888 B
Markdown
18 lines
888 B
Markdown
# Code Structure
|
|
|
|
- `eval`: The evaluation utilities.
|
|
- `lang`: The frontend language.
|
|
- `srt`: The backend engine for running local models. (SRT = SGLang Runtime).
|
|
- `test`: The test utilities.
|
|
- `api.py`: The public APIs.
|
|
- `bench_offline_throughput.py`: Benchmark the performance in the offline mode.
|
|
- `bench_one_batch.py`: Benchmark the latency of running a single static batch without a server.
|
|
- `bench_one_batch_server.py`: Benchmark the latency of running a single batch with a server.
|
|
- `bench_serving.py`: Benchmark online serving with dynamic requests.
|
|
- `check_env.py`: Check the environment variables and dependencies.
|
|
- `global_config.py`: The global configs and constants.
|
|
- `launch_server.py`: The entry point for launching a local server.
|
|
- `profiler.py`: The profiling entry point to send profile requests.
|
|
- `utils.py`: Common utilities.
|
|
- `version.py`: Version info.
|