This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
13
Packages
Projects
Releases
Wiki
Activity
Files
733de6be31e20242421ea5cbc2a66582711d33f0
sglang
/
python
/
sglang
/
srt
/
hardware_backend
/
npu
T
History
Todobe
733de6be31
[NPU]Support GPT-OSS for NPU (
#14197
)
2026-01-19 04:13:41 +08:00
..
attention
[NPU]Support GPT-OSS for NPU (
#14197
)
2026-01-19 04:13:41 +08:00
graph_runner
[Feature] npu support enable_torch_compile for torchair backend (
#13410
)
2025-12-16 09:23:51 +08:00
modules
[DeepSeek v3.2] opt Context Parallelism: support fused moe, multi batch and fp8 kvcache (
#13959
)
2026-01-02 23:49:14 +08:00
moe
[NPU] optimization for dsv3.2 (
#14572
)
2025-12-12 14:52:16 +08:00
quantization
[NPU] NPU quantization refactoring & more quantization formats support (
#14504
)
2026-01-15 04:25:15 +08:00
allocator_npu.py
[NPU][Bugfix] move free_page logics to cpu (
#16608
)
2026-01-08 12:00:18 +08:00
cmo.py
…
memory_pool_npu.py
[NPU] [BUGFIX] [CRITICAL!] Fix NPU inference (torch_npu._npu_reshape_and_cache() crash) (
#15484
)
2025-12-20 14:53:46 +08:00
utils.py
[NPU] fix for NPU memory settings logic (
#15258
)
2025-12-16 17:04:22 -08:00