This website requires JavaScript.
Explore
Help
Sign In
chenchenghao
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
12
Packages
Projects
Releases
Wiki
Activity
Files
fb04e7e3c8c51d80ec4c091ed7c34d86a810a081
sglang
/
python
/
sglang
/
srt
/
hardware_backend
/
npu
T
History
hw-csong
261860e17b
[NPU][Bugfix] move free_page logics to cpu (
#16608
)
2026-01-08 12:00:18 +08:00
..
attention
[NPU] update Mixed chunk op to FIA (
#15518
)
2025-12-26 12:17:40 +08:00
graph_runner
[Feature] npu support enable_torch_compile for torchair backend (
#13410
)
2025-12-16 09:23:51 +08:00
modules
[DeepSeek v3.2] opt Context Parallelism: support fused moe, multi batch and fp8 kvcache (
#13959
)
2026-01-02 23:49:14 +08:00
moe
[NPU] optimization for dsv3.2 (
#14572
)
2025-12-12 14:52:16 +08:00
quantization
[NPU] Support w4a8 with activation clip (
#14736
)
2025-12-27 16:19:46 +08:00
allocator_npu.py
[NPU][Bugfix] move free_page logics to cpu (
#16608
)
2026-01-08 12:00:18 +08:00
cmo.py
…
memory_pool_npu.py
[NPU] [BUGFIX] [CRITICAL!] Fix NPU inference (torch_npu._npu_reshape_and_cache() crash) (
#15484
)
2025-12-20 14:53:46 +08:00
utils.py
[NPU] fix for NPU memory settings logic (
#15258
)
2025-12-16 17:04:22 -08:00