Logo
Explore Help
Sign In
chenchenghao/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 13 Packages Projects Releases Wiki Activity
Files
733de6be31e20242421ea5cbc2a66582711d33f0
sglang/python/sglang/srt/hardware_backend/npu
T
History
Todobe 733de6be31 [NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
..
attention
[NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
graph_runner
[Feature] npu support enable_torch_compile for torchair backend (#13410)
2025-12-16 09:23:51 +08:00
modules
[DeepSeek v3.2] opt Context Parallelism: support fused moe, multi batch and fp8 kvcache (#13959)
2026-01-02 23:49:14 +08:00
moe
[NPU] optimization for dsv3.2 (#14572)
2025-12-12 14:52:16 +08:00
quantization
[NPU] NPU quantization refactoring & more quantization formats support (#14504)
2026-01-15 04:25:15 +08:00
allocator_npu.py
[NPU][Bugfix] move free_page logics to cpu (#16608)
2026-01-08 12:00:18 +08:00
cmo.py
…
memory_pool_npu.py
[NPU] [BUGFIX] [CRITICAL!] Fix NPU inference (torch_npu._npu_reshape_and_cache() crash) (#15484)
2025-12-20 14:53:46 +08:00
utils.py
[NPU] fix for NPU memory settings logic (#15258)
2025-12-16 17:04:22 -08:00
Powered by Gitea Version: 1.27.2 Page: 1785ms Template: 182ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API