Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 11 Packages Projects Releases Wiki Activity
Files
733de6be31e20242421ea5cbc2a66582711d33f0
sglang/python/sglang/srt/hardware_backend/npu
History
Todobe 733de6be31 [NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
..
attention
[NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
graph_runner
[Feature] npu support enable_torch_compile for torchair backend (#13410)
2025-12-16 09:23:51 +08:00
modules
[DeepSeek v3.2] opt Context Parallelism: support fused moe, multi batch and fp8 kvcache (#13959)
2026-01-02 23:49:14 +08:00
moe
[NPU] optimization for dsv3.2 (#14572)
2025-12-12 14:52:16 +08:00
quantization
[NPU] NPU quantization refactoring & more quantization formats support (#14504)
2026-01-15 04:25:15 +08:00
allocator_npu.py
[NPU][Bugfix] move free_page logics to cpu (#16608)
2026-01-08 12:00:18 +08:00
cmo.py
[NPU][1/N] NPU basic functions refactor and new modelslim quant type (#13359)
2025-12-04 16:15:31 +08:00
memory_pool_npu.py
[NPU] [BUGFIX] [CRITICAL!] Fix NPU inference (torch_npu._npu_reshape_and_cache() crash) (#15484)
2025-12-20 14:53:46 +08:00
utils.py
[NPU] fix for NPU memory settings logic (#15258)
2025-12-16 17:04:22 -08:00
Powered by Gitea Version: 1.25.5 Page: 405ms Template: 2ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API