Logo
Explore Help
Sign In
chenchenghao/sglang
1
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
Files
8fb45523f339f2eb3ea6815934d12afd43778b8b
sglang/python/sglang/srt/hardware_backend/npu
History
Todobe 733de6be31 [NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
..
attention
[NPU]Support GPT-OSS for NPU (#14197)
2026-01-19 04:13:41 +08:00
graph_runner
[Feature] npu support enable_torch_compile for torchair backend (#13410)
2025-12-16 09:23:51 +08:00
modules
[DeepSeek v3.2] opt Context Parallelism: support fused moe, multi batch and fp8 kvcache (#13959)
2026-01-02 23:49:14 +08:00
moe
[NPU] optimization for dsv3.2 (#14572)
2025-12-12 14:52:16 +08:00
quantization
[NPU] NPU quantization refactoring & more quantization formats support (#14504)
2026-01-15 04:25:15 +08:00
allocator_npu.py
[NPU][Bugfix] move free_page logics to cpu (#16608)
2026-01-08 12:00:18 +08:00
cmo.py
[NPU][1/N] NPU basic functions refactor and new modelslim quant type (#13359)
2025-12-04 16:15:31 +08:00
memory_pool_npu.py
[NPU] [BUGFIX] [CRITICAL!] Fix NPU inference (torch_npu._npu_reshape_and_cache() crash) (#15484)
2025-12-20 14:53:46 +08:00
utils.py
[NPU] fix for NPU memory settings logic (#15258)
2025-12-16 17:04:22 -08:00
Powered by Gitea Version: 1.25.5 Page: 412ms Template: 4ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API