Logo
Explore Help
Register Sign In
minke.yu/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 13 Packages Projects Releases Wiki Activity
15,937 Commits 3 Branches 0 Tags
00a219f6c9e5b38ec7beddcfcb99266601182443
Commit Graph
10 Commits
This Branch
This Branch
All Branches
Author SHA1 Message Date
cen121212 b78d3999b5 【NPU】fix decode MTP + eagle shape error (#32791) 2026-07-30 21:34:16 +08:00
cen121212 0ffed946f2 [NPU] Add extra topk_weights input in deepep ll dispatch (#29480) 2026-07-09 09:23:29 +08:00
cen121212 9305d10099 [NPU] adapt_fused_rope_qk_mqa_optimize (#28872) 2026-06-25 08:55:52 +08:00
cen121212cen121212Even Zhou
b421e60eed 【NPU】【bugfix】fix server error when mtp unquant (#26389)
Co-authored-by: cen121212 <luochen23@huawei.com>
Co-authored-by: Even Zhou <even.y.zhou@outlook.com>
2026-05-30 15:01:19 +03:00
cen121212 461bc8af49 [NPU][Doc] Update GLM-5 docs, enabling deepep by default (#23708) 2026-05-08 11:12:35 +08:00
cen121212 b1e1fe8eee 【NPU】【bugfix】accuracy fix when enable both nsa cp and prefixcache (#23268) 2026-04-28 09:08:28 +08:00
cen121212 ba6d54d0f0 [NPU] GLM-5 optimize with fused kernels (#18617) 2026-03-30 22:48:15 +08:00
cen121212 fc543df289 [NPU] qwen3_vl encoder support graph 2026-03-09 10:13:35 +08:00
cen121212 0c2993eed0 Optimize Qwen3-VL video memory usage (#16366) 2026-01-22 09:10:08 +08:00
cen121212 25b48564c3 [NPU][Bugfix] fix Qwen3-VL-30B-A3B-Instruct accuracy loss (#15597) 2025-12-31 15:57:38 +08:00
Powered by Gitea Version: 1.27.3 Page: 339ms Template: 5ms
GitHub Default Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API