Logo
Explore Help
Register Sign In
minke.yu/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
12,612 Commits 3 Branches 0 Tags
65e9f81c7d4e00fe6b1c8800ee8e724737d756e5
Commit Graph
8 Commits
This Branch
This Branch
All Branches
Author SHA1 Message Date
Yuzhen ZhouandByron Hsu 4a279d9c36 [R3] Avoid implicit CUDA sync in routed experts DP slicing (#24550)
Co-authored-by: Byron Hsu <byronhsu1230@gmail.com>
2026-05-06 18:37:36 -07:00
Yuzhen Zhou 6b876a7710 [ROCM][RL] Shuffle Weight In-Place to Preserve Parameter Attributes (#21825) 2026-04-02 23:43:55 -07:00
Yuzhen Zhou b719219de9 [ROCm] Use unreg path for aiter custom all-reduce during CUDA graph capture (#20155) 2026-03-09 01:09:04 -07:00
Yuzhen Zhou 63003a39cf [BUG] Support tuple hidden_states from fused MXFP4/FP8 quantization (#19643) 2026-03-02 20:39:06 -08:00
Yuzhen Zhou 4f3dc1ef5b [ROCm] Use unreg path for custom all-reduce during CUDA graph capture (#19162) 2026-02-22 23:27:31 -08:00
Yuzhen Zhou 2169025b77 turn off dit_layerwise_offload for wan on rocm (#17569) 2026-01-23 15:22:42 +08:00
Yuzhen ZhouSabre ShaoYusheng SuHubert Luxsun
4bf06635fc [diffusion] multi-platform: support diffusion on amd and fix encoder loading on MI325 (#13760)
Co-authored-by: Sabre Shao <sabre.shao@amd.com>
Co-authored-by: Yusheng (Ethan) Su <yushengsu.thu@gmail.com>
Co-authored-by: Hubert Lu <Hubert.Lu@amd.com>
Co-authored-by: xsun <sunxiao04@gmail.com>
2025-12-19 15:38:46 +08:00
Yuzhen Zhou 0380ca82ef Add Batch‑Invariant RMSNorm (#12144) 2025-10-28 21:05:57 -07:00
Powered by Gitea Version: 1.27.3 Page: 300ms Template: 3ms
GitHub Default Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API