Logo
Explore Help
Register Sign In
minke.yu/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 14 Packages Projects Releases Wiki Activity
Files
7431f35fd8a93c4a8a193d372f7697b55cf9b6d0
sglang/python/sglang/srt/model_loader
T
History
Jinzhen Linguzekai01Julian Huang墨楼Claude Opus 4.8Peng ZhangXiaoyu Zhang
423b8485fb [Quantization] add humming quantization kernel (#23754)
Co-authored-by: guzekai01 <zekai01@antgroup.com>
Co-authored-by: Julian Huang <huangzhilin.hzl@gmail.com>
Co-authored-by: 墨楼 <huangzhilin.hzl@antgroup.com>
Co-authored-by: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-authored-by: Peng Zhang <aniz1905@gmail.com>
Co-authored-by: Xiaoyu Zhang <1182563586@qq.com>
2026-07-14 08:42:56 +08:00
..
__init__.py
chore: add vLLM SPDX copyright headers to ported files (#25182)
2026-05-13 15:17:30 -07:00
ci_weight_validation.py
fix: accept 0-indexed safetensors shard names in CI weight validator (#24237)
2026-05-02 00:58:15 -07:00
loader.py
[Quantization] add humming quantization kernel (#23754)
2026-07-14 08:42:56 +08:00
remote_instance_weight_loader_utils.py
Fix remote weight info nnode>1 and dp>1 (#17389)
2026-03-31 21:17:18 +08:00
utils.py
fix: Fix DSR1 perf regression due to unnecessarily falling back to triton gemm (#28073)
2026-06-15 09:45:09 -04:00
weight_utils.py
[Model] Support Qwen3.6 ModelOpt mixed NVFP4 (#27906)
2026-07-05 21:31:15 -07:00
Powered by Gitea Version: 1.27.3 Page: 437ms Template: 3ms
GitHub Default Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API