This website requires JavaScript.
Explore
Help
Register
Sign In
minke.yu
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
15
Packages
Projects
Releases
Wiki
Activity
Files
5da265de30e8168067be5915fbbf8904e39abb5d
sglang
/
python
/
sglang
/
srt
/
model_loader
T
History
Trevor Morris
5da265de30
[NVIDIA] Fix FP8 gemm performance with fp16 models (MInimax-M2.5) (
#22300
)
2026-06-07 02:45:00 +00:00
..
__init__.py
chore: add vLLM SPDX copyright headers to ported files (
#25182
)
2026-05-13 15:17:30 -07:00
ci_weight_validation.py
fix: accept 0-indexed safetensors shard names in CI weight validator (
#24237
)
2026-05-02 00:58:15 -07:00
loader.py
[AMD][MXFP4] Online MXFP4 quantization 1/N - dense and MOE models w. original BF16 weight (
#18005
)
2026-06-03 12:55:24 -07:00
remote_instance_weight_loader_utils.py
Fix remote weight info nnode>1 and dp>1 (
#17389
)
2026-03-31 21:17:18 +08:00
utils.py
[NVIDIA] Fix FP8 gemm performance with fp16 models (MInimax-M2.5) (
#22300
)
2026-06-07 02:45:00 +00:00
weight_utils.py
[AMD][MXFP4] Online MXFP4 quantization 1/N - dense and MOE models w. original BF16 weight (
#18005
)
2026-06-03 12:55:24 -07:00