This website requires JavaScript.
Explore
Help
Register
Sign In
minke.yu
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
13
Packages
Projects
Releases
Wiki
Activity
Files
9d878c1f3e1e5bd1ac7be9417e5b3cd6308cd23a
sglang
/
python
/
sglang
/
srt
/
distributed
T
History
Yuan Luo
and
luoyuan.luo
019517a356
[VLM] Support ViT Piecewise CUDA Graph for Qwen3-VL (
#15320
)
...
Co-authored-by: luoyuan.luo <
luoyuan.luo@antgroup.com
>
2025-12-20 21:00:07 +08:00
..
device_communicators
[amd] Add deterministic all-reduce kernel for AMD (ROCm) (
#15340
)
2025-12-18 23:36:03 -08:00
__init__.py
Roll back to use vllm custom allreduce (
#3006
)
2025-01-20 04:03:15 -08:00
communication_op.py
Sync distributed package from vllm 0.6.4.post1 (
#3010
)
2025-01-20 04:57:14 -08:00
naive_distributed.py
vlm: enforce pybase64 for image and str encode/decode (
#10700
)
2025-10-21 19:05:32 +08:00
parallel_state.py
[VLM] Support ViT Piecewise CUDA Graph for Qwen3-VL (
#15320
)
2025-12-20 21:00:07 +08:00
utils.py
Optimize uneven PP layer distribution logic to improve PP performance (
#13977
)
2025-11-26 22:05:03 +08:00