Logo
Explore Help
Register Sign In
minke.yu/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
Files
bca8ed4afc03166c55a3b5e1680d5ae21be2fa13
sglang/python/sglang/srt/distributed
T
History
Shu Wang c7c03ec53b [NVIDIA] Add flashinfer MNNVL backend for allreduce only (#30700)
2026-08-11 16:46:46 -07:00
..
device_communicators
Consolidate CUDA VMM allocation helpers (#34199)
2026-08-10 18:11:11 -07:00
__init__.py
Roll back to use vllm custom allreduce (#3006)
2025-01-20 04:03:15 -08:00
bootstrap.py
[NVIDIA] Add flashinfer MNNVL backend for allreduce only (#30700)
2026-08-11 16:46:46 -07:00
communication_op.py
[AMD] Add fused all-reduce RMSNorm per-group quant for Qwen3.5 FP8 (#24651)
2026-07-22 07:33:03 -07:00
communication_tags.py
Fix: add grammar sync in PP for structured output (#30747)
2026-07-12 02:54:50 +08:00
naive_distributed.py
vlm: enforce pybase64 for image and str encode/decode (#10700)
2025-10-21 19:05:32 +08:00
parallel_state_wrapper.py
config: derive the runner's DCP topology from its ParallelState (#34133)
2026-08-09 01:18:24 -07:00
parallel_state.py
[NVIDIA] Add flashinfer MNNVL backend for allreduce only (#30700)
2026-08-11 16:46:46 -07:00
utils.py
[refactor] Retire the legacy config accessor and the remaining process singletons (#30493)
2026-07-09 02:10:47 -07:00
Powered by Gitea Version: 1.27.3 Page: 778ms Template: 3ms
GitHub Default Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API