Logo
Explore Help
Register Sign In
minke.yu/sglang
Watch 1
Star 0
Fork 0
Code Issues Pull Requests Actions 12 Packages Projects Releases Wiki Activity
Files
ff1ce11348673eea9b51165bd6b2da7238c958b4
sglang/test/registered/kernel
T
History
Kevin MiClaude Fable 5.1Haocheng XiMick
ff1ce11348 [diffusion] model: support VDN-H3 with a hybrid_window_attn_h3 backend (#37903)
Co-authored-by: Claude Fable 5.1 <noreply@anthropic.com>
Co-authored-by: Haocheng Xi <xihc@berkeley.edu>
Co-authored-by: Mick <mickjagger19@icloud.com>
2026-09-12 11:36:32 +08:00
..
attention
[Diffusion][MiniMax-H3] Add SM90 Sage compute for SubBlock sparse attention (#37982)
2026-09-09 20:00:03 +08:00
diffusion
[diffusion] model: support VDN-H3 with a hybrid_window_attn_h3 backend (#37903)
2026-09-12 11:36:32 +08:00
embeddings
support qwen 3.8 flash next (#37500)
2026-09-08 13:56:21 -07:00
hyperconnection
support qwen 3.8 flash next (#37500)
2026-09-08 13:56:21 -07:00
jit
support qwen 3.8 flash next (#37500)
2026-09-08 13:56:21 -07:00
ops/attention
Store mamba prefix-cache checkpoints at the configured SSM state dtype (#34820)
2026-09-09 15:25:57 +08:00
qsa
fix(qsa): make the paged sparse-decode gather memory-safe (zero-fill scratch, int64 offsets, dequant FP8 on gather) (#38851)
2026-09-11 15:47:25 -07:00
quantization
[AMD] Fix weight checking for AITER-shuffled block FP8 weights (#34330)
2026-09-11 15:49:56 -07:00
speculative
[AMD] Parallelize aiter spec-decode KV index building over token blocks (#37659)
2026-09-09 17:53:53 -07:00
Powered by Gitea Version: 1.27.3 Page: 480ms Template: 4ms
GitHub Default Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API