This website requires JavaScript.
Explore
Help
Register
Sign In
minke.yu
/
sglang
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
12
Packages
Projects
Releases
Wiki
Activity
Files
35eb7cf8d62396e9bfd16900c1dc97d76e0cab02
sglang
/
test
/
manual
/
attention
T
History
HuangJi
2a0cb2f04e
[Diffusion][MiniMax-H3] Add SM120 Sage compute for SubBlock sparse attention (
#40116
)
2026-09-20 16:40:28 +08:00
..
test_fa3.py
[misc] Use --cuda-graph-max-bs-decode in tests, examples, and docs (
#29591
)
2026-06-28 18:38:28 -07:00
test_flashattn_backend.py
fix(attention): read per-runner kv cache dtype off model_runner (
#32251
)
2026-07-23 20:08:57 -07:00
test_flashattn_mla_backend.py
fix(attention): read per-runner kv cache dtype off model_runner (
#32251
)
2026-07-23 20:08:57 -07:00
test_local_attn.py
[misc] Use --cuda-graph-max-bs-decode in tests, examples, and docs (
#29591
)
2026-06-28 18:38:28 -07:00
test_prefix_chunk_info.py
feat(model_runner): remove pool/backend refs from ForwardBatch via ForwardContext (
#25983
)
2026-05-21 14:01:49 -07:00
test_subblock_sage_fp8_sm120.py
[Diffusion][MiniMax-H3] Add SM120 Sage compute for SubBlock sparse attention (
#40116
)
2026-09-20 16:40:28 +08:00
test_trtllm_mla_backend.py
[CI][RFC] Replace black-jupyter with ruff-format (
#37210
)
2026-09-02 19:46:08 -07:00