![github-actions[bot]](/assets/img/avatar_default.png) Jimmy Shongandgithub-actions[bot]
|
54acffc864
|
Eval accuracy gpqa aime25 mixins (#27102)
Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>
|
2026-06-14 00:49:12 -07:00 |
|
Jimmy Shong
|
0ef39784ef
|
[Bugfix] Gate DP-attention even-token padding to CP-enabled configs (#26911)
|
2026-06-03 02:06:52 -04:00 |
|
Jimmy Shong
|
716e670d3d
|
[bugfix]: size CuteDSL MoE allgather buffers for the worst-case forward (#26696)
|
2026-05-30 00:27:20 -07:00 |
|
Jimmy Shong
|
f838adb7d4
|
bench_serving: add Zipfian shared-prefix sampling to generated-shared-prefix (#26378)
|
2026-05-28 14:39:46 -07:00 |
|
Jimmy Shong
|
1a85586738
|
[Fix]: Restrict Kimi-K2.5 shared-experts fusion to Quark MXFP4 checkpoints (#25974)
|
2026-05-21 13:07:45 -07:00 |
|
Jimmy Shong
|
daade9cc00
|
[Fix] Probe speculative draft config via sglang get_config (#25428)
|
2026-05-15 22:01:44 -07:00 |
|
Jimmy Shong
|
a741d0cc56
|
[CI] Lower mem-fraction-static for GLM-5.1 FP8 8-GPU test to 0.85 (#25453)
|
2026-05-15 20:14:47 -07:00 |
|
![github-actions[bot]](/assets/img/avatar_default.png) Jimmy Shongandgithub-actions[bot]
|
fd3eb77d45
|
[Cookbook]: add Laguna-XS.2 (Poolside) (#24730)
Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>
|
2026-05-12 16:06:26 +01:00 |
|
Jimmy Shong
|
e9a15b95da
|
[Fix] Disable FlashInfer allreduce fusion under deterministic inference (#24629)
|
2026-05-10 20:04:52 -05:00 |
|
![gemini-code-assist[bot]](/assets/img/avatar_default.png) Jimmy Shongandgemini-code-assist[bot]
|
fa8985486e
|
[test/fix]: isolate VLM MMMU eval output dirs to fix nightly-4-gpu cross-test pollution (#24623)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-05-08 15:01:53 -07:00 |
|
Jimmy Shong
|
096ad02b06
|
[Model] Laguna-XS.2 Model Support (#24204)
|
2026-05-09 05:43:13 +08:00 |
|
Jimmy Shong
|
3d31ac2672
|
[Fix] FP8 Qwen3-Next quant error by removing fallback fused shards (#23973)
|
2026-04-29 17:33:47 -04:00 |
|
 Jimmy ShongandSGLang CI
|
68a8ed9b11
|
[Fix/Kernel] Add JIT rmsnorm_hf kernel to fix transformers backend MMLU accuracy regression (#22931)
Co-authored-by: SGLang CI <ci@sglang.ai>
|
2026-04-23 12:00:31 +08:00 |
|
Jimmy Shong
|
28e915b474
|
[Bugfix] Preserve auto-detected quant_config for GLM NextN draft model (#22823)
|
2026-04-15 13:25:36 -07:00 |
|
Jimmy Shong
|
e83560562b
|
Update CI Permissions (#22826)
|
2026-04-14 15:13:31 -07:00 |
|