Alison Shao
|
0c474273c5
|
Fix gpt_oss_common import path and migrate core tests (#16426)
|
2026-01-07 12:58:32 -08:00 |
|
fzyzcjy
|
d874c8bba4
|
Tiny support http headers in bench serving (#16606)
|
2026-01-07 10:15:17 +08:00 |
|
Junrong Lin
|
bc2f40bebc
|
[test] Add mamba cache release/resume memory test (#14215)
|
2026-01-06 15:51:10 +08:00 |
|
Alison Shao
|
73398e22d6
|
ci: migrate VLM tests to test/registered/vlm/ (#16415)
|
2026-01-05 21:25:29 -08:00 |
|
fzyzcjy
|
c105a3124b
|
Support multi-round conversations in bench_serving (#6135)
|
2026-01-06 11:59:39 +08:00 |
|
Netanel Haber
|
bebd625ba1
|
EVS Framework: Support NemotronH_Nano_VL_V2 (#14051)
|
2026-01-05 16:18:07 +08:00 |
|
Douglas Yang
|
87699d48eb
|
fix: only publish trace from tp 0 (#16411)
|
2026-01-04 12:17:08 -08:00 |
|
![gemini-code-assist[bot]](/assets/img/avatar_default.png) Hudson Xingandgemini-code-assist[bot]
|
f4ab2ec5be
|
Add unified metrics collection framework (v1) (#16064)
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
|
2026-01-03 16:30:38 -08:00 |
|
Baizhou Zhang
|
f35b5da521
|
[CI] Append test variant name to markdown report header in nightly test (#16166)
|
2025-12-31 00:09:24 +08:00 |
|
Vladimir Serov
|
9e263c2162
|
[LoRA] Torch native backend: rework implementation and updated tests (#15187)
|
2025-12-30 11:48:51 +08:00 |
|
Liangsheng Yin
|
a435f55d18
|
Tiny print launch command with shlex (#16010)
|
2025-12-29 11:26:46 +08:00 |
|
lif
|
5969be2f06
|
Apply fixture-kit mode to MMMUVLMMixin (#15615)
|
2025-12-28 17:22:05 +08:00 |
|
 Alison ShaoandKangyan-Zhou
|
0e536600e8
|
Refactor: separate CI-specific weight validation into dedicated module (#15216)
Co-authored-by: Kangyan-Zhou <zky314343421@gmail.com>
|
2025-12-27 20:50:39 -08:00 |
|
Lianmin Zheng
|
183b65190a
|
Clean up logging (#15919)
|
2025-12-27 15:27:12 -08:00 |
|
Ke Bao
|
faecd37ed4
|
Add Mimo-v2-flash model to ci test (#15887)
|
2025-12-27 14:18:08 +08:00 |
|
Liangsheng Yin
|
9ad546d7e8
|
Tiny cleanup the models' name in test_utils (#15920)
|
2025-12-27 14:13:23 +08:00 |
|
Douglas Yang
|
17e65466de
|
fix: nightly fix b200 gpqa (#15745)
|
2025-12-24 10:29:45 -08:00 |
|
Liangsheng Yin
|
159b128357
|
Tiny add flush for CI crash locating (#15769)
|
2025-12-24 22:47:10 +08:00 |
|
michael-amd
|
e7b09efc0a
|
[AMD] Add AMD Nightly Performance & VLMs Accuracy Tests (#15500)
|
2025-12-23 19:03:27 -08:00 |
|
Liangsheng Yin
|
bd572360f3
|
Tiny apply gsm8k mixin to ngram test (#15606)
|
2025-12-24 01:30:26 +08:00 |
|
 Yubo WangandLiangsheng Yin
|
762846531f
|
Fix Illegal Memory Access when fa3 + spec + topk + page_size > 1 (#15469)
Co-authored-by: Liangsheng Yin <lsyincs@gmail.com>
|
2025-12-24 00:13:57 +08:00 |
|
Douglas Yang
|
f9dd90ac35
|
fix: increasing H200 test timeout (#15600)
|
2025-12-23 01:00:37 -08:00 |
|
Alison Shao
|
ac42797cf7
|
[CI] Enable retry logic for flaky CI tests (#14983)
|
2025-12-22 22:23:42 -08:00 |
|
Alison Shao
|
989d4b3012
|
[CI] Migrate nightly tests to test/registered/ (#15582)
|
2025-12-22 22:16:32 -08:00 |
|
Liangsheng Yin
|
beae3f961c
|
Adapt fixture-kit to gsm8k mixin (#15599)
|
2025-12-22 14:19:35 +08:00 |
|
 
|
bed301a5ac
|
[Feature] Enable return routed experts (#12162)
Co-authored-by: yizhang2077 <1109276519@qq.com>
Co-authored-by: Liangsheng Yin <lsyincs@gmail.com>
|
2025-12-21 15:16:43 +08:00 |
|
 ![gemini-code-assist[bot]](/assets/img/avatar_default.png)     
|
1f1f05a85e
|
vlm: refactor engine vlm params and support processor output as input (#14091)
Co-authored-by: Mick <mickjagger19@icloud.com>
Co-authored-by: zhaochenyang20 <zhaochenyang20@gmail.com>
Co-authored-by: Xinyuan Tong <115166877+JustinTong0323@users.noreply.github.com>
Co-authored-by: BenYao21 <cyao22@asu.edu>
Co-authored-by: minleminzui <minleminzui@gmail.com>
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
Co-authored-by: 赵晨阳 <zhaochen20@outlook.com>
|
2025-12-20 18:31:24 +08:00 |
|
Alison Shao
|
4128d4f5cb
|
[CI] Migrate LoRA tests to test/registered/lora/ (#15176)
|
2025-12-17 13:19:42 -08:00 |
|
 
|
45a959d3e9
|
[PP] Add pp support for Qwen3-VL (#12333)
Signed-off-by: Xuchun Shang <xuchun.shang@gmail.com>
Signed-off-by: Kun(llfl) <i@imux.top>
Signed-off-by: Kun(llfl) <llfl@linux.alibaba.com>
Co-authored-by: kun-llfl <i@imux.top>
Co-authored-by: Kun(llfl) <llfl@linux.alibaba.com>
|
2025-12-17 16:03:58 +08:00 |
|
 
|
3e4d431a44
|
[Feature] Add AIME25 dataset support for SGLang simple_eval (#14990)
Co-authored-by: zkexorability <zkexorability@gmail.com>
Co-authored-by: Baizhou Zhang <sobereddiezhang@gmail.com>
|
2025-12-15 21:59:40 -08:00 |
|
Douglas Yang
|
9e9a61691e
|
ci: adding errors to Github summary (#14778)
|
2025-12-14 21:08:16 -08:00 |
|
Hanming Lu
|
e61dabf5e4
|
[Qwen3-next] support mamba radix cache for overlap scheduler (#14792)
|
2025-12-14 18:54:16 -08:00 |
|
Baizhou Zhang
|
ab3ffd1c8e
|
Add nightly accuracy test for DeepSeek V3.2 (#14935)
|
2025-12-13 12:11:16 -08:00 |
|
Yuhao Yang
|
06b58c5dc5
|
fix flaky image access in ci by switching to raw content url (#14940)
|
2025-12-13 10:52:06 -08:00 |
|
Liangsheng Yin
|
90e7d4f78f
|
Tiny adjust CI run suite (#15074)
|
2025-12-14 00:05:19 +08:00 |
|
Kangyan-Zhou
|
b243154614
|
Fix CI by reverting incorrect metric check logic (#15004)
|
2025-12-12 10:07:38 -08:00 |
|
Liangsheng Yin
|
c660d8dfd0
|
Re-org eagle unit tests (#14909)
|
2025-12-12 12:25:39 +09:00 |
|
Alison Shao
|
0aa3dec5c7
|
Fix black formatting in ci_utils.py (#14932)
|
2025-12-11 17:28:52 -08:00 |
|
Alison Shao
|
e59435c34b
|
Add retry logic for scheduled CI tests (#14771)
|
2025-12-11 16:59:58 -08:00 |
|
 b8zhongandBrayden Zhong
|
6107268fe7
|
extend timeout for b200 test (#14925)
Co-authored-by: Brayden Zhong <b8zhong@users.noreply.github.com>
|
2025-12-11 15:43:53 -08:00 |
|
Liangsheng Yin
|
543d62d11a
|
Introduce server_fixtures in sglang.test (#14899)
|
2025-12-11 22:30:33 +09:00 |
|
 Vladimir221andronnie_zheng
|
27032cecd9
|
[Ascend]Support of piecewise graph compilation for prefill on NPU (#12287)
Co-authored-by: ronnie_zheng <zl19940307@163.com>
|
2025-12-11 21:10:07 +08:00 |
|
Yuhao Yang
|
b62fe8504c
|
fix nightly vlm ci : restore original eval for requests without regex (#14875)
|
2025-12-10 23:13:25 -08:00 |
|
Binyao Jiang
|
312df1d6c0
|
Fix TestGLM41VPPAccuracy test flakiness (#14848)
|
2025-12-10 16:59:58 -08:00 |
|
 Yuan Luoandluoyuan.luo
|
03836d85d2
|
[GLM-4.6V] Support Pipeline Parallelism for GLM-4.6V & GLM-4.1V (#14720)
Co-authored-by: luoyuan.luo <luoyuan.luo@antgroup.com>
|
2025-12-10 16:40:12 +08:00 |
|
 
|
0c63fb9420
|
[Feature] Add LoRA support for embedding layers (#14177)
Co-authored-by: Baizhou Zhang <sobereddiezhang@gmail.com>
Co-authored-by: Beichen-Ma <bm685@cornell.edu>
|
2025-12-09 15:53:33 -08:00 |
|
Xiaoyu Zhang
|
53d170883a
|
Add fuse_marlin_moe test to ci and add new ep test (#14686)
|
2025-12-09 20:17:38 +08:00 |
|
 b8zhongandBrayden Zhong
|
3b47973af8
|
[CI] Tiny speed up VLM CI (#14517)
Co-authored-by: Brayden Zhong <b8zhong@users.noreply.github.com>
|
2025-12-07 13:30:41 -08:00 |
|
Hanming Lu
|
e592ee6545
|
[Qwen3-next] remove heuristics and add radix cache kl test (#14520)
|
2025-12-06 12:11:40 -08:00 |
|
Cherry_ming
|
1808df48fe
|
[NPU]add nightly-test-npu (#14143)
|
2025-12-05 00:43:35 +08:00 |
|