Junlin Wu
|
09ecb9aaaa
|
📝 [NPU] Use vendor-neutral wording in quantization comments (#36768)
|
2026-08-29 23:05:35 +03:00 |
|
 Junlin Wuandronnie_zheng
|
308bc1228b
|
📝 [NPU] Clean up quantization comments (#34829)
Co-authored-by: ronnie_zheng <zl19940307@163.com>
|
2026-08-20 22:03:58 +03:00 |
|
Junlin Wu
|
9ab88380c1
|
👥 chore(codeowners): Update codeowners for NPU quantization (#31917)
|
2026-07-29 12:10:01 +03:00 |
|
 
|
f05c92fb6d
|
✨ [llm][npu][quant] Add W8A8 MXFP8 quantization for Qwen3 MoE on Ascend NPU (#30768)
Co-authored-by: Артем Савкин <58187114+OrangeRedeng@users.noreply.github.com>
Co-authored-by: ronnie_zheng <zl19940307@163.com>
|
2026-07-29 10:39:36 +03:00 |
|
Junlin Wu
|
d6fcfe02d6
|
🐛 [llm][npu][quant] Fix ModelSlim MXFP4 packed weight loading (#32013)
|
2026-07-29 11:34:41 +08:00 |
|
Junlin Wu
|
bbd2a3fe4a
|
✨ [llm][npu][quant] Add W4A4 MXFP4 quantization support for Qwen3 Dense on Ascend NPU (#23795)
|
2026-07-17 09:06:30 +03:00 |
|
Junlin Wu
|
aae04b1241
|
📝 docs(diffusion): add MXFP4 quantization docs (#25904)
|
2026-05-25 10:24:30 +03:00 |
|
 Junlin Wuandronnie_zheng
|
4c9f31b85e
|
✨ [diffusion][npu][quant] Add MXFP4 quantization support for Wan2.2 Diffusion on Ascend NPU (#22338)
Co-authored-by: ronnie_zheng <zl19940307@163.com>
|
2026-05-19 07:46:52 +03:00 |
|
Junlin Wu
|
a623ee4cb5
|
📝 docs(diffusion): add MXFP8 quantization docs for Wan2.2 on Ascend NPU (#24918)
|
2026-05-11 08:13:34 +03:00 |
|
![github-actions[bot]](/assets/img/avatar_default.png) 
|
80a6014243
|
✨ [diffusion][npu][quant] Add MXFP8 quantization support for Wan2.2 Diffusion on Ascend NPU (#20922)
Co-authored-by: ronnie_zheng <zl19940307@163.com>
Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>
|
2026-05-07 21:30:56 +03:00 |
|