 
|
67e9d287ee
|
[Quantization] Support Quark Dense + MoE FP8 & FP8 PTPC (#10485)
Co-authored-by: HAI <hixiao@gmail.com>
Co-authored-by: kk <43161300+kkHuang-amd@users.noreply.github.com>
|
2025-11-13 08:16:00 -08:00 |
|
Bowen Bao
|
cd4b39a900
|
[quantization] Properly ignore quantization for layers excluded in quant_config (#11205)
|
2025-10-07 14:06:05 -07:00 |
|
 Bowen BaoandHaiShaw
|
baee08601b
|
[quantization] Enable aiter mxfp4 fused_moe for Quark (#10048)
Co-authored-by: HaiShaw <hixiao@gmail.com>
|
2025-10-05 19:51:34 -07:00 |
|
 Bowen BaoandHAI
|
c7a104c12b
|
[quantization] Fix scale remapping for mllama4 (#10042)
Co-authored-by: HAI <hixiao@gmail.com>
|
2025-10-05 19:51:15 -07:00 |
|