[Kernel] Add H20 block-FP8 MoE configs for GLM-5.3-Flash EP4/EP8 (#38913)
This commit is contained in:
@@ -107,6 +107,7 @@ def get_model_config(
|
||||
"Glm4MoeForCausalLM",
|
||||
"Glm4MoeLiteForCausalLM",
|
||||
"GlmMoeDsaForCausalLM",
|
||||
"Glm5NextForConditionalGeneration",
|
||||
"KimiVLForConditionalGeneration",
|
||||
"MistralLarge3ForCausalLM",
|
||||
]:
|
||||
|
||||
Reference in New Issue
Block a user