[Kernel] Add fused MoE Triton configs for Qwen3.8-Flash-Next FP8 on NVIDIA H200 NVL (TP2+EP2) (#38116)
This commit is contained in:
@@ -84,6 +84,7 @@ def get_model_config(
|
||||
"Qwen3VLMoeForConditionalGeneration",
|
||||
"Qwen3_5MoeForCausalLM",
|
||||
"Qwen3_5MoeForConditionalGeneration",
|
||||
"Qwen4ExpForConditionalGeneration",
|
||||
"InternS2PreviewForConditionalGeneration",
|
||||
"MellumForCausalLM",
|
||||
]:
|
||||
|
||||
Reference in New Issue
Block a user