13 Commits
Author SHA1 Message Date
gaopengffandMa Mingfei 317da0964e [Intel XPU] support prefill only models for xpu (#35072)
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
2026-08-24 16:25:47 +08:00
gaopengffandMa Mingfei 56834422a1 [Intel XPU] Add xpu pass for biased_topk and hash_topk (#33323)
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
2026-08-24 12:18:22 +08:00
gaopengff bab1dd0d12 [Intel XPU] Enable (biased) grouped topk for xpu (#31126) 2026-07-20 09:35:08 +08:00
gaopengff dc1e46ec8f [Intel GPU]Add sycl mrope pass for xpu device (#27646) 2026-06-11 12:25:55 +08:00
gaopengff aa510bda45 Support specific pass of bias_grouped_topk for xpu (#26349) 2026-06-03 13:13:48 +08:00
gaopengff eda21f6839 Add fused_rope and for xpu (#25773) 2026-06-03 09:41:42 +08:00
gaopengff 65fe32379e [Intel GPU]Support fused_topk for XPU (#24641) 2026-05-20 10:34:29 +08:00
gaopengff f4393bf3f6 Fix correctness test issue for bench_one_batch (#20650) 2026-03-15 20:05:36 -07:00
gaopengff 7541da15d2 Fix prefill latency performance drop of bench serving (#14592) 2026-01-29 21:28:17 -08:00
gaopengff 077ca70ee4 [Intel XPU]Add xpu support for get_device_memory_capacity (#13895) 2025-11-26 20:55:52 -08:00
gaopengff aeac622058 [Intel XPU]support xgrammar backend for intel xpu (#13245) 2025-11-24 16:48:00 +08:00
gaopengff 4e234b4cf9 [Intel XPU]Update pytorch xpu to 2.9 (#12363) 2025-11-06 10:34:23 -08:00
gaopengff 4cc725ac1c [Intel]Add 'intel_xpu' attention backend for llama4 (#11051) 2025-11-06 10:33:41 -08:00