 Polisetty V R K Jyothendra VarmaandMa Mingfei
|
62b3c8e177
|
[Intel GPU] Guard tvm_ffi import in dsv4 online mtp module under TYPE_CHECKING to fix import error on XPU (#28531)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
|
2026-06-22 13:56:21 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
fbbf559de2
|
fix bench_one_batch by extending array with array not list (#28732)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-06-21 08:28:41 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
d7c8b9ab9f
|
[Intel GPU] Enable fused_experts in fp8.py for quantized models on XPU (#27533)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-06-09 09:22:30 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
fd94bd30b8
|
[Intel GPU] DeepSeek V4 2/N: Fix tvm ffi import (#26118)
|
2026-05-24 12:59:49 -07:00 |
|
Polisetty V R K Jyothendra Varma
|
80680dc3fe
|
[Intel GPU] 1/N Fix tilelang import in deepseek v4 rope as optional (#25128)
|
2026-05-22 18:23:18 +08:00 |
|
 Polisetty V R K Jyothendra VarmaandBrayden Zhong
|
52d4c697bb
|
Fix fused_moe import for non-NPU devices (#25076)
Co-authored-by: Brayden Zhong <b8zhong@uwaterloo.ca>
|
2026-05-12 23:05:51 +03:00 |
|
 Polisetty V R K Jyothendra VarmaandMa Mingfei
|
50ed01674e
|
fix is_arch_support_pdl function usage (#24600)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
|
2026-05-09 09:39:34 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
9dfb1d2ebe
|
[Intel GPU] Fix flash_mla_get_workspace_size call in intel_xpu (#24372)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-05-07 13:45:32 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
fdfc46f3a5
|
[Intel GPU] Enable DeepSeek V3.2 inference on XPU (#24356)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-05-05 20:47:40 +08:00 |
|
  ![gemini-code-assist[bot]](/assets/img/avatar_default.png)
|
da7f890788
|
[Intel GPU] Integrate flash_mla_decode in Intel XPU attention backend (#23557)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
Co-authored-by: gemini-code-assist[bot] <176961590+gemini-code-assist[bot]@users.noreply.github.com>
Co-authored-by: Kangyan-Zhou <zky314343421@gmail.com>
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
|
2026-05-01 07:21:28 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
214c35b031
|
[Intel GPU] Update xpu.Dockerfile to python 3.12 version (#23367)
|
2026-04-23 09:23:52 +08:00 |
|
 Polisetty V R K Jyothendra VarmaandMa Mingfei
|
7d2c11970c
|
[Intel GPU] Upgrade pytorch xpu version to 2.11 (#21908)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
|
2026-04-13 13:16:24 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
599cce4d82
|
[Intel GPU] import flash_attn functions from sgl_kernel only (#22438)
|
2026-04-10 15:10:00 +08:00 |
|
Polisetty V R K Jyothendra Varma
|
f0303fd07e
|
[Intel GPU] Enable DeepSeek R1 inference on XPU (#18461)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-03-29 22:35:59 -07:00 |
|
Polisetty V R K Jyothendra Varma
|
b2dd104ade
|
[Intel GPU] Upgrade pytorch xpu version to 2.10 (#20254)
Signed-off-by: P V R K Jyothendra Varma <polisetty.v.r.k.jyothendra.varma@intel.com>
|
2026-03-10 18:47:25 -07:00 |
|
Polisetty V R K Jyothendra Varma
|
71e4d3b6bc
|
[Intel GPU] fix import error to run DeepSeek-V2-Lite model with BF16 on XPU (#10858)
|
2026-01-29 21:53:53 -08:00 |
|
 Polisetty V R K Jyothendra VarmaandMa Mingfei
|
858dc80aff
|
[Intel GPU] fix device in DeepseekScalingRotaryEmbedding to run DeepSeek-V2-Lite BF16 on XPU (#10021)
Co-authored-by: Ma Mingfei <mingfei.ma@intel.com>
|
2026-01-29 21:21:38 -08:00 |
|