[llvm-branch-commits] [llvm] [AMDGPU][InstCombine] Fold zero dot operands to accumulator (PR #225003)

Matt Arsenault via llvm-branch-commits llvm-branch-commits at lists.llvm.org
Mon Sep 21 05:31:12 PDT 2026


================
@@ -1974,6 +1974,10 @@ GCNTTIImpl::instCombineIntrinsic(InstCombiner &IC, IntrinsicInst &II) const {
   case Intrinsic::amdgcn_udot4:
   case Intrinsic::amdgcn_sdot8:
   case Intrinsic::amdgcn_udot8: {
+    if (match(II.getArgOperand(0), m_Zero()) ||
----------------
arsenm wrote:

If it doesn't work the canonicalize to RHS should be done separately 

https://github.com/llvm/llvm-project/pull/225003


More information about the llvm-branch-commits mailing list