[llvm-branch-commits] [llvm] [AMDGPU][InstCombine] Fold zero dot operands to accumulator (PR #225003)
Matt Arsenault via llvm-branch-commits
llvm-branch-commits at lists.llvm.org
Mon Sep 21 05:31:12 PDT 2026
================
@@ -1974,6 +1974,10 @@ GCNTTIImpl::instCombineIntrinsic(InstCombiner &IC, IntrinsicInst &II) const {
case Intrinsic::amdgcn_udot4:
case Intrinsic::amdgcn_sdot8:
case Intrinsic::amdgcn_udot8: {
+ if (match(II.getArgOperand(0), m_Zero()) ||
----------------
arsenm wrote:
If it doesn't work the canonicalize to RHS should be done separately
https://github.com/llvm/llvm-project/pull/225003
More information about the llvm-branch-commits
mailing list