[llvm] [AMDGPU] Split the true16 fcopysign pattern by uniformity (PR #226898)

Matt Arsenault via llvm-commits llvm-commits at lists.llvm.org
Mon Sep 28 14:00:37 PDT 2026


================
@@ -2464,10 +2464,16 @@ def : GCNPat <
 }
 let True16Predicate = UseRealTrue16Insts in {
 def : GCNPat <
-  (fcopysign fp16vt:$src0, fp16vt:$src1),
+  (UniformBinFrag<fcopysign> fp16vt:$src0, fp16vt:$src1),
+  (EXTRACT_SUBREG (V_BFI_B32_e64 (S_MOV_B32 (i32 0x00007fff)),
----------------
arsenm wrote:

I don't understand selecting a VALU instruction from a UniformBinFrag pattern. This is just going to force readfirstlane insertion by SIFixSGPRCopies and we definitely don't want that 

https://github.com/llvm/llvm-project/pull/226898


More information about the llvm-commits mailing list