[llvm] [NVPTX] Add support for f32x2 mixed-precision add/sub (PR #221957)
Srinivasa Ravi via llvm-commits
llvm-commits at lists.llvm.org
Wed Sep 9 22:47:02 PDT 2026
================
@@ -2592,6 +2658,26 @@ foreach rnd = FPRoundingModes in {
(f32 (fpextend type:$a)),
(f32 (fneg f32:$b)), rnd_imm))]>,
Requires<[SM100]>;
+
+ foreach t = [F16X2RT, BF16X2RT] in {
+ // combineFAddWithNeg has already folded the fneg into a sub node.
+ defvar SubOp = !cast<SDNode>("sub_" # rnd);
+
+ def INT_NVVM_MIXED_SUB_ # rnd # _f32x2_ # t.PtxType :
----------------
Wolfram70 wrote:
Fixed, thanks!
https://github.com/llvm/llvm-project/pull/221957
More information about the llvm-commits
mailing list