[llvm] [NVPTX] Fold vector FMAs in the IR peephole pass (PR #224018)
Sjoerd Meijer via llvm-commits
llvm-commits at lists.llvm.org
Mon Sep 21 01:43:51 PDT 2026
sjoerdmeijer wrote:
Sorry, high level question first, but looking at the pattern, I was expecting DAGCombine to handle this case, and I think it does:
https://godbolt.org/z/4Ej1a5zM6
So, without this patch, we are already generating a `FMA_F32x2`.
I understand that with this patch, we generate a FMA earlier in the flow and that this may have some benefit, but that's why I wanted to ask about the motivation and if you do see benefits doing this?
https://github.com/llvm/llvm-project/pull/224018
More information about the llvm-commits
mailing list