[llvm] [AMDGPU] Gate rootn(x, +-2) -> sqrt/rsqrt fold on nsz/ninf (PR #200578)
Matt Arsenault via llvm-commits
llvm-commits at lists.llvm.org
Sat May 30 10:54:45 PDT 2026
arsenm wrote:
> _rootn_fast keeps the unconditional fold
Note the __fast variants are supposed to have correct edge case handling, only reduced accuracy
https://github.com/llvm/llvm-project/pull/200578
More information about the llvm-commits
mailing list