[all-commits] [llvm/llvm-project] c2ba46: AMDGPU: Skip last corrections in afn f64 reciprocal
Matt Arsenault via All-commits
all-commits at lists.llvm.org
Wed Apr 22 12:27:45 PDT 2026
Branch: refs/heads/users/arsenm/amdgpu/skip-last-corrections-fast-f64-rcp
Home: https://github.com/llvm/llvm-project
Commit: c2ba462706b2fc889b2d831fa2570b1f7bc2f90c
https://github.com/llvm/llvm-project/commit/c2ba462706b2fc889b2d831fa2570b1f7bc2f90c
Author: Matt Arsenault <Matthew.Arsenault at amd.com>
Date: 2026-04-22 (Wed, 22 Apr 2026)
Changed paths:
M llvm/lib/Target/AMDGPU/AMDGPULegalizerInfo.cpp
M llvm/lib/Target/AMDGPU/SIISelLowering.cpp
M llvm/test/CodeGen/AMDGPU/GlobalISel/fdiv.f64.ll
M llvm/test/CodeGen/AMDGPU/fdiv.f64.ll
M llvm/test/CodeGen/AMDGPU/fneg-combines.new.ll
M llvm/test/CodeGen/AMDGPU/llvm.amdgcn.rcp.ll
M llvm/test/CodeGen/AMDGPU/rsq.f64.ll
Log Message:
-----------
AMDGPU: Skip last corrections in afn f64 reciprocal
Device libs has a fast reciprocal macro that is close
to the fast division expansion, but skips the last terms
compared to the full division.
The basic reciprocal handling has identical output to this
macro. The negative reciprocal case has different fneg placement
and smaller code size, but I believe should be the same.
To unsubscribe from these emails, change your notification settings at https://github.com/llvm/llvm-project/settings/notifications
More information about the All-commits
mailing list