[llvm] [AMDGPU][SIMemoryLegalizer] Ensure LDS -> DMA hazards respect fences (PR #220288)

Krzysztof Drewniak via llvm-commits llvm-commits at lists.llvm.org
Fri Sep 11 14:50:15 PDT 2026


================
@@ -8140,8 +8140,20 @@ in table :ref:`amdgpu-amdhsa-memory-model-code-sequences-gfx6-gfx9-table`.
                                                              is being released.
 
                                                          2. buffer/global/flat_atomic
-     fence        release      - singlethread *none*     *none*
+     fence        release      - singlethread *none*     1. s_waitcnt lgkmcnt(0)
                                - wavefront
+
+                                                           - Emit only on GFX9,
+                                                             before
----------------
krzysz00 wrote:

... Ok, hold on, why would that only be needed on gfx9? Can't we get this same hazard on gfx1250 - or are we requiring people who mix sync and async ops to explicitly do a workgroup fence (and then promising to not demote it)?

https://github.com/llvm/llvm-project/pull/220288


More information about the llvm-commits mailing list