[llvm] [AMDGPU] DAG Mutation to solve load bunching in outer product matrix multiplications (PR #203095)
Lukas Sommer via llvm-commits
llvm-commits at lists.llvm.org
Thu Sep 10 07:14:51 PDT 2026
================
----------------
sommerlukas wrote:
So it's currently sorted by the smallest `MinPos` of each fragment's subloads, right? What is the secondary sorting criteria/tie breaker? Would it make sense to use the `MaxPos` or the number of registers in this fragment as tie-breaker, e.g., to expand the window for two smaller fragments instead of a single one?
https://github.com/llvm/llvm-project/pull/203095
More information about the llvm-commits
mailing list