[llvm] [AMDGPU] DAG Mutation to solve load bunching in outer product matrix multiplications (PR #203095)
Axel Sorenson via llvm-commits
llvm-commits at lists.llvm.org
Wed Sep 9 17:48:34 PDT 2026
================
----------------
axelcool1234 wrote:
Good catch. Instead of min/max, I went a little further and directly used each load's latency and reciprocal throughput when needed.
https://github.com/llvm/llvm-project/pull/203095
More information about the llvm-commits
mailing list