[llvm] [AMDGPU][GlobalISel] Eliminate READANYLANE round trips in mixed build_vector/merge (PR #210449)

Petar Avramovic via llvm-commits llvm-commits at lists.llvm.org
Tue Jul 28 06:26:07 PDT 2026


https://github.com/petar-avramovic commented:

mostly looks fine.

Now this comment is not really related to this patch but
copies of large sgpr to vgpr are in general terrible. first (sgpr) reg alloc will try to have contiguous sgprs for copy src even that was not really necessary. Final ISA will end up with bunch of moves for pretty much no reason.
So I was thinking to try and move all large register sgpr to vgpr copies to vgpr by element copy. But then the problem is that there are many regressions in cases where user instr could have used sgpr. So don't know, maybe this kind of combine should happen much later, was thinking near SIFoldOperands

https://github.com/llvm/llvm-project/pull/210449


More information about the llvm-commits mailing list