swdev380865.ll - OpenGrok history log for /llvm-project/llvm/test/CodeGen/AMDGPU/swdev380865.ll

Revision (<<< Hide revision tags) (Show revision tags >>>)	Date	Author	Comments
Revision tags: llvmorg-18.1.8, llvmorg-18.1.7, llvmorg-18.1.6, llvmorg-18.1.5, llvmorg-18.1.4, llvmorg-18.1.3, llvmorg-18.1.2, llvmorg-18.1.1, llvmorg-18.1.0, llvmorg-18.1.0-rc4, llvmorg-18.1.0-rc3, llvmorg-18.1.0-rc2, llvmorg-18.1.0-rc1, llvmorg-19-init, llvmorg-17.0.6, llvmorg-17.0.5, llvmorg-17.0.4
# fe8335ba	30-Oct-2023	Stanislav Mekhanoshin <rampitec@users.noreply.github.com>	[AMDGPU] Select 64-bit imm moves if can be encoded as 32 bit operand (#70395) This allows folding of 64-bit operands if fit into 32-bit. Fixes https://github.com/llvm/llvm-project/issues/67781
# a0eb6b88	27-Oct-2023	Christudasan Devadasan <christudasan.devadasan@amd.com>	[AMDGPU] Try to fix the block prologs broken by RA inserted instructions (#69924) The insertion point determined by RA while attempting spills and liverange split at the beginning of a block goes w [AMDGPU] Try to fix the block prologs broken by RA inserted instructions (#69924) The insertion point determined by RA while attempting spills and liverange split at the beginning of a block goes wrong at times, and the newly inserted vector instructions are placed before the exec-mask restore instruction which is wrong. It occurs mainly due to the dependency on isBasicBlockPrologue that doesn't account early inserted instructions (spills and splits) during RA and causes the block prolog break. A better approach for deciding the insertion point should be worked out. For now, improving the helper function to consider all possible early insertions. This patch includes the spill instructions. The copies associated with liverange split should also be included in the block prolog. show more ...
Revision tags: llvmorg-17.0.3
# 7b3bbd83	09-Oct-2023	Jay Foad <jay.foad@amd.com>	Revert "[CodeGen] Really renumber slot indexes before register allocation (#67038)" This reverts commit 2501ae58e3bb9a70d279a56d7b3a0ed70a8a852c. Reverted due to various buildbot failures.
# 2501ae58	09-Oct-2023	Jay Foad <jay.foad@amd.com>	[CodeGen] Really renumber slot indexes before register allocation (#67038) PR #66334 tried to renumber slot indexes before register allocation, but the numbering was still affected by list entries [CodeGen] Really renumber slot indexes before register allocation (#67038) PR #66334 tried to renumber slot indexes before register allocation, but the numbering was still affected by list entries for instructions which had been erased. Fix this to make the register allocator's live range length heuristics even less dependent on the history of how instructions have been added to and removed from SlotIndexes's maps. show more ...
Revision tags: llvmorg-17.0.2
# e816c89c	02-Oct-2023	JP Lehr <JanPatrick.Lehr@amd.com>	Revert "InlineSpiller: Consider if all subranges are the same when avoiding redundant spills" This reverts commit d8127b2ba8a87a610851b9a462f2fc2526c36e37.
Revision tags: llvmorg-17.0.1, llvmorg-17.0.0, llvmorg-17.0.0-rc4, llvmorg-17.0.0-rc3, llvmorg-17.0.0-rc2, llvmorg-17.0.0-rc1, llvmorg-18-init, llvmorg-16.0.6, llvmorg-16.0.5, llvmorg-16.0.4, llvmorg-16.0.3, llvmorg-16.0.2, llvmorg-16.0.1
# d8127b2b	28-Mar-2023	Matt Arsenault <Matthew.Arsenault@amd.com>	InlineSpiller: Consider if all subranges are the same when avoiding redundant spills This avoids some redundant spills of subranges, and avoids a compile failure. This greatly reduces the numbers of InlineSpiller: Consider if all subranges are the same when avoiding redundant spills This avoids some redundant spills of subranges, and avoids a compile failure. This greatly reduces the numbers of spills in a loop. The main range is not informative when multiple instructions are needed to fully define a register. A common scenario is a lowered reg_sequence where every subregister is sequentially defined, but each def changes the main range's value number. If we look at specific lanes at the use index, we can see the value is actually the same. In this testcase, there are a large number of materialized 64-bit constant defs which are hoisted outside of the loop by MachineLICM. These are feeding REG_SEQUENCES, which is not considered rematerializable inside the loop. After coalescing, the split constant defs produce main ranges with an apparent phi def. There's no phi def if you look at each individual subrange, and only half of the register is really redefined to a constant. Fixes: SWDEV-380865 https://reviews.llvm.org/D147079 show more ...
# 7ac532ef	29-Sep-2023	Yashwant Singh <114088807+yashssh@users.noreply.github.com>	[AMDGPU] Introduce AMDGPU::SGPR_SPILL asm comment flag (#67091) Use this flag to give more context to implicit def comments in assembly. Reviewed on phabricator: https://reviews.llvm.org/D153754
# e0919b18	13-Sep-2023	Jay Foad <jay.foad@amd.com>	[CodeGen] Renumber slot indexes before register allocation (#66334) RegAllocGreedy uses SlotIndexes::getApproxInstrDistance to approximate the length of a live range for its heuristics. Renumbering [CodeGen] Renumber slot indexes before register allocation (#66334) RegAllocGreedy uses SlotIndexes::getApproxInstrDistance to approximate the length of a live range for its heuristics. Renumbering all slot indexes with the default instruction distance ensures that this estimate will be as accurate as possible, and will not depend on the history of how instructions have been added to and removed from SlotIndexes's maps. This also means that enabling -early-live-intervals, which runs the SlotIndexes analysis earlier, will not cause large amounts of churn due to different register allocator decisions. show more ...
# 4d42e8b5	28-Jul-2023	Matt Arsenault <Matthew.Arsenault@amd.com>	Reapply "[CodeGen]Allow targets to use target specific COPY instructions for live range splitting" This reverts commit a496c8be6e638ae58bb45f13113dbe3a4b7b23fd. The workaround in c26dfc81e254c78dc2 Reapply "[CodeGen]Allow targets to use target specific COPY instructions for live range splitting" This reverts commit a496c8be6e638ae58bb45f13113dbe3a4b7b23fd. The workaround in c26dfc81e254c78dc23579cf3d1336f77249e1f6 should work around the underlying problem with SUBREG_TO_REG. show more ...
# a496c8be	26-Jul-2023	Vitaly Buka <vitalybuka@google.com>	Revert "[CodeGen]Allow targets to use target specific COPY instructions for live range splitting" And dependent commits. Details in D150388. This reverts commit 825b7f0ca5f2211ec3c93139f98d1e24048 Revert "[CodeGen]Allow targets to use target specific COPY instructions for live range splitting" And dependent commits. Details in D150388. This reverts commit 825b7f0ca5f2211ec3c93139f98d1e24048c225c. This reverts commit 7a98f084c4d121244ef7286bc6503b6a181d446e. This reverts commit b4a62b1fa546312d882fa12dfdcd015177d66826. This reverts commit b7836d856206ec39509d42529f958c920368166b. No conflicts in the code, few tests had conflicts in autogenerated CHECKs: llvm/test/CodeGen/Thumb2/mve-float32regloops.ll llvm/test/CodeGen/AMDGPU/fix-frame-reg-in-custom-csr-spills.ll Reviewed By: alexfh Differential Revision: https://reviews.llvm.org/D156381 show more ...
# 7a98f084	17-May-2023	Christudasan Devadasan <Christudasan.Devadasan@amd.com>	[AMDGPU][SILowerSGPRSpills] Spill SGPRs to virtual VGPRs Currently, the custom SGPR spill lowering pass spills SGPRs into physical VGPR lanes and the remaining VGPRs are used by regalloc for vector [AMDGPU][SILowerSGPRSpills] Spill SGPRs to virtual VGPRs Currently, the custom SGPR spill lowering pass spills SGPRs into physical VGPR lanes and the remaining VGPRs are used by regalloc for vector regclass allocation. This imposes many restrictions that we ended up with unsuccessful SGPR spilling when there won't be enough VGPRs and we are forced to spill the leftover into memory during PEI. The custom spill handling during PEI has many edge cases and often breaks the compiler time to time. This patch implements spilling SGPRs into virtual VGPR lanes. Since we now split the register allocation for SGPRs and VGPRs, the virtual registers introduced for the spill lanes would get allocated automatically in the subsequent regalloc invocation for VGPRs. Spill to virtual registers will always be successful, even in the high-pressure situations, and hence it avoids most of the edge cases during PEI. We are now left with only the custom SGPR spills during PEI for special registers like the frame pointer which is an unproblematic case. Differential Revision: https://reviews.llvm.org/D124196 show more ...
# 7dcb9c0f	20-Mar-2023	Matt Arsenault <Matthew.Arsenault@amd.com>	InlineSpiller: Consider copy bundles when looking for snippet copies This was looking for full copies produced by SplitKit, but SplitKit introduces copy bundles if not all lanes are live. The scan f InlineSpiller: Consider copy bundles when looking for snippet copies This was looking for full copies produced by SplitKit, but SplitKit introduces copy bundles if not all lanes are live. The scan for uses needs to look at bundles, not individual instructions. This is a prerequisite to avoiding some redundant spills due to subregisters which will help avoid an allocation failure in a future patch. show more ...
Revision tags: llvmorg-16.0.0, llvmorg-16.0.0-rc4
# 051112a3	06-Mar-2023	Matt Arsenault <Matthew.Arsenault@amd.com>	AMDGPU: Add baseline test for SWDEV-380865 This demonstrates really bad rematerialization support for 64-bit constants which need to be split into 32-bit pieces.