chore(compiler): set SMP for Compiler ww28 SMP 2.4 - #6964
Draft
ronlieb wants to merge 2 commits into
Draft
Conversation
llvm 64fa6b89b8d5 spirv 28fa9b05 hipify fd6bd537 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-HOTSWAP JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2150 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2241 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2458 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2500 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2403 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2452 JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2491 CP's LCOMPILER-2491 64fa6b89b8d5 [MLIR][OpenMP] Restore debug-info fixup for declare target function. (#3654) LCOMPILER-2382 ef833c9b6390 Revert "Revert "device-libs: Move trig edge case to input (#1700)"" CP's: LCOMPILER-2452 4b2adbad0efb Pass `-O0` by default for HIP no-RDC compilations LCOMPILRE-2403 7f4d33f95b31 [AMDGPU] Fix instruction size of LDS-DMA buffer loads (#211302) LCOMPILER-2500 1d9f8c45183c Revert "[clang] Handle constructor closures with consteval default args (#203554) LCOMPILER-2458 6c3996842e34 Revert "[AMDGPU] Select the high-half 16-bit packing idiom to v_lshl_or_b32 (#206058) LCOMPILER-2382 21c4ca9d9a6e device-libs: Guard trig reduction quadrant index against NaN UB (#2752) "" 96c977988aa3 [DAGCombiner] Fold NaN-guard fptosi/fptoui select to saturating variant (#201435) LCOMPILER-2241 795861910c68 Revert "[AMDGPU] Enable runtime loop unrolling (#194924)" LCOMPILER-2150 973d77e112f4 Revert "device-libs: Move trig edge case to input (#1700)" CP's: f4e59fb506e5 (HEAD -> users/rlieberm/compiler-ww-28-SMP-2) [SLP][TTI][AMDGPU] Add TTI hook preferSLPInstCountCheck for per-target opt-out (#199696) 87335791b9a4 [AMDGPU] Fix CFI emission when scratch instructions are used to spill bf49a71aa259 [OffloadBundler] Bound compressed bundles by header size, not magic scan (#206745) 66ad914b6bb0 [Offload] Make compressed offload bundle header little-endian (#206744) COMGR HotSwap: a3bde579ffaf [Comgr] Force --offload-new-driver in unbundle-compressed spirv test b7759a1fd9a5 COMGR: scale hotswap rewrites for large code objects 02c9570d9283 COMGR: eliminate add-pc from far hotswap trampolines ef69a85ca584 [Comgr] fix program header offset in addKernelEntryTrampolineSymbols (#3357) 8f0b71632c44 [comgr][hotswap] make gfx1250 semantic rewrites fail-safe (#3346) 5b9ed4b0f168 [comgr][hotswap] preserve DS2 operand dependencies (#3345) b1688ed1f72d [comgr][hotswap] remove duplicate PoolVAddr declaration (#3359) 84f599a82d07 [comgr][hotswap] harden gfx1250 A0 trampoline foundations (#3344) e11ac472bfb8 [Comgr][Hotswap] Skip entry trampoline when prologue already has unclaused-VMEM workaround c3b8a0cda37b [Comgr] Assert hotswap RSRC1 field layout assumptions (#3072) d1a00eaba9e1 [comgr][hotswap] Test entry-trampoline rewrite is idempotent for stub symbols fbfaa80be39d [comgr][hotswap] Add <kernel>.stub symbols for entry trampolines 50106237933a [AMDGPU] comgr: wrap long comment lines to 80 columns 6932309fb284 [AMDGPU] comgr: trim stochastic rounding comment per review cb45344c9ec6 [AMDGPU] comgr: update SR FP8 modifier lit test for carry-propagation fix b02c0ccb0913 [AMDGPU] comgr: fix SR FP8 carry-propagation in E5M3 stochastic rounding bfef5feda161 AMDGPU: apply gfx1250 hotswap mask workarounds c1330e837683 AMDGPU: add strict hotswap rewrite plumbing 7e238120546e [Comgr][hotswap] Decline far s_add_pc_i64 trampolines (#3305) 046f71ecb959 [Comgr][hotswap] Append trampoline pool at a fresh vaddr instead of shifting post-.text addresses/symbols (#3252)
✅ All Checks Passed — Ready for Review
📖 Need help? See the Policy FAQ for details on every check and how to fix failures. |
|
🎉 All checks passed! This PR is ready for review. |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
set submodule pointer for Compiler ww28 SMP 2.4
llvm 64fa6b89b8d5
spirv 28fa9b05
hipify fd6bd537
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-HOTSWAP
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2150
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2241
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2458
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2500
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2403
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2452
JIRA ID : https://amd-hub.atlassian.net/browse/LCOMPILER-2491
CP's
LCOMPILER-2491 64fa6b89b8d5 [MLIR][OpenMP] Restore debug-info fixup for declare target function. (#3654)
LCOMPILER-2382 ef833c9b6390 Revert "Revert "device-libs: Move trig edge case to input (#1700)""
CP's:
LCOMPILER-2452 4b2adbad0efb Pass
-O0by default for HIP no-RDC compilationsLCOMPILRE-2403 7f4d33f95b31 [AMDGPU] Fix instruction size of LDS-DMA buffer loads (#211302)
LCOMPILER-2500 1d9f8c45183c Revert "[clang] Handle constructor closures with consteval default args (#203554)
LCOMPILER-2458 6c3996842e34 Revert "[AMDGPU] Select the high-half 16-bit packing idiom to v_lshl_or_b32 (#206058)
LCOMPILER-2382 21c4ca9d9a6e device-libs: Guard trig reduction quadrant index against NaN UB (#2752)
"" 96c977988aa3 [DAGCombiner] Fold NaN-guard fptosi/fptoui select to saturating variant (#201435)
LCOMPILER-2241 795861910c68 Revert "[AMDGPU] Enable runtime loop unrolling (#194924)"
LCOMPILER-2150 973d77e112f4 Revert "device-libs: Move trig edge case to input (#1700)"
CP's:
f4e59fb506e5 (HEAD -> users/rlieberm/compiler-ww-28-SMP-2) [SLP][TTI][AMDGPU] Add TTI hook preferSLPInstCountCheck for per-target opt-out (#199696)
87335791b9a4 [AMDGPU] Fix CFI emission when scratch instructions are used to spill
bf49a71aa259 [OffloadBundler] Bound compressed bundles by header size, not magic scan (#206745)
66ad914b6bb0 [Offload] Make compressed offload bundle header little-endian (#206744)
COMGR HotSwap:
a3bde579ffaf [Comgr] Force --offload-new-driver in unbundle-compressed spirv test
b7759a1fd9a5 COMGR: scale hotswap rewrites for large code objects
02c9570d9283 COMGR: eliminate add-pc from far hotswap trampolines
ef69a85ca584 [Comgr] fix program header offset in addKernelEntryTrampolineSymbols (#3357)
8f0b71632c44 [comgr][hotswap] make gfx1250 semantic rewrites fail-safe (#3346)
5b9ed4b0f168 [comgr][hotswap] preserve DS2 operand dependencies (#3345)
b1688ed1f72d [comgr][hotswap] remove duplicate PoolVAddr declaration (#3359)
84f599a82d07 [comgr][hotswap] harden gfx1250 A0 trampoline foundations (#3344)
e11ac472bfb8 [Comgr][Hotswap] Skip entry trampoline when prologue already has unclaused-VMEM workaround
c3b8a0cda37b [Comgr] Assert hotswap RSRC1 field layout assumptions (#3072)
d1a00eaba9e1 [comgr][hotswap] Test entry-trampoline rewrite is idempotent for stub symbols
fbfaa80be39d [comgr][hotswap] Add .stub symbols for entry trampolines
50106237933a [AMDGPU] comgr: wrap long comment lines to 80 columns
6932309fb284 [AMDGPU] comgr: trim stochastic rounding comment per review
cb45344c9ec6 [AMDGPU] comgr: update SR FP8 modifier lit test for carry-propagation fix
b02c0ccb0913 [AMDGPU] comgr: fix SR FP8 carry-propagation in E5M3 stochastic rounding
bfef5feda161 AMDGPU: apply gfx1250 hotswap mask workarounds
c1330e837683 AMDGPU: add strict hotswap rewrite plumbing
7e238120546e [Comgr][hotswap] Decline far s_add_pc_i64 trampolines (#3305)
046f71ecb959 [Comgr][hotswap] Append trampoline pool at a fresh vaddr instead of shifting post-.text addresses/symbols (#3252)