llvm-capstone

mirror of https://github.com/capstone-engine/llvm-capstone.git synced 2024-11-30 09:01:19 +00:00

Author	SHA1	Message	Date
Fangrui Song	557299c9b6	[ELF][test] Test relocations referencing weak symbol, which is defined relative to a section discarded by /DISCARD/	2023-10-14 14:59:10 -07:00
Fangrui Song	4fb49f44fd	[ELF][test] Test relocations referencing symbols relative to sections discarded by /DISCARD/	2023-10-14 14:30:44 -07:00
Fangrui Song	60b3e05967	[ELF] Restore the --call-graph-profile-sort=hfsort default before #68638 The high time complexity of cache-directed sort is a real issue and is not appropriate as the default, at least for now (https://github.com/llvm/llvm-project/pull/68638#issuecomment-1760918891).	2023-10-12 22:58:42 -07:00
Kazu Hirata	4a0ccfa865	Use llvm::endianness::{big,little,native} (NFC) Note that llvm::support::endianness has been renamed to llvm::endianness while becoming an enum class as opposed to an enum. This patch replaces support::{big,little,native} with llvm::endianness::{big,little,native}.	2023-10-12 21:21:45 -07:00
Jacek Caban	85d0fbeaae	[lld] Don't allow -dynamicbase:no on ARM64EC.	2023-10-11 18:09:00 +02:00
Kazu Hirata	b8885926f8	Use llvm::endianness::{big,little,native} (NFC) Note that llvm::support::endianness has been renamed to llvm::endianness while becoming an enum class as opposed to an enum. This patch replaces llvm::support::{big,little,native} with llvm::endianness::{big,little,native}.	2023-10-10 22:54:51 -07:00
spupyrev	d5c1d735ad	[ELF] Making cdsort default for function reordering (#68638 ) CDSort function reordering outperforms the existing default heuristic ( hfsort/C^3) in terms of the performance of generated binaries while being (almost) as fast. Thus, the suggestion is to change the default. The speedup is up to 1.5% perf for large front-end binaries, and can be moderate/neutral for "small" benchmarks. High-level perf impact on two selected binaries: clang-10 binary (built with LTO+AutoFDO/CSSPGO): wins on top of C^3 in [0.3%..0.8%] rocksDB-8 binary (built with LTO+CSSPGO): wins on top of C^3 in [0.8%..1.5%] More detailed measurements on the clang binary is at [here](https://reviews.llvm.org/D152834#4445042)	2023-10-10 09:06:31 -07:00
Mitch Phillips	144d127bef	[lld] [MTE] Drop MTE globals for fully static executables, not ban (#68217 ) Integrating MTE globals on Android revealed a lot of cases where libraries are built as both archives and DSOs, and they're linked into fully static and dynamic executables respectively. MTE globals doesn't work for fully static executables. They need a dynamic loader to process the special R_AARCH64_RELATIVE relocation semantics with the encoded offset. Fully static executables that had out-of-bounds derived symbols (like 'int* foo_end = foo[16]') crash under MTE globals w/ static executables. So, LLD in its current form simply errors out when you try and compile a fully static executable that has a single MTE global variable in it. It seems like a much better idea to simply have LLD not do the special work for MTE globals in fully static contexts, and to drop any unnecessary metadata. This means that you can build archives with MTE globals and link them into both fully-static and dynamic executables.	2023-10-10 17:32:10 +02:00
Martin Storsjö	3548b79557	[LLD] [MinGW] Handle the --dll option (#68575 ) Treat this as an alias for the --shared option. In practice in GNU ld, they aren't exact aliases but there are small subtle differences in what happens and what doesn't, when they are set, but those differences are probably not intended by users who might be using the --dll option (which gets passed by -mdll in the compiler driver), and the differences are within the area where small details differ between LLD and GNU ld anyway.	2023-10-09 23:19:41 +03:00
Kazu Hirata	d7b18d5083	Use llvm::endianness{,::little,::native} (NFC) Now that llvm::support::endianness has been renamed to llvm::endianness, we can use the shorter form. This patch replaces llvm::support::endianness with llvm::endianness.	2023-10-09 00:54:47 -07:00
Ben Shi	488a62f86e	[lld][ELF][AVR] Add range check for R_AVR_13_PCREL (#67636 ) Some large AVR programs (for devices without long jump) may exceed 128KiB, and lld should give explicit errors other than generate wrong executables silently.	2023-10-06 18:20:03 +08:00
Alexandre Ganea	356139bd02	[LLD][COFF] Add support for `--time-trace` (#68236 ) This adds support for generating Chrome-tracing .json profile traces in the LLD COFF driver. Also add the necessary time scopes, so that the profile trace shows in great detail which tasks are executed. As an example, this is what we see when linking a Unreal Engine executable: ![image](https://github.com/llvm/llvm-project/assets/37383324/b2e26eb4-9d37-4cf9-b002-48b604e7dcb7)	2023-10-05 22:33:58 -04:00
Arthur Eubanks	9d6ec280fc	[lld/ELF] Don't relax R_X86_64_(REX_)GOTPCRELX when offset is too far For each R_X86_64_(REX_)GOTPCRELX relocation, check that the offset to the symbol is representable with 2^32 signed offset. If not, add a GOT entry for it and set its expr to R_GOT_PC so that we emit the GOT load instead of the relaxed lea. Do this in finalizeAddressDependentContent() where we iteratively attempt this (e.g. RISCV uses this for relaxation, ARM uses this to insert thunks). Decided not to do the opposite of inserting GOT entries initially and removing them when relaxable because removing GOT entries isn't simple. One drawback of this approach is that if we see any GOTPCRELX relocation, we'll create an empty .got even if it's not required in the end. Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D157020	2023-10-04 13:03:56 -07:00
Samira Bazuzi	8899b713ea	[LLD][COFF] Mark operator== const to avoid ambiguity in C++20. (#68119 ) C++20 will automatically generate an operator== with reversed operand order, which is ambiguous with the written operator== when one argument is marked const and the other isn't.	2023-10-04 11:39:20 -04:00
Martin Storsjö	503bc5f661	[LLD] [COFF] Fix handling of comdat .drectve sections (#68116 ) This can happen when manually emitting strings into .drectve sections with `__attribute__((section(".drectve")))`, which is a way to emulate `#pragma comment(linker, "...")` for mingw compilers, without requiring building with -fms-extensions. Normally, this doesn't generate any comdat, but if compiled with -fsanitize=address, this section does get turned into a comdat section. This fixes #67261. This issue can be seen as a regression; a change in the "lli" tool in 17.x triggers this case, if compiled with ASAN enabled, triggering this unsupported corner case in LLD. With this change, LLD can handle it.	2023-10-04 10:54:50 +03:00
Matheus Izvekov	d12b99a431	[lld][COFF][LTO] Implement /lldemit:asm option (#67079 ) With this new option, assembly code will be emited instead of object code. This is analogous to the `--lto-emit-asm` option in the ELF linker.	2023-10-04 01:24:54 +02:00
Sam Clegg	afe957ea95	[WebAssembly] Allow absolute symbols in the linking section (symbol table) (#67493 ) Fixes a crash in `-Wl,-emit-relocs` where the linker was not able to write linker-synthetic absolute symbols to the symbol table. This change adds a new symbol flag (`WASM_SYMBOL_ABS`), which means that the symbol's offset is absolute and not relative to a given segment. Such symbols include `__stack_low` and `__stack_low`. Note that wasm object files never contains such symbols, only binaries linked with `-Wl,-emit-relocs`. Fixes: #67111	2023-10-03 13:16:16 -07:00
Alexandre Ganea	a2ef046a2d	[LLD][ELF] Import `ObjFile::importCmseSymbols` at call site (#68025 ) Before this patch, with MSVC I was seeing: ``` [304/334] Building CXX object tools\lld\ELF\CMakeFiles\lldELF.dir\InputFiles.cpp.obj C:\git\llvm-project\lld\ELF\InputFiles.h(327): warning C4661: 'void lld:🧝:ObjFile<llvm::object::ELF32LE>::importCmseSymbols(void)': no suitable definition provided for explicit template instantiation request C:\git\llvm-project\lld\ELF\InputFiles.h(291): note: see declaration of 'lld:🧝:ObjFile<llvm::object::ELF32LE>::importCmseSymbols' C:\git\llvm-project\lld\ELF\InputFiles.h(327): warning C4661: 'void lld:🧝:ObjFile<llvm::object::ELF32LE>::redirectCmseSymbols(void)': no suitable definition provided for explicit template instantiation request C:\git\llvm-project\lld\ELF\InputFiles.h(292): note: see declaration of 'lld:🧝:ObjFile<llvm::object::ELF32LE>::redirectCmseSymbols' ``` This patch removes `redirectCmseSymbols` which is not defined. And it imports `importCmseSymbols` in InputFiles.cpp, because it is already explicitly instantiated in ARM.cpp.	2023-10-03 10:12:09 -04:00
simpal01	3cde1d8000	[ELF] Handle relocations in synthetic .eh_frame with a non-zero offset within the output section (#65966 ) When the .eh_frame section is placed at a non-zero offset within its output section, the relocation value within .eh_frame are computed incorrectly. We had similar issue in .ARM.exidx section and it has been fixed already in https://reviews.llvm.org/D148033. While applying the relocation using S+A-P, the value of P (the location of the relocation) is getting wrong. P is: P = SecAddr + rel.offset, But SecAddr points to the starting address of the outputsection rather than the starting address of the eh frame section within that output section. This issue affects all targets which generates .eh_frame section. Hence fixing in all the corresponding targets it affecting.	2023-10-03 10:20:14 +01:00
Matheus Izvekov	52d35345b2	[LLD/COFF] Fix link failure due to missing component. This was broken by `3923e61b96` We should link BitWriter into LLD/COFF since now it uses `WriteBitcodeToFile`.	2023-10-03 00:33:36 +02:00
Matheus Izvekov	3923e61b96	[lld][COFF][LTO] Implement /lldemit:llvm option (#66964 ) With this new option, bitcode will be emited instead of object code. This is analogous to the `--plugin-opt=emit-llvm` option in the ELF linker.	2023-10-02 22:54:43 +02:00
Alexandre Ganea	dd8cfdeef9	[LLD][COFF] Delete unused field `DebugSHandler::source`	2023-10-02 16:22:54 -04:00
Alexandre Ganea	52d9f768b6	[LLD][COFF] Remove unused `DebugSHandler::recordStringTableReferences`	2023-10-02 16:22:54 -04:00
Martin Storsjö	f906fd53b5	[LLD] [COFF] Restore the current dir as the first entry in the search path (#67857 ) Before `af744f0b84`, the first entry among the search paths was the empty string, indicating searching in (or starting from) the current directory. After `af744f0b84`, the toolchain/clang specific lib directories were added at the head of the search path. This would cause lookups of literal file names or relative paths to match paths in the toolchain, if there are coincidental files with similar names there, even if they would be find in the current directory as well. Change addClangLibSearchPaths to append to the list like all other operations on searchPaths - but move the invocation of the function to the right place in the sequence. This fixes #67779.	2023-10-02 13:47:30 +03:00
Martin Storsjö	7d7d9e462a	[LLD] [COFF] Clarify -print-search-path for the empty string element (#67856 ) Also switch the test case to use -NEXT to strictly match all lines.	2023-10-02 13:43:56 +03:00
David Spickett	ffc67bb360	Revert "[Flang] [FlangRT] Introduce FlangRT project as solution to Flang's runtime LLVM integration" This reverts commit `6403287eff`. This is failing on all but 1 of Linaro's flang builders. CMake Error at /home/tcwg-buildbot/worker/clang-aarch64-full-2stage/llvm/flang-rt/unittests/CMakeLists.txt:37 (message): Target llvm_gtest not found.	2023-10-02 09:02:05 +00:00
Paul Scoropan	6403287eff	[Flang] [FlangRT] Introduce FlangRT project as solution to Flang's runtime LLVM integration See discourse thread https://discourse.llvm.org/t/rfc-support-cmake-option-to-control-link-type-built-for-flang-runtime-libraries/71602/18 for full details. Flang-rt is the new library target for the flang runtime libraries. It builds the Flang-rt library (which contains the sources of FortranRuntime and FortranDecimal) and the Fortran_main library. See documentation in this patch for detailed description (flang-rt/docs/GettingStarted.md). This patch aims to: - integrate Flang's runtime into existing llvm infrasturcture so that Flang's runtime can be built similarly to other runtimes via the runtimes target or via the llvm target as an enabled runtime - decouple the FortranDecimal library sources that were used by both compiler and runtime so that different build configurations can be applied for compiler vs runtime - add support for running flang-rt testsuites, which were created by migrating relevant tests from `flang/test` and `flang/unittest` to `flang-rt/test` and `flang-rt/unittest`, using a new `check-flang-rt` target. - provide documentation on how to build and use the new FlangRT runtime Reviewed By: DanielCChen Differential Revision: https://reviews.llvm.org/D154869	2023-09-30 12:35:33 -04:00
Nico Weber	94d9455157	[lld] Fix REQUIRES line in new test Noticed by chapuni: https://github.com/llvm/llvm-project/pull/67445#pullrequestreview-1645422024 Fixes a test failure if X86 isn't in LLVM_TARGETS_TO_BUILD.	2023-09-27 19:55:16 -04:00
Matheus Izvekov	8ff77a8f04	[NFC][LLD] Refactor some copy-paste into the Common library (#67598 )	2023-09-28 00:06:48 +02:00
Arthur Eubanks	706b04e79b	[lld/ELF][test] Add test for .got too far away for unrelaxed R_X86_64_(REX_)GOTPCRELX (#67609 )	2023-09-27 14:52:51 -07:00
kyulee-com	e04bf91111	[lld-macho][NFC] Remove redundant checks (#67450 ) `ignoreAutoLinkOptions` checks run both in `parseLCLinkerOptions` and `resolveLCLinkerOptions`. Convert the latter check to an assert.	2023-09-26 14:59:18 -07:00
Nico Weber	210e8984fe	[lld/mac] Resolve defined symbols before undefined symbols in bitcode (#67445 ) Ports https://reviews.llvm.org/D106293 to bitcode, or https://github.com/llvm/llvm-project/commit/bd448f01a6 from ELF to MachO. See also #59162 for some vaguely related discussion.	2023-09-26 15:31:51 -04:00
spupyrev	904b3f66f5	[ELF] A new code layout algorithm for function reordering [3a/3] We are brining a new algorithm for function layout (reordering) based on the call graph (extracted from a profile data). The algorithm is an improvement of top of a known heuristic, C^3. It tries to co-locate hot and frequently executed together functions in the resulting ordering. Unlike C^3, it explores a larger search space and have an objective closely tied to the performance of instruction and i-TLB caches. Hence, the name CDS = Cache-Directed Sort. The algorithm can be used at the linking or post-linking (e.g., BOLT) stage. Refer to https://reviews.llvm.org/D152834 for the actual implementation of the reordering algorithm. This diff adds a linker option to replace the existing C^3 heuristic with CDS. The new behavior can be turned on by passing "--use-cache-directed-sort". (the plan is to make it default in a next diff) Perf-impact clang-10 binary (built with LTO+AutoFDO/CSSPGO): wins on top of C^3 in [0.3%..0.8%] rocksDB-8 binary (built with LTO+CSSPGO): wins on top of C^3 in [0.8%..1.5%] Note that function layout affects the perf the most on older machines (with smaller instruction/iTLB caches) and when huge pages are not enabled. The impact on newer processors with huge pages enabled is likely neutral/minor. Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D152840	2023-09-26 06:24:34 -07:00
Fangrui Song	8c556b7e2b	[ELF] Change --call-graph-profile-sort to accept an argument Change the FF form --call-graph-profile-sort to --call-graph-profile-sort={none,hfsort}. This will be extended to support llvm/lib/Transforms/Utils/CodeLayout.cpp. --call-graph-profile-sort is not used in the wild but --no-call-graph-profile-sort is (Chromium). Make --no-call-graph-profile-sort an alias for --call-graph-profile-sort=none. Reviewed By: rahmanl Differential Revision: https://reviews.llvm.org/D159544	2023-09-25 09:49:40 -07:00
David Truby	247b7d0684	[cmake] Add LLVM_FORCE_VC_REVISION option (#67125 ) This patch adds a LLVM_FORCE_VC_REVISION option to force a custom VC revision to be included instead of trying to fetch one from a git command. This is helpful in environments where git is not available or is non-functional but the vc revision is available through some other means.	2023-09-25 14:32:52 +01:00
Fangrui Song	c2cd66db74	[ELF][test] Improve GOTPCRELX tests	2023-09-21 15:59:14 -07:00
Matheus Izvekov	a5e280bc6b	[NFC] clang-format lld/COFF/Driver.cpp and lld/Common/Filesystem.cpp In order to reduce noise for a MR.	2023-09-21 13:19:04 +02:00
Pierre van Houtryve	fe2f67e4ba	[AMDGPU] Remove Code Object V2 (#65715 ) Code Object V2 has been deprecated for more than a year now. We can safely remove it from LLVM. - [clang] Remove support for the `-mcode-object-version=2` option. - [lld] Remove/refactor tests that were still using COV2 - [llvm] Update AMDGPUUsage.rst - Code Object V2 docs are left for informational purposes because those code objects may still be supported by the runtime/loaders for a while. - [AMDGPU] Remove COV2 emission capabilities. - [AMDGPU] Remove `MetadataStreamerYamlV2` which was only used by COV2 - [AMDGPU] Update all tests that were still using COV2 - They are either deleted or ported directly to code object v4 (as v3 is also planned to be removed soon).	2023-09-21 12:00:45 +02:00
Fangrui Song	f5b42eaadb	[ELF] -r --compress-debug-sections: update implicit addends for .rel.debug_* referencing STT_SECTION symbols (#66804 ) https://reviews.llvm.org/D48929 updated addends for non-SHF_ALLOC sections relocated by REL for -r links, but the patch did not update the addends when --compress-debug-sections={zlib,zstd} is used (#66738). https://reviews.llvm.org/D116946 handled tombstone values in debug sections in relocatable links. As a side effect, both relocateNonAllocForRelocatable (using `sec->relocations`) and relocatenonNonAlloc (using raw REL/RELA) may run. Actually, we can adjust the condition in relocatenonAlloc to completely replace relocateNonAllocForRelocatable. This patch implements this idea and fixes #66738. As relocateNonAlloc processes the raw relocations like copyRelocations() does, the condition `if (config->relocatable && type != target.noneRel)` in `copyRelocations` (commit `08d6a3f133`, modified by https://reviews.llvm.org/D62052) can be made specific to SHF_ALLOC sections. As a side effect, we can now report diagnostics for PC-relative relocations for -r. This is a less useful diagnostic that is not worth too much code. As https://github.com/ClangBuiltLinux/linux/issues/1937 has violations, just suppress the warning for -r. Tested by commit `561b98f9e0`.	2023-09-20 14:50:13 -07:00
Fangrui Song	561b98f9e0	[ELF][test] Improve non-abs-reloc.s to test non-STT_SECTION local symbol	2023-09-20 14:40:32 -07:00
Fangrui Song	0de0b6dded	[ELF] Postpone "unable to move location counter backward" error (#66854 ) The size of .ARM.exidx may shrink across `assignAddress` calls. It is possible that the initial iteration has a larger location counter, causing `__code_size = __code_end - .; osec : { . += __code_size; }` to report an error, while the error would have been suppressed for subsequent `assignAddress` iterations. Other sections like .relr.dyn may change sizes across `assignAddress` calls as well. However, their initial size is zero, so it is difficiult to trigger a similar error. Similar to https://reviews.llvm.org/D152170, postpone the error reporting. Fix #66836. While here, add more information to the error message.	2023-09-20 09:06:45 -07:00
Fangrui Song	309d1c43bd	[ELF][test] Add a test to demonstrate #66836	2023-09-20 09:04:41 -07:00
Fangrui Song	a317afaf00	[ELF][test] Improve -r tests for local symbols	2023-09-19 23:44:55 -07:00
Fangrui Song	678c1f142c	[ELF] Remove a R_ARM_PCA special case from relocateNonAlloc https://reviews.llvm.org/D75042 added a special case about R_ARM_PCA to relocateNonAlloc. This is untested and actually unused in the wild.	2023-09-19 21:04:50 -07:00
Fangrui Song	a495b2f8cb	[ELF][test] Improve tests about non-SHF_ALLOC sections relocated by non-ABS relocations	2023-09-19 21:02:27 -07:00
Haojian Wu	71c83fb8b6	[LLD] Improve the lit tests added by `272bd6f9cc` The `cp` command will copy the permission bits from the original file to the new one, which will cause permission denied (no written access) for the following "echo" command in some system. Switch to use `cat` which is more robust.	2023-09-19 13:47:08 +02:00
Fangrui Song	345f532f3f	[ELF][test] Improve relocations referencing STT_SECTION tests for -r Also add a --compress-debug-sections=zlib test to demonstrate issue #66738	2023-09-18 22:46:56 -07:00
modimo	272bd6f9cc	[WPD][LLD] Add option to validate RTTI is enabled on all native types and prevent devirtualization on types with native RTTI Discussion about this approach: https://discourse.llvm.org/t/rfc-safer-whole-program-class-hierarchy-analysis/65144/18 When enabling WPD in an environment where native binaries are present, types we want to optimize can be derived from inside these native files and devirtualizing them can lead to correctness issues. RTTI can be used as a way to determine all such types in native files and exclude them from WPD providing a safe checked way to enable WPD. The approach is: 1. In the linker, identify if RTTI is available for all native types. If not, under `--lto-validate-all-vtables-have-type-infos` `--lto-whole-program-visibility` is automatically disabled. This is done by examining all .symtab symbols in object files and .dynsym symbols in DSOs for vtable (_ZTV) and typeinfo (_ZTI) symbols and ensuring there's always a match for every vtable symbol. 2. During thinlink, if `--lto-validate-all-vtables-have-type-infos` is set and RTTI is available for all native types, identify all typename (_ZTS) symbols via their corresponding typeinfo (_ZTI) symbols that are used natively or outside of our summary and exclude them from WPD. Testing: ninja check-all large Meta service that uses boost, glog and libstdc++.so runs successfully with WPD via --lto-whole-program-visibility. Previously, native types in boost caused incorrect devirtualization that led to crashes. Reviewed By: MaskRay, tejohnson Differential Revision: https://reviews.llvm.org/D155659	2023-09-18 15:51:49 -07:00
Sam Clegg	7bac0bc115	[lld][WebAssembly] Improve error message on adding LTO object post-LTO. NFC (#66688 ) We have been getting errors from emscripten users where including the name of the symbol that triggered the inclusion would be useful in the diagnosis. e.g: https://github.com/emscripten-core/emscripten/issues/20275	2023-09-18 14:11:49 -07:00
Fangrui Song	bd288ebf5f	[ELF] Make -t work with --format=binary This is natural and matches GNU ld.	2023-09-16 00:41:23 -07:00
Fangrui Song	71d7e69c56	[ELF] Implement getImplicitAddend and enable checkDynamicRelocsDefault for Hexagon	2023-09-15 23:10:17 -07:00
Fangrui Song	df979228ef	[ELF] Implement getImplicitAddend and enable checkDynamicRelocsDefault for AMDGPU	2023-09-15 22:59:08 -07:00
Fangrui Song	0cbe49eade	[ELF] Implement getImplicitAddend and enable checkDynamicRelocsDefault for PPC32	2023-09-15 22:49:18 -07:00
Fangrui Song	1b65b159da	[ELF] Enable checkDynamicRelocsDefault for PPC64 .plt and .branch_lt have the type of SHT_NOBITS and may be relocated by dynamic relocations with non-zero addends. They should be skipped for the --check-dynamic-relocations check, as --apply-dynamic-relocs does not apply. A side effect is that -z rel does not work for the two sections. Added two --apply-dynamic-relocs --check-dynamic-relocations tests. Also checked linking a PPC64 clang.	2023-09-15 22:38:18 -07:00
Fangrui Song	d5682e5777	[ELF] checkDynamicRelocsDefault: invert the condition Fixes: `64c9dbbf4e`	2023-09-15 22:25:09 -07:00
Fangrui Song	64c9dbbf4e	[ELF] checkDynamicRelocsDefault: invert the condition Most targets support --check-dynamic-relocations --apply-dynamic-relocs, so it makes sense to invert the condition. We want new ports to support getImplicitAddend.	2023-09-15 21:06:19 -07:00
Fangrui Song	1bd5df7af6	[ELF] Correct a comment about ^=. NFC GNU ld added ^= support in July 2023.	2023-09-15 17:52:48 -07:00
Fangrui Song	9eb1f2d0ac	[ELF] Remove a special case from ExprValue::getSectionOffset. NFC The special cases added by https://reviews.llvm.org/D42911 (defsym-reserved-syms.s) are unneeded after https://reviews.llvm.org/D66717	2023-09-15 17:45:52 -07:00
Saleem Abdulrasool	401296f697	Object: account for short output names (#66532 ) The import library thunk name suffix uses the stem of the file. We currently would attempt to trim the suffix by dropping the trailing 4 characters (under the assumption that the output name was `.lib`). This now uses the `llvm::sys::path` API for computing the stem. This avoids an assertion failure when the name is less the 4 characters and assertions are enabled.	2023-09-15 12:58:23 -07:00
Aaron Ballman	cf3b12c795	Fix lld Sphinx build This addresses issues found by: https://lab.llvm.org/buildbot/#/builders/69/builds/41718	2023-09-15 15:15:52 -04:00
Julien Schueller	715bc51190	[LLD] [MinGW] Ignore --sort-common option (#66336 ) like done in the ELF side this would allow to use archlinux default mingw flags: `-Wl,-O1,--sort-common,--as-needed -fstack-protector` (on archlinux packages use the GNU linker by default)	2023-09-15 11:13:27 +03:00
Douglas Yung	72bbac4738	Mark test added in `5a58e98` as requiring ppc, not x86 since it tries to use the powerpc64le target.	2023-09-14 14:47:20 -07:00
Arthur Eubanks	0a1aa6cda2	[NFC][CodeGen] Change CodeGenOpt::Level/CodeGenFileType into enum classes (#66295 ) This will make it easy for callers to see issues with and fix up calls to createTargetMachine after a future change to the params of TargetMachine. This matches other nearby enums. For downstream users, this should be a fairly straightforward replacement, e.g. s/CodeGenOpt::Aggressive/CodeGenOptLevel::Aggressive or s/CGFT_/CodeGenFileType::	2023-09-14 14:10:14 -07:00
Fangrui Song	5a58e98c20	[ELF] Align the end of PT_GNU_RELRO associated PT_LOAD to a common-page-size boundary (#66042 ) Close #57618: currently we align the end of PT_GNU_RELRO to a common-page-size boundary, but do not align the end of the associated PT_LOAD. This is benign when runtime_page_size >= common-page-size. However, when runtime_page_size < common-page-size, it is possible that `alignUp(end(PT_LOAD), page_size) < alignDown(end(PT_GNU_RELRO), page_size)`. In this case, rtld's mprotect call for PT_GNU_RELRO will apply to unmapped regions and lead to an error, e.g. ``` error while loading shared libraries: cannot apply additional memory protection after relocation: Cannot allocate memory ``` To fix the issue, add a padding section .relro_padding like mold, which is contained in the PT_GNU_RELRO segment and the associated PT_LOAD segment. The section also prevents strip from corrupting PT_LOAD program headers. .relro_padding has the largest `sortRank` among RELRO sections. Therefore, it is naturally placed at the end of `PT_GNU_RELRO` segment in the absence of `PHDRS`/`SECTIONS` commands. In the presence of `SECTIONS` commands, we place .relro_padding immediately before a symbol assignment using DATA_SEGMENT_RELRO_END (see also https://reviews.llvm.org/D124656), if present. DATA_SEGMENT_RELRO_END is changed to align to max-page-size instead of common-page-size. Some edge cases worth mentioning: * ppc64-toc-addis-nop.s: when PHDRS is present, do not append .relro_padding * avoid-empty-program-headers.s: when the only RELRO section is .tbss, it is not part of PT_LOAD segment, therefore we do not append .relro_padding. --- Close #65002: GNU ld from 2.39 onwards aligns the end of PT_GNU_RELRO to a max-page-size boundary (https://sourceware.org/PR28824) so that the last page is protected even if runtime_page_size > common-page-size. In my opinion, losing protection for the last page when the runtime page size is larger than common-page-size is not really an issue. Double mapping a page of up to max-common-page for the protection could cause undesired VM waste. Internally we had users complaining about 2MiB max-page-size applying to shared objects. Therefore, the end of .relro_padding is padded to a common-page-size boundary. Users who are really anxious can set common-page-size to match their runtime page size. --- 17 tests need updating as there are lots of change detectors.	2023-09-14 10:33:11 -07:00
Fangrui Song	e057d8973c	[ELF][PPC64] Use the regular placement for .branch_lt The currently rule places .branch_lt after .data, which does not make sense. The original contributor probably wanted to place .branch_lt before .got/.toc, but failed to notice that .got/.toc are RELRO and placed earlier. Remove the special case so that .branch_lt is actually closer to .toc, alleviating the distance issue.	2023-09-13 19:15:42 -07:00
Fangrui Song	a700ed6909	[ELF][test] Test the order of PPC64 .got, .toc, and .branch_lt	2023-09-13 19:08:22 -07:00
Saiyedul Islam	466a8149b3	Revert "[AMDGPU] Make default AMDHSA Code Object Version to be 5 (#65410 )" (#66060 ) This reverts commit `0a8d17e79b`.	2023-09-12 15:13:59 +05:30
Saiyedul Islam	0a8d17e79b	[AMDGPU] Make default AMDHSA Code Object Version to be 5 (#65410 ) Also update LIT tests and docs. For more details, see https://llvm.org/docs/AMDGPUUsage.html#code-object-v5-metadata Reviewed By: arsenm, jhuber6 Github PR: #65410 Differential Revision: https://reviews.llvm.org/D129818	2023-09-12 13:53:31 +05:30
Fangrui Song	c5ccae4f18	[ELF][test] Make tests less sensitive of addresses/number of sections	2023-09-11 21:01:36 -07:00
Fangrui Song	d1b418f552	[ELF] Reset two member variables in Ctx::reset Otherwise they are dangling if lldMain is called more than once.	2023-09-11 13:51:01 -07:00
Vy Nguyen	595cd45a66	[lld-macho][nfc]Add bounds on sections and subsections check before attempting to dereferencing iterators. Runnign some tests with asan built of LLD would throw errors similar to the following: AddressSanitizer:DEADLYSIGNAL #0 0x55d8e6da5df7 in operator() /mnt/ssd/repo/lld/llvm-project/lld/MachO/Arch/ARM64.cpp:612 #1 0x55d8e6daa514 in operator() /mnt/ssd/repo/lld/llvm-project/lld/MachO/Arch/ARM64.cpp:650 Differential Revision: https://reviews.llvm.org/D157027	2023-09-11 14:19:25 -04:00
Fangrui Song	65a15a56d5	[ELF] Respect orders of symbol assignments and DEFINED (#65866 ) Fix #64600: the currently implementation is minimal (see https://reviews.llvm.org/D83758), and an assignment like `__TEXT_REGION_ORIGIN__ = DEFINED(__TEXT_REGION_ORIGIN__) ? __TEXT_REGION_ORIGIN__ : 0;` (used by avr-ld[1]) leads to a value of zero (default value in `declareSymbol`), which is unexpected. Assign orders to symbol assignments and references so that for a script-defined symbol, the `DEFINED` results match users' expectation. I am unclear about GNU ld's exact behavior, but this hopefully matches its behavior in the majority of cases. [1]: https://sourceware.org/git/?p=binutils-gdb.git;a=blob;f=ld/scripttempl/avr.sc	2023-09-11 10:54:49 -07:00
Ellis Hoag	30e688e6d0	[lld][MachO] Add option to suppress mismatch profile errors (#65551 ) Both ELF and COFF support `--no-lto-pgo-warn-mismatch` in https://reviews.llvm.org/D104431 to suppress warnings due to mismatching profile hashes. As profiles go stale, it becomes likely that some function's CFGs will change so that their profiles can no longer be used. This commit adds the linker option `--no-pgo-warn-mismatch` to suppress these warnings. Note that we do have the LLVM backend flag `no-pgo-warn-mismatch` `3df1a64eba/llvm/lib/Transforms/Instrumentation/PGOInstrumentation.cpp (L210)` but that is set to true by default during LTO `3df1a64eba/llvm/include/llvm/LTO/Config.h (L76-L77)`	2023-09-11 09:13:55 -07:00
Fangrui Song	fcc761bd0e	[ELF][test] Make tests less sensitive to addresses/number of sections	2023-09-11 00:26:22 -07:00
Fangrui Song	21ac457f3f	[ELF] Priorize the last catch-all pattern in version scripts When there are multiple catch-all patterns (i.e. a single `*`), GNU ld and gold select the last pattern. Match their behavior. This change was inspired by a correction made by Michael Kerrisk to a blog post I wrote at https://maskray.me/blog/2020-11-26-all-about-symbol-versioning , following the current lld rules. Note: GNU ld prefers global: patterns to local: patterns, which might seem awkward (https://www.airs.com/blog/archives/300). gold doesn't follow this behavior, and we do not either.	2023-09-09 23:47:01 -07:00
Job Noorman	649cac3b62	[ELF][RISCV] Implement --emit-relocs with relaxation Linker relaxation may change relocations (offsets and types). However, when --emit-relocs is used, relocations are simply copied from the input section causing a mismatch with the corresponding (relaxed) code section. This patch fixes this as follows: for non-relocatable RISC-V binaries, `InputSection::copyRelocations` reads relocations from the relocated section's `relocations` array (since this gets updated by the relaxation code). For all other cases, relocations are read from the input section directly as before. In order to reuse as much code as possible, and to keep the diff small, the original `InputSection::copyRelocations` is changed to accept the relocations as a range of `Relocation` objects. This means that, in the general case when reading from the input section, raw relocations need to be converted to `Relocation`s first, which introduces quite a bit of boiler plate. It also means there's a slight code size increase due to the extra instantiations of `copyRelocations` (for both range types). Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D159082	2023-09-09 12:06:39 +02:00
Vy Nguyen	84a2155921	[lld-macho]Limit "cannot-export-hidden-symbol" warnings to only 3 to avoid crowding logs. Details: We often use wildcard symbols in the exported_symbols list, and sometimes they match autohide symbols, which triggers these "cannot export hidden symbols" warnings that can be a bit noisy. It'd be more user-friendly if LLD could truncate these. Differential Revision: https://reviews.llvm.org/D159095	2023-09-07 13:50:55 -04:00
aeubanks	a07d4c0365	[lld/ELF,gold] Remove transitionary opaque pointer flags (#65529 ) This was only useful during the transition when mixing non-opaque-pointer and opaque-pointer IR, now everything uses opaque pointers.	2023-09-06 15:07:37 -07:00
Daniel Paoliello	f2f36c9b29	Emit line numbers in CodeView for trailing (after `ret`) blocks from inlined functions Issue Details: When building up line information for CodeView debug info, LLVM attempts to gather the "range" of instructions within a function as these are printed together in a single record. If there is an inlined function, then those lines are attributed to the original function to enable generating `S_INLINESITE` records. However, this thus requires there to be instructions from the inlining function after the inlined function otherwise the instruction range would not include the inlined function. Fix Details: Include any inlined functions when finding the extent of a function in `getFunctionLineEntries` Reviewed By: rnk Differential Revision: https://reviews.llvm.org/D159226	2023-09-06 10:19:30 -07:00
Teresa Johnson	bbe8cd1333	[LTO] Remove module id from summary index The module paths string table mapped to both an id sequentially assigned during LTO linking, and the module hash. The former is leftover from before the module hash was added for caching and subsequently replaced use of the module id when renaming promoted symbols (to avoid affects due to link order changes). The sequentially assigned module id was not removed, however, as it was still a convenience when serializing to/from bitcode and assembly. This patch removes the module id from this table, since it isn't strictly needed and can lead to confusion on when it is appropriate to use (e.g. see fix in D156525). It also takes a (likely not significant) amount of overhead. Where an integer module id is needed (e.g. bitcode writing), one is assigned on the fly. There are a couple of test changes since the paths are now sorted alphanumerically when assigning ids on the fly during assembly writing, in order to ensure deterministic behavior. Differential Revision: https://reviews.llvm.org/D156730	2023-09-01 13:43:08 -07:00
Amy Kwan	698b45aa90	[PowerPC][lld] Account for additional X-Forms -> D-Form/DS-Forms load/stores when relaxing initial-exec to local-exec D153645 added additional X-Form load/stores that can be generated for TLS accesses. However, these added instructions have not been accounted for in lld. As a result, lld does not know how to handle them and cannot relax initial-exec to local-exec when the initial-exec sequence contains these additional load/stores. This patch aims to resolve https://github.com/llvm/llvm-project/issues/64424. Differential Revision: https://reviews.llvm.org/D158197	2023-08-31 08:45:10 -05:00
Takuya Shimizu	01b88dd66d	[NFC] Remove unused variables declared in conditions D152495 makes clang warn on unused variables that are declared in conditions like `if (int var = init) {}` This patch is an NFC fix to suppress the new warning in llvm,clang,lld builds to pass CI in the above patch. Differential Revision: https://reviews.llvm.org/D158016	2023-08-30 10:05:06 +09:00
Martin Storsjö	0b51e64830	[LLD] [MinGW] Hook up more LTO options Many of these options can be passed to the linker by the Clang driver based on other options passed to Clang, after `a23bf1786b`. Before commit `5c92c9f34a`, these were ignored by lld, but now we're erroring out on the unrecognized options. The ELF linker has even more LTO options available, but not all of these are currently settable via the lld-link option interface, and some aren't set automatically by Clang but only if the user manually passes them - and thus probably aren't in wide use so far. (Previously LLD/MinGW would have accepted them silently but ignored them though.) Differential Revision: https://reviews.llvm.org/D158887	2023-08-28 00:28:46 +03:00
Sam Clegg	93adcb770b	[lld][WebAssembly] Add `--table-base` setting This is similar to `--global-base` but determines where to place the table segments rather than that data segments. See https://github.com/emscripten-core/emscripten/issues/20097 Differential Revision: https://reviews.llvm.org/D158892	2023-08-25 16:10:38 -07:00
Martin Storsjö	5c92c9f34a	[LLD] [MinGW] Pass LLVM specific LTO options through to lld-link This is modelled after how the ELF driver does it (see e.g. https://reviews.llvm.org/D78158 for the source of the ELF implementation); we need to intercept some options, but ignore GCC specific ones that GCC passes regardless of whether GCC is using LTO or not. This is the second part of the fix for https://github.com/mstorsjo/llvm-mingw/issues/349. Differential Revision: https://reviews.llvm.org/D158412	2023-08-24 23:15:26 +03:00
Kazu Hirata	4efd1e0d3a	[lld] Do not include StringSwitch.h (NFC)	2023-08-23 09:20:14 -07:00
Kyungwoo Lee	a033184abb	[lld-macho] Stricter Bitcode Symbol Resolution LLD resolves symbols before performing LTO compilation, assuming that the symbols in question are resolved by the resulting object files from LTO. However, there is a scenario where the prevailing symbols might be resolved incorrectly due to specific assembly symbols not appearing in the symbol table of the bitcode. This patch deals with such a scenario by generating an error instead of silently allowing a mis-linkage. If a prevailing symbol is resolved through post-loaded archives via LC linker options, a warning will now be issued. Reviewed By: #lld-macho, thevinster Differential Revision: https://reviews.llvm.org/D158003	2023-08-22 12:03:17 -07:00
Shoaib Meenai	97e39f96c8	[ELF] Add -Bsymbolic-non-weak This adds a new -Bsymbolic option that directly binds all non-weak symbols. There's a couple of reasons motivating this: * The new flag will match the default behavior on Mach-O, so you can get consistent behavior across platforms. * We have use cases for which making weak data preemptible is useful, but we don't want to pessimize access to non-weak data. (For a large internal app, we measured 2000+ data symbols whose accesses would be unnecessarily pessimized by `-Bsymbolic-functions`.) Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D158322	2023-08-21 09:11:51 -07:00
Paul Kirth	14e3bec8fc	Reland "[lld] Preliminary fat-lto-object support" This patch adds support to lld for --fat-lto-objects. We add a new --fat-lto-objects option to LLD, and slightly change how it chooses input files in the driver when the option is set. Fat LTO objects contain both LTO compatible IR, as well as generated object code. This allows users to defer the choice of whether to use LTO or not to link-time. This is a feature available in GCC for some time, and makes the existing -ffat-lto-objects option functional in the same way as GCC's. If the --fat-lto-objects option is passed to LLD and the input files are fat object files, then the linker will chose the LTO compatible bitcode sections embedded within the fat object and link them together using LTO. Otherwise, standard object file linking is done using the assembly section in the object files. The previous version of this patch had a missing `REQUIRES: x86` line in `fatlto.invalid.s`. Additionally, it was reported that this patch caused a test failure in `export-dynamic-symbols.s`, however, `29112a9946` disabled the `export-dynamic-symbols.s` test on Windows due to a quotation difference between platforms, unrelated to this patch. Original RFC: https://discourse.llvm.org/t/rfc-ffat-lto-objects-support/63977 Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D146778	2023-08-18 22:51:25 +00:00
Sam Clegg	8e44f037dc	[lld][WebAssembly] Add support for -soname This change writes the module name to the name section of the wasm binary. We use the `-soname` argument to determine the name and we default the output file basename if this option is not specified. In the future we will likely want to embed the soname in the dylink section too, but this the first step in supporting `-soname`. Differential Revision: https://reviews.llvm.org/D158001	2023-08-15 18:33:45 -07:00
Sam Clegg	cb5bc75680	[lld][WebAssembly] Always include bss segments when `--emit-relocs` is used If we don't do this then we end up with symbols that refer to non-existent segments. Differential Revision: https://reviews.llvm.org/D158025	2023-08-15 17:03:40 -07:00
Justin Bogner	dcb6d212fd	Reapply "[Option] Add "Visibility" field and clone the OptTable APIs to use it" This reverts commit `4e3b89483a`, with fixes for places I'd missed updating in lld and lldb. I've also renamed OptionVisibility::Default to "DefaultVis" to avoid ambiguity since the undecorated name has to be available anywhere Options.inc is included. Original message follows: This splits OptTable's "Flags" field into "Flags" and "Visibility", updates the places where we instantiate Option tables, and adds variants of the OptTable APIs that use Visibility mask instead of Include/Exclude flags. We need to do this to clean up a bunch of complexity in the clang driver's option handling - there's a whole slew of flags like CoreOption, NoDriverOption, and FlangOnlyOption there today to try to handle all of the permutations of flags that the various drivers need, but it really doesn't scale well, as can be seen by things like the somewhat recently introduced CLDXCOption. Instead, we'll provide an additive model for visibility that's separate from the other flags. For things like "HelpHidden", which is used as a "subtractive" modifier for option visibility, we leave that in "Flags" and handle it as a special case. Note that we don't actually update the users of the Include/Exclude APIs here or change the flags that exist in clang at all - that will come in a follow up that refactors clang's Options.td to use the increased flexibility this change allows. Differential Revision: https://reviews.llvm.org/D157149	2023-08-15 01:16:58 -07:00
Justin Bogner	4e3b89483a	Revert "[Option] Add "Visibility" field and clone the OptTable APIs to use it" this is failing on bots, reverting to investigate. This reverts commit `a16104e6da`.	2023-08-14 13:31:02 -07:00
Justin Bogner	a16104e6da	[Option] Add "Visibility" field and clone the OptTable APIs to use it This splits OptTable's "Flags" field into "Flags" and "Visibility", updates the places where we instantiate Option tables, and adds variants of the OptTable APIs that use Visibility mask instead of Include/Exclude flags. We need to do this to clean up a bunch of complexity in the clang driver's option handling - there's a whole slew of flags like CoreOption, NoDriverOption, and FlangOnlyOption there today to try to handle all of the permutations of flags that the various drivers need, but it really doesn't scale well, as can be seen by things like the somewhat recently introduced CLDXCOption. Instead, we'll provide an additive model for visibility that's separate from the other flags. For things like "HelpHidden", which is used as a "subtractive" modifier for option visibility, we leave that in "Flags" and handle it as a special case. Note that we don't actually update the users of the Include/Exclude APIs here or change the flags that exist in clang at all - that will come in a follow up that refactors clang's Options.td to use the increased flexibility this change allows. Differential Revision: https://reviews.llvm.org/D157149	2023-08-14 13:24:54 -07:00
Kyungwoo Lee	484c961ccd	[lld-macho] Postprocess LC Linker Option LLD resolves symbols regardless of LTO modes early when reading and parsing input files in order. The object files built from LTO passes are appended later. Because LLD eagerly resolves the LC linker options while parsing a new object file (and its chain of dependent libraries), the prior decision on pending prevailing symbols (belonging to some bitcode files) can change to ones in those native libraries that are just loaded. This patch delays processing LC linker options until all the native object files are added after LTO is done, similar to LD64. This way we preserve the decision on prevailing symbols LLD made, regardless of LTO modes. - When parsing a new object file in `parseLinkerOptions()`, it just parses LC linker options in the header, and saves those contents to `unprocessedLCLinkerOptions`. - After LTO is finished, `resolveLCLinkerOptions()` is called to recursively load dependent libraries, starting with initial linker options collected in `unprocessedLCLinkerOptions` (which also updates during recursions) Reviewed By: #lld-macho, int3 Differential Revision: https://reviews.llvm.org/D157716	2023-08-13 13:39:04 -07:00
Sam Clegg	7a8ee92fa8	[lld][WebAssembly] Process stub libraries in a loop When stub libraries trigger the fetching of new object files we can potentially introduce new undefined symbols so process the stub in loop until no new objects are pulled in. Differential Revision: https://reviews.llvm.org/D153466	2023-08-11 11:07:40 -07:00
namazso	e335c78ec2	[lld][COFF] Remove incorrect flag from EHcont table Fixes EHCont implementation in LLD. Closes #64570 Reviewed By: rnk Differential Revision: https://reviews.llvm.org/D157623	2023-08-10 16:17:38 -04:00
Vy Nguyen	2d873d5aa4	[lld-macho]Rework error-checking in peeking at first-member in archive to avoid segfault. Details: calling getMemoryBufferRef() on an empty archive can trigger segfault so the code should check before calling this. this seems like a bug in the Archive API but that can be fixed separately. P.S: follow up to D156468 Differential Revision: https://reviews.llvm.org/D157300	2023-08-09 09:54:32 -04:00
Fangrui Song	ebb0a21099	[GlobPattern] Update invalid glob pattern diagnostic for unmatched '[' Update this diagnostic to mention the reason (unmatched '['), matching the other diagnostic about stray '\'. The original pattern is omitted, as some users may mention the original pattern in another position, not repeating it.	2023-08-08 20:25:10 -07:00
Weining Lu	8a31f7ddb8	[lld][LoongArch] Support the R_LARCH_PCREL20_S2 relocation type `R_LARCH_PCREL20_S2` is a new added relocation type in LoongArch ELF psABI v2.10 [1] which is not corvered by D138135 except `R_LARCH_64_PCREL`. A motivation to support `R_LARCH_PCREL20_S2` in lld is to build the runtime of .NET core (a.k.a `CoreCLR`) in which strict PC-relative semantics need to be guaranteed [2]. The normal `pcalau12i + addi.d` approach doesn't work because the code will be copied to other places with different "page" and offsets. To achieve this, we can use `pcaddi` with explicit `R_LARCH_PCREL20_S2` reloc to address +-2MB PC-relative range with 4-bytes aligned. [1]: https://github.com/loongson/la-abi-specs/releases/tag/v2.10 [2]: https://github.com/dotnet/runtime/blob/release/7.0/src/coreclr/vm/loongarch64/asmhelpers.S#L307 Reviewed By: xen0n, MaskRay Differential Revision: https://reviews.llvm.org/D156772	2023-08-09 09:55:27 +08:00

1 2 3 4 5 ...

16341 Commits