llvm-project

Commit Graph

Author	SHA1	Message	Date
Yaxun (Sam) Liu	10eb3bf2d4	Skip -fPIE for AMDGPU and HIP toolchain AMDGPU toolchain does not support -fPIE, therefore skip it if specified by driver. Differential Revision: https://reviews.llvm.org/D88425	2020-09-28 22:03:18 -04:00
Valentin Clement	bbb5dc4923	[mlir][openacc] Add acc.data operation verifier Add a basic verifier for the data operation following the restriction from the standard. Reviewed By: kiranchandramohan Differential Revision: https://reviews.llvm.org/D88334	2020-09-28 21:22:32 -04:00
Nathan Ridge	cc6d1f8029	[clangd] When finding refs for a renaming alias, do not return refs to underlying decls Fixes https://github.com/clangd/clangd/issues/515 Differential Revision: https://reviews.llvm.org/D87225	2020-09-28 21:18:31 -04:00
Mehdi Amini	9f9f89d44b	Remove dependency from LLVM Dialect on the OpenMP dialect The OmpDialect is in practice optional during translation to LLVM IR: the code is tolerant to have a "nullptr" when not present / needed. The dependency still exists on the export to LLVMIR. Reviewed By: ftynse Differential Revision: https://reviews.llvm.org/D88351	2020-09-29 01:12:01 +00:00
LLVM GN Syncbot	727c4223d7	[gn build] Port `54d9f743c8`	2020-09-29 00:24:06 +00:00
Richard Smith	c375635d05	Ensure that we don't compute linkage for an anonymous class too early if it has a member whose name is the same as a builtin. Fixes a regression from the introduction of BuiltinAttr.	2020-09-28 17:22:40 -07:00
Jan Korous	6fd8c69049	[clang] Update warning-wall.c test Follow-up to 1e86d637eb4f: [clang] Selectively ena/disa-ble format-insufficient-args warning	2020-09-28 17:19:51 -07:00
Ruiling Song	73805329ba	[RegisterCoalescer] Pass Undefs to extendToIndices() When extending the subranges, the reaching-def may be an undefs. When extending such kind of subrange, it will try to search for the reaching def first. If the reaching def is an undef and we did not provide 'Undefs', The findReachingDefs() will fail with message: "Use of $noreg does not have a corresponding definition on every path: LLVM ERROR: Use not jointly dominated by defs." So we computeSubRangeUndefs() and pass the result to extendToIndices(). Reviewed By: arsenm Differential Revision: https://reviews.llvm.org/D87744	2020-09-29 08:14:24 +08:00
Zahira Ammarguellat	efd04721c9	BuildVectorType with a dependent (array) type is crashing the compiler - Fix for PR-47542 Reviewed By: rnk Differential Revision: https://reviews.llvm.org/D88150	2020-09-28 17:10:32 -07:00
Yonghong Song	54d9f743c8	BPF: move AbstractMemberAccess and PreserveDIType passes to EP_EarlyAsPossible Move abstractMemberAccess and PreserveDIType passes as early as possible, right after clang code generation. Currently, compiler may transform the above code p1 = llvm.bpf.builtin.preserve.struct.access(base, 0, 0); p2 = llvm.bpf.builtin.preserve.struct.access(p1, 1, 2); a = llvm.bpf.builtin.preserve_field_info(p2, EXIST); if (a) { p1 = llvm.bpf.builtin.preserve.struct.access(base, 0, 0); p2 = llvm.bpf.builtin.preserve.struct.access(p1, 1, 2); bpf_probe_read(buf, buf_size, p2); } to p1 = llvm.bpf.builtin.preserve.struct.access(base, 0, 0); p2 = llvm.bpf.builtin.preserve.struct.access(p1, 1, 2); a = llvm.bpf.builtin.preserve_field_info(p2, EXIST); if (a) { bpf_probe_read(buf, buf_size, p2); } and eventually assembly code looks like reloc_exist = 1; reloc_member_offset = 10; //calculate member offset from base p2 = base + reloc_member_offset; if (reloc_exist) { bpf_probe_read(bpf, buf_size, p2); } if during libbpf relocation resolution, reloc_exist is actually resolved to 0 (not exist), reloc_member_offset relocation cannot be resolved and will be patched with illegal instruction. This will cause verifier failure. This patch attempts to address this issue by do chaining analysis and replace chains with special globals right after clang code gen. This will remove the cse possibility described in the above. The IR typically looks like %6 = load @llvm.sk_buff:0:50$0:0:0:2:0 %7 = bitcast %struct.sk_buff* %2 to i8* %8 = getelementptr i8, i8* %7, %6 for a particular address computation relocation. But this transformation has another consequence, code sinking may happen like below: PHI = <possibly different @preserve__access_globals> %7 = bitcast %struct.sk_buff %2 to i8* %8 = getelementptr i8, i8* %7, %6 For such cases, we will not able to generate relocations since multiple relocations are merged into one. This patch introduced a passthrough builtin to prevent such optimization. Looks like inline assembly has more impact for optimizaiton, e.g., inlining. Using passthrough has less impact on optimizations. A new IR pass is introduced at the beginning of target-dependent IR optimization, which does: - report fatal error if any reloc global in PHI nodes - remove all bpf passthrough builtin functions Changes for existing CORE tests: - for clang tests, add "-Xclang -disable-llvm-passes" flags to avoid builtin->reloc_global transformation so the test is still able to check correctness for clang generated IR. - for llvm CodeGen/BPF tests, add "opt -O2 <ir_file> \| llvm-dis" command before "llc" command since "opt" is needed to call newly-placed builtin->reloc_global transformation. Add target triple in the IR file since "opt" requires it. - Since target triple is added in IR file, if a test may produce different results for different endianness, two tests will be created, one for bpfeb and another for bpfel, e.g., some tests for relocation of lshift/rshift of bitfields. - field-reloc-bitfield-1.ll has different relocations compared to old codes. This is because for the structure in the test, new code returns struct layout alignment 4 while old code is 8. Align 8 is more precise and permits double load. With align 4, the new mechanism uses 4-byte load, so generating different relocations. - test intrinsic-transforms.ll is removed. This is used to test cse on intrinsics so we do not lose metadata. Now metadata is attached to global and not instruction, it won't get lost with cse. Differential Revision: https://reviews.llvm.org/D87153	2020-09-28 16:56:22 -07:00
David Tenty	ee80615b5c	[clang][driver][AIX] Set compiler-rt as default rtlib Reviewed By: hubert.reinterpretcast Differential Revision: https://reviews.llvm.org/D88182	2020-09-28 19:45:43 -04:00
ogiroux	665dc4012b	Attempt to clear some msan errors in the libcxx atomic tests.	2020-09-28 16:34:41 -07:00
Diego Caballero	93936da904	[mlir][Affine][VectorOps] Fix super vectorizer utility (D85869) Adding missing code that should have been part of "D85869: Utility to vectorize loop nest using strategy." Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D88346	2020-09-28 16:24:11 -07:00
Kostya Kortchinsky	f668a84b58	[scudo][standalone] Remove unused atomic_compare_exchange_weak `atomic_compare_exchange_weak` is unused in Scudo, and its associated test is actually wrong since the weak variant is allowed to fail spuriously (thanks Roland). This lead to flakes such as: ``` [ RUN ] ScudoAtomicTest.AtomicCompareExchangeTest ../../zircon/third_party/scudo/src/tests/atomic_test.cpp:98: Failure: Expected atomic_compare_exchange_weak(reinterpret_cast<T >(&V), &OldVal, NewVal, memory_order_relaxed) is true. Expected: true Which is: 01 Actual : atomic_compare_exchange_weak(reinterpret_cast<T >(&V), &OldVal, NewVal, memory_order_relaxed) Which is: 00 ../../zircon/third_party/scudo/src/tests/atomic_test.cpp💯 Failure: Expected atomic_compare_exchange_weak( reinterpret_cast<T >(&V), &OldVal, NewVal, memory_order_relaxed) is false. Expected: false Which is: 00 Actual : atomic_compare_exchange_weak( reinterpret_cast<T >(&V), &OldVal, NewVal, memory_order_relaxed) Which is: 01 ../../zircon/third_party/scudo/src/tests/atomic_test.cpp:101: Failure: Expected OldVal == NewVal. Expected: NewVal Which is: 24 Actual : OldVal Which is: 42 [ FAILED ] ScudoAtomicTest.AtomicCompareExchangeTest (0 ms) [----------] 2 tests from ScudoAtomicTest (1 ms total) ``` So I am removing this, if someone ever needs the weak variant, feel free to add it back with a test that is not as terrible. This test was initially ported from sanitizer_common, but their weak version calls the strong version, so it works for them. Differential Revision: https://reviews.llvm.org/D88443	2020-09-28 16:25:14 -07:00
Jan Korous	1e86d637eb	[clang] Selectively ena/disa-ble format-insufficient-args warning Differential Revision: https://reviews.llvm.org/D87176	2020-09-28 16:24:50 -07:00
Mehdi Amini	e72d792c14	Guard `find_library(tensorflow_c_api ...)` by checking for TENSORFLOW_C_LIB_PATH to be set by the user Also have CMake fails if the user provides a TENSORFLOW_C_LIB_PATH but we can't find TensorFlow at this path. At the moment the CMake script tries to figure if TensorFlow is available on the system and enables support for it. This is in general not desirable to customize build features this way and instead it is preferable to let the user opt-in explicitly into the features they want to enable. This is in line with other optional external dependencies like Z3. There are a few reasons to this but amongst others: - reproducibility: making features "magically" enabled based on whether we find a package on the system or not makes it harder to handle bug reports from users. - user control: they can't have TensorFlow on the system and build LLVM without TensorFlow right now. They also would suddenly distribute LLVM with a different set of features unknowingly just because their build machine environment would change subtly. Right now this is motivated by a user reporting build failures on their system: .../mesa-git/llvm-git/src/llvm-project/llvm/lib/Analysis/TFUtils.cpp:23:10: fatal error: tensorflow/c/c_api.h: No such file or directory 23 \| #include "tensorflow/c/c_api.h" \| ^~~~~~ It looks like we detected TensorFlow at configure time but couldn't set all the paths correctly. Differential Revision: https://reviews.llvm.org/D88371	2020-09-28 22:15:55 +00:00
Philip Reames	e46d74b589	[CVP] Allow two transforms in one invocation For a call site which had both constant deopt operands and nonnull arguments, we were missing the opportunity to recognize the later by bailing early. This is somewhat of a speculative fix. Months ago, I'd had a private report of performance and compile time regressions from the deopt operand folding. I never received a test case. However, the only possibility I see was that after that change CVP missed the nonnull fold, and we end up with a pass ordering/missed simplification issue. So, since it's a real issue, fix it and hope.	2020-09-28 15:11:42 -07:00
Fangrui Song	bd08a87cfe	[EHStreamer] Simplify sharedTypeIDs with std::mismatch (Note that EMStreamer.cpp is largely under tested. The only test checking the prefix sharing is CodeGen/WebAssembly/eh-lsda.ll)	2020-09-28 15:05:59 -07:00
Sean Silva	a975be0e00	[mlir][shape] Make conversion passes more consistent. - use select-ops to make the lowering simpler - change style of FileCheck variables names to be consistent - change some variable names in the code to be more explicit Differential Revision: https://reviews.llvm.org/D88258	2020-09-28 14:55:42 -07:00
Petr Hosek	2d657d1bd7	[libcxx] Don't pass -s to libtool This flag is the default in libtool on Darwin, and it's not supported by llvm-libtool-darwin causing a build failure. Differential Revision: https://reviews.llvm.org/D88449	2020-09-28 14:50:09 -07:00
Louis Dionne	d092c91288	[libc++] Fix constexpr dynamic allocation on GCC 10 We're technically not allowed by the Standard to call ::operator new in constexpr functions like __libcpp_allocate. Clang doesn't seem to complain about it, but GCC does.	2020-09-28 17:44:31 -04:00
Craig Topper	e53196b1e8	[X86] Add support for calling SimplifyDemandedBits on the input of PDEP with a constant mask. We can do several optimizations for PDEP using computeKnownBits and SimplifyDemandedBits -If the MSBs of the output aren't demanded, those MSBs of the mask input aren't demanded either. We need to keep the most significant demanded bit of the mask and any mask bits before it. -The number of possible ones in the mask determines how many bits of the lsbs of the other operand are demanded. Any bits of the mask we don't demand by the previous rule should not be counted. -The result will have zeros in any position that the mask is zero. -Since non-mask input bits can only be output in the original position or a higher bit position, the result will have at least as many trailing zeroes as the non-mask input. Differential Revision: https://reviews.llvm.org/D87883	2020-09-28 14:21:30 -07:00
Craig Topper	e5ef523ee4	[X86] Add tests for D87883. NFC	2020-09-28 14:21:29 -07:00
Amara Emerson	082321909e	[GlobalISel] Add support for lowering of vector G_SELECT and use for AArch64. The lowering is a port of the SDAG expansion. Differential Revision: https://reviews.llvm.org/D88364	2020-09-28 14:00:46 -07:00
David Tenty	25affb04aa	[CMake][AIX] Limit tools in external project build This is a follow on to D85329 which disabled some llvm tools in the runtimes build due to XCOFF64 limitations. This change disables them in other external project builds as well, when no list of tools is specified in the arguments. Reviewed By: hubert.reinterpretcast, stevewan Differential Revision: https://reviews.llvm.org/D88310	2020-09-28 16:59:25 -04:00
Nico Weber	d897351335	[gn build] Re-run CompletionModelCodegen when input json files change	2020-09-28 16:58:00 -04:00
Aaron Ballman	e7549dafcd	Fix a think-o with the numerical suffixes in the docs for init_priority.	2020-09-28 16:52:58 -04:00
Jonas Devlieghere	974551d37d	[lldb] Add print_function import	2020-09-28 13:51:11 -07:00
Amara Emerson	5aa56b2429	Revert "Revert "[AArch64][GlobalISel] Add selection support for <8 x s16> G_INSERT_VECTOR_ELT with GPR scalar."" This isn't a real with the codegen, it's a previously known bug in clang which causes non-deterministic failures due to garbage bits in undef registers being used in saturating instructions. I'm disabling the result checking for the test until this issue is resolved. This reverts commit `6c8168324b`.	2020-09-28 13:44:51 -07:00
Aart Bik	e9628955f5	[mlir] [VectorOps] Relaxed restrictions on vector.reduction types even more Recently, restrictions on vector reductions were made more relaxed by accepting any width signless integer and floating-point. This CL relaxes the restriction even more by including unsigned and signed integers. Reviewed By: bkramer Differential Revision: https://reviews.llvm.org/D88442	2020-09-28 13:38:03 -07:00
Craig Topper	288c5776c9	[X86] Use inlineasm flag output for the _bittest* intrinsics. Instead of expliciting emitting a setc in the inline asm instructions, we can use flag output. This allows the backend to use the flag directly if it is needed by a branch. Previously we needed a test instruction to convert the register back to a flag. If the flag can't be used directly, the backend will emit a setcc. Differential Revision: https://reviews.llvm.org/D87888	2020-09-28 13:33:22 -07:00
Simon Pilgrim	2f768a68a1	[InstCombine] Regenerate cast tests. NFC.	2020-09-28 21:32:12 +01:00
Eric Astor	bd19876dc6	[COFF] Aliases resolve directly to defined external targets Avoid introducing unnecessary indirection for weak-external references. We only need to introduce ".weak.<SYMBOL>.default" when referencing a symbol that is defined, but not external. Reviewed By: mstorsjo Differential Revision: https://reviews.llvm.org/D88305	2020-09-28 16:12:45 -04:00
Louis Dionne	59f8ac3eb4	[libc++] Replace uses of __libcpp_allocate by std::allocator<> Both are equivalent, however std::allocator can appear in constant expressions and is higher level.	2020-09-28 16:09:42 -04:00
Louis Dionne	93ba33066c	[libc++] Add UNSUPPORTED markup to atomic test in single-threaded mode	2020-09-28 16:09:15 -04:00
Louis Dionne	46fdaac098	[libc++] Fix heap UaF issue in coroutine test This wasn't being flagged by older versions of ASAN, but it is now.	2020-09-28 16:09:15 -04:00
Benjamin Kramer	b59dff4b16	[wasm] Move WasmTraits.h to BinaryFormat There's no dependency on Object in there and this avoids a cyclic dependency between libMC and libObject.	2020-09-28 22:07:28 +02:00
Sanjay Patel	c37a8acef6	[CostModel] remove hack for intrinsic cost based on cost type This hack seems to only have been necessary because of the constructor bug noted in `33125cffd`. Once again, it's hard to prove NFC, but that's the hope...	2020-09-28 15:58:42 -04:00
Jason Molenda	6e54918db7	Once we've found a firmware binary and loaded it, don't search more Add the flag in ProcessMachCore::DoLoadCore that stops additional searches for the binaries when we have an LC_NOTE identifying the firmware/standalone binary as the correct one & we have loaded it successfully.	2020-09-28 12:51:23 -07:00
Jonas Devlieghere	8b95bd3310	[lldb] Enable markdown support for documentation This enables support for writing LLDB documentation in markdown in addition to reStructured text. We already had documentation written in markdown (StructuredDataPlugins and DarwinLog) which will now also be available on the website.	2020-09-28 12:51:15 -07:00
Baptiste Saleil	0156914275	[PowerPC] Legalize v256i1 and v512i1 and implement load and store of these types This patch legalizes the v256i1 and v512i1 types that will be used for MMA. It implements loads and stores of these types. v256i1 is a pair of VSX registers, so for this type, we load/store the two underlying registers. v512i1 is used for MMA accumulators. So in addition to loading and storing the 4 associated VSX registers, we generate instructions to prime (copy the VSX registers to the accumulator) after loading and unprime (copy the accumulator back to the VSX registers) before storing. This patch also adds the UACC register class that is necessary to implement the loads and stores. This class represents accumulator in their unprimed form and allow the distinction between primed and unprimed accumulators to avoid invalid copies of the VSX registers associated with primed accumulators. Differential Revision: https://reviews.llvm.org/D84968	2020-09-28 14:39:37 -05:00
Sanjay Patel	33125cffda	[CostModel] fill in arguments as part of intrinsic attribute constructor This appears to be an error of code duplication - instead of one constructor variant calling another, we have N similar but not identical versions. I think this is 'NFC' based on the current callers, but it's hard to tell or guess the intent in all cases.	2020-09-28 15:27:45 -04:00
Paweł Bylica	0c82fa677f	[python][tests] Fix string comparison with "is"	2020-09-28 21:11:50 +02:00
Peter Collingbourne	e851aeb0a5	scudo: Re-order Allocator fields for improved performance. NFCI. Move smaller and frequently-accessed fields near the beginning of the data structure in order to improve locality and reduce the number of instructions required to form an access to those fields. With this change I measured a ~5% performance improvement on BM_malloc_sql_trace_default on aarch64 Android devices (Pixel 4 and DragonBoard 845c). Differential Revision: https://reviews.llvm.org/D88350	2020-09-28 11:51:45 -07:00
Aart Bik	54759cefdb	[mlir] [VectorOps] changes to printing support for integers (1) simplify integer printing logic by always using 64-bit print (2) add index support (since vector<16xindex> is planned to be added) (3) adjust naming convention print_x -> printX Reviewed By: bkramer Differential Revision: https://reviews.llvm.org/D88436	2020-09-28 11:43:31 -07:00
Jon Roelofs	83dc53d30c	[AArch64] reuse another map iterator. NFC	2020-09-28 11:30:21 -07:00
Amara Emerson	6c8168324b	Revert "[AArch64][GlobalISel] Add selection support for <8 x s16> G_INSERT_VECTOR_ELT with GPR scalar." This reverts commit `b5e87c9ef2` as it seems to have broken a bot.	2020-09-28 11:25:19 -07:00
Utkarsh Saxena	9b1666f3ce	[clangd] Rename evaluate() to evaluateHeuristics() Since we have 2 scoring functions (heuristics and decision forest), renaming the existing evaluate() function to be more descriptive of the Heuristics being evaluated in it. Differential Revision: https://reviews.llvm.org/D88431	2020-09-28 20:05:01 +02:00
Dominic Chen	06e68f05da	[AddressSanitizer] Copy type metadata to prevent miscompilation When ASan and e.g. Dead Virtual Function Elimination are enabled, the latter will rely on type metadata to determine if certain virtual calls can be removed. However, ASan currently does not copy type metadata, which can cause virtual function calls to be incorrectly removed. Differential Revision: https://reviews.llvm.org/D88368	2020-09-28 13:56:05 -04:00
Simon Pilgrim	d047bb1cf6	[InstCombine] Add trunc(shr(trunc(x),c)) non-uniform vector tests	2020-09-28 18:53:38 +01:00

1 2 3 4 5 ...

367576 Commits All Branches Search

367576 Commits

All Branches