llvm-project

Commit Graph

Author	SHA1	Message	Date
Jordan Rupprecht	e41aaea262	[NFC][libObject] clang-format Archive{.h,.cpp} In preparation for D100651	2021-05-27 16:48:40 -07:00
Andrea Di Biagio	57646d38d5	[MCA] Minor changes to the InOrderIssueStage. NFC The constructor of InOrderIssueStage no longer takes as input a reference to the target scheduling model. The stage can always query the subtarget to obtain a reference to the scheduling model. The ResourceManager is no longer stored internally as a unique_ptr. Moved a couple of method definitions to the .cpp file.	2021-05-28 00:33:59 +01:00
Arthur Eubanks	8086f9d87e	[ConstFold] Simplify a load's GEP operand through local aliases MSVC-style RTTI produces loads through a GEP of a local alias which itself is a GEP. Currently we aren't able to devirtualize any virtual calls when MSVC RTTI is enabled. This patch attempts to simplify a load's GEP operand by calling SymbolicallyEvaluateGEP() with an option to look through local aliases. Differential Revision: https://reviews.llvm.org/D101100	2021-05-27 16:04:19 -07:00
Craig Topper	0fa5aac292	[RISCV] Teach VSETVLI insertion to look through PHIs to prove we don't need to insert a vsetvli. If an instruction's AVL operand is a PHI node in the same block, we may be able to peek through the PHI to find vsetvli instructions that produce the AVL in other basic blocks. If we can prove those vsetvli instructions have the same VTYPE and were the last vsetvli in their respective blocks, then we don't need to insert a vsetvli for this pseudo instruction. Reviewed By: rogfer01 Differential Revision: https://reviews.llvm.org/D103277	2021-05-27 15:34:08 -07:00
Arthur Eubanks	2d2a902078	[SanCov] Properly set ABI parameter attributes Arguments need to have the proper ABI parameter attributes set. Followup to D101806. Reviewed By: rnk Differential Revision: https://reviews.llvm.org/D103288	2021-05-27 15:27:21 -07:00
Roman Lebedev	ee544b8d86	[NFC][X86][Codegen] Re-autogenerate a few tests to reduce noise in future changes	2021-05-28 00:58:01 +03:00
Aart Bik	ef1cc4e7ae	[mlir][capi] fix build issue with "all passes" registration Some builds exposed missing dependences on trafo/conv passes. Reviewed By: jpienaar Differential Revision: https://reviews.llvm.org/D103283	2021-05-27 14:57:21 -07:00
Ryan Prichard	b834d63094	[sanitizer] Android ELF TLS is supported from Q (API 29) Reviewed By: oontvoo, MaskRay Differential Revision: https://reviews.llvm.org/D103214	2021-05-27 14:53:49 -07:00
River Riddle	8cbbc5d00b	[mlir-lsp-server] Add support for processing split files MLIR tools very commonly use `// -----` to split a file into distinct sub documents, that are processed separately. This revision adds support to mlir-lsp-server for splitting MLIR files based on this sigil, and processing them separately. Differential Revision: https://reviews.llvm.org/D102660	2021-05-27 14:42:37 -07:00
Andrea Di Biagio	50770d8de5	[MCA] Refactor the InOrderIssueStage stage. NFCI Moved the logic that checks for RAW hazards from the InOrderIssueStage to the RegisterFile. Changed how the InOrderIssueStage keeps track of backend stalls. Stall events are now generated from method notifyStallEvent(). No functional change intended.	2021-05-27 22:28:04 +01:00
Quinn Pham	62b5df7fe2	[PowerPC] Added multiple PowerPC builtins This is the first in a series of patches to provide builtins for compatibility with the XL compiler. Most of the builtins already had intrinsics and only needed to be implemented in the front end. Intrinsics were created for the three iospace builtins, eieio, and icbt. Pseudo instructions were created for eieio and iospace_eieio to ensure that nops were inserted before the eieio instruction. Reviewed By: nemanjai, #powerpc Differential Revision: https://reviews.llvm.org/D102443	2021-05-27 16:23:03 -05:00
Reid Kleckner	109aac9212	[PDB] Enable parallel ghash type merging by default Ghashing is probably going to be faster in most cases, even without precomputed ghashes in object files. Here is my table of results linking clang.pdb: ------------------------------- \| threads \| GHASH \| NOGHASH \| ------------------------------- \| j1 \| 51.031s \| 25.141s \| \| j2 \| 31.079s \| 22.109s \| \| j4 \| 18.609s \| 23.156s \| \| j8 \| 11.938s \| 21.984s \| \| j28 \| 8.375s \| 18.391s \| ------------------------------- This shows that ghashing is faster if at least four cores are available. This may make the linker slower if most cores are busy in the middle of a build, but in that case, the linker probably isn't on the critical path of the build. Incremental build performance is arguably more important than highly contended batch build link performance. The -time output indicates that ghash computation is the dominant factor: Input File Reading: 924 ms ( 1.8%) GC: 689 ms ( 1.3%) ICF: 527 ms ( 1.0%) Code Layout: 414 ms ( 0.8%) Commit Output File: 24 ms ( 0.0%) PDB Emission (Cumulative): 49938 ms ( 94.8%) Add Objects: 46783 ms ( 88.8%) Global Type Hashing: 38983 ms ( 74.0%) GHash Type Merging: 5640 ms ( 10.7%) Symbol Merging: 2154 ms ( 4.1%) Publics Stream Layout: 188 ms ( 0.4%) TPI Stream Layout: 18 ms ( 0.0%) Commit to Disk: 2818 ms ( 5.4%) -------------------------------------------------- Total Link Time: 52669 ms (100.0%) We can speed that up with a faster content hash (not SHA1). Differential Revision: https://reviews.llvm.org/D102888	2021-05-27 14:19:36 -07:00
Craig Topper	020df692d8	[RISCV] Fix typo, use addImm instead of addReg.	2021-05-27 14:04:51 -07:00
River Riddle	d47dd11071	[mlir] Add support for querying the ModRef behavior from the AliasAnalysis class This allows for checking if a given operation may modify/reference/or both a given value. Right now this API is limited to Value based memory locations, but we should expand this to include attribute based values at some point. This is left for future work because the rest of the AliasAnalysis API also has this restriction. Differential Revision: https://reviews.llvm.org/D101673	2021-05-27 13:57:29 -07:00
Martin Storsjö	b3ceffdf35	[libcxx] [test] Convert an XFAIL LIBCXX-WINDOWS-FIXME into UNSUPPORTED with explanation Differential Revision: https://reviews.llvm.org/D103149	2021-05-27 23:51:24 +03:00
Martin Storsjö	0e4cf807ae	[clang] [MinGW] Don't mark emutls variables as DSO local These actually can be automatically imported from another DLL. (This works properly as long as the actual implementation of emutls is linked dynamically from e.g. libgcc; if the implementation comes from compiler-rt or a statically linked libgcc, it doesn't work as intended.) This fixes PR50146 and https://github.com/msys2/MINGW-packages/issues/8706 (fixing calling std::call_once in a dynamically linked libstdc++); since `f731839584` the dso_local attribute on the TLS variable affected the actual generated code for accessing the emutls variable. The dso_local attribute on the emutls variable made those accesses to use 32 bit relative addressing in code, which requires runtime pseudo relocations in the text section, and breaks entirely if the actual other variable ends up loaded too far away in the virtual address space. Differential Revision: https://reviews.llvm.org/D102970	2021-05-27 23:51:22 +03:00
Louis Dionne	aad878f112	[libc++] NFC: Make it easier for vendors to extend the run-buildbot script	2021-05-27 16:51:47 -04:00
Erich Keane	cb66bf2c6d	Replace 'magic static' with a member variable for SCYL kernel names I discovered when merging the __builtin_sycl_unique_stable_name into my downstream that it is actually possible for the cc1 invocation to have more than 1 Sema instance, if you pass it multiple input files, each gets its own Sema instance and thus ASTContext instance. The result was that the call to Filter the SYCL kernels was using an ItaniumMangleContext stored via a 'magic static', so it had an invalid reference to ASTContext when processing the 2nd failure. The failure is unfortunately flakey/transient, but the test that fails was added anyway. The magic-static was switched to a unique_ptr member variable in ASTContext that is initialized when needed.	2021-05-27 13:46:31 -07:00
Sanjay Patel	0d5219feb9	[x86] add tests for extend of vector compare; NFC	2021-05-27 16:35:15 -04:00
Roman Lebedev	9712b16763	[NFC][X86][Codegen] vector-interleaved-store-i16-stride-5.ll: precisely match the actual IR Now that i've reimplemented the testcase generator to produce actual IR (https://godbolt.org/z/s7PM8E6v9), it turns out that this was the only discrepancy from what the LV would produce.	2021-05-27 23:25:15 +03:00
Adrian Prantl	f3869a5c32	Support stripping indirectly referenced DILocations from !llvm.loop metadata in stripDebugInfo(). This patch fixes an oversight in https://reviews.llvm.org/D96181 and also takes into account loop metadata pointing to other MDNodes that point into the debug info. rdar://78487175 Differential Revision: https://reviews.llvm.org/D103220	2021-05-27 13:23:33 -07:00
Georgeta Igna	50f17e9d31	[analyzer] RetainCountChecker: Disable reference counting for OSMetaClass. It is a reference-counted class but it uses different methods for that and the checker doesn't understand them yet. Differential Revision: https://reviews.llvm.org/D103081	2021-05-27 13:12:19 -07:00
Eugene Zhulenev	8f23fac4da	[mlir:Async] Convert assertions to async errors only inside async functions Differential Revision: https://reviews.llvm.org/D103278	2021-05-27 12:49:00 -07:00
Walter Erquinigo	32bacb7410	[lldb][intel-pt] Remove old plugin Now that LLDB proper has built-in support for intel-pt traces, we can remove the old plugin written by Intel. It has less features and it's hard to work with. As a test, I ran "ninja lldbIntelFeatures" and it worked. Differential Revision: https://reviews.llvm.org/D102866	2021-05-27 12:16:22 -07:00
Craig Topper	d7ae2438b9	[RISCV] Add a test showing missed opportunity to avoid a vsetvli in a loop. This is another case we need to look through a phi to prove.	2021-05-27 11:30:25 -07:00
Louis Dionne	8d7d7f340e	[libc++] NFC: Refactor raw_storage_iterator test to use UNSUPPORTED markup The test would previously disable itself using `#if TEST_STD_VER` instead of using UNSUPPORTED markup.	2021-05-27 14:23:32 -04:00
Vitaly Buka	c261edb277	[NFC][scudo] Check zeros on smaller allocations 1Tb counting was the slowest test under the QEMU with MTE.	2021-05-27 11:14:26 -07:00
Jacques Pienaar	5618a5a059	[mlir] Update cmake variable post D102976	2021-05-27 11:11:58 -07:00
Eugene Zhulenev	9136b7d075	[mlir] AsyncRefCounting: check that LivenessBlockInfo is not nullptr Differential Revision: https://reviews.llvm.org/D103270	2021-05-27 10:54:21 -07:00
Saleem Abdulrasool	4cc5a97101	MC: mark `dump` with `LLVM_DUMP_METHOD` Mark the `ELFRelocationEntry::dump` method as `LLVM_DUMP_METHOD` to annotate it properly as used to prevent the function being dead stripped away. This allows use of `dump` in the debugger. This is purely to improve the developer experience.	2021-05-27 10:47:39 -07:00
Vitaly Buka	eb69763ad8	[NFC][scudo] Rename internal function	2021-05-27 10:41:07 -07:00
Louis Dionne	b6399e85d8	Revert "[libc++] NFC: Parenthesize expression to satisfy GCC 11" That fix was actually incorrect and caused tests to start failing.	2021-05-27 13:42:39 -04:00
Roman Lebedev	bafbec8535	[NFC][X86][Codegen] Re-autogenerate check lines in a few tests to remove noise from future changes	2021-05-27 20:29:50 +03:00
Simon Pilgrim	90d25808c4	[CostModel][X86] Improve accuracy of sext/zext to 256-bit vector costs on AVX1 targets Determined from llvm-mca analysis (btver2 vs bdver2 vs sandybridge), the split+extends+concat sequence on AVX1 capable targets are cheaper than the #ops that the cost was previously based on.	2021-05-27 18:17:50 +01:00
Craig Topper	527cd01314	[RISCV] Teach vsetvli insertion to use vsetvl x0, x0 form when we can tell that VLMAX and AVL haven't changed. This can help avoid needing a virtual register for the vsetvl output when the AVL is X0. For other register AVLs it can shorter the live range of the AVL register if it isn't needed later. There's probably no advantage when AVL is a 5 bit immediate that can use vsetivli. But do it anyway for consistency. Reviewed By: rogfer01 Differential Revision: https://reviews.llvm.org/D103215	2021-05-27 10:11:38 -07:00
thomasraoux	750799b7bc	[mlir][NFC] Don't outline kernel in MMA integration tests This matches better how other gpu integration tests are done. Differential Revision: https://reviews.llvm.org/D103099	2021-05-27 09:43:54 -07:00
Eugene Zhulenev	d8c84d2a4e	[mlir] Async: Add error propagation support to async groups Depends On D103109 If any of the tokens/values added to the `!async.group` switches to the error state, than the group itself switches to the error state. Reviewed By: mehdi_amini Differential Revision: https://reviews.llvm.org/D103203	2021-05-27 09:35:11 -07:00
Craig Topper	a105d3024e	[X86] Fold (shift undef, X)->0 for vector shifts by immediate. We could previously do this by accident through the later call to getTargetConstantBitsFromNode I think, but that only worked if N0 had a single use. This patch makes it explicit for undef and doesn't have a use count check. I think this is needed to move the (shl X, 1)->(add X, X) fold to isel for PR50468. We need to be sure X won't be IMPLICIT_DEF which might prevent the same vreg from being used for both operands. Differential Revision: https://reviews.llvm.org/D103192	2021-05-27 09:31:47 -07:00
Craig Topper	b5f8ac2682	[X86] Pre-commit tests for D103192. NFC	2021-05-27 09:31:47 -07:00
Eugene Zhulenev	39957aa424	[mlir] Add error state and error propagation to async runtime values Depends On D103102 Not yet implemented: 1. Error handling after synchronous await 2. Error handling for async groups Will be addressed in the followup PRs Reviewed By: mehdi_amini Differential Revision: https://reviews.llvm.org/D103109	2021-05-27 09:28:47 -07:00
Marco Elver	4fbc66cd6d	[Clang] Enable __has_feature(coverage_sanitizer) Like other sanitizers, enable __has_feature(coverage_sanitizer) if clang has enabled at least one SanitizerCoverage instrumentation type. Because coverage instrumentation selection is not handled via normal -fsanitize= (and thus not in SanitizeSet), passing this information through to LangOptions required propagating the already parsed -fsanitize-coverage= options from CodeGenOptions through to LangOptions in FixupInvocation(). Reviewed By: aaron.ballman Differential Revision: https://reviews.llvm.org/D103159	2021-05-27 18:24:21 +02:00
Eugene Zhulenev	c412979cde	[mlir] Async reference counting for block successors with divergent reference counted liveness Support reference counted values implicitly passed (live) only to some of the successors. Example: if branched to ^bb2 token will leak, unless `drop_ref` operation is properly created ``` ^entry: %token = async.runtime.create : !async.token cond_br %cond, ^bb1, ^bb2 ^bb1: async.runtime.await %token async.runtime.drop_ref %token br ^bb2 ^bb2: return ``` Reviewed By: mehdi_amini Differential Revision: https://reviews.llvm.org/D103102	2021-05-27 09:21:59 -07:00
maekawatoshiki	2165360003	[LoopUnrollAndJam] Change LoopUnrollAndJamPass to LoopNest pass This patch changes LoopUnrollAndJamPass from FunctionPass to LoopNest pass. The next patch will utilize LoopNest to effectively handle loop nests. Reviewed By: Whitney Differential Revision: https://reviews.llvm.org/D99149	2021-05-28 01:17:23 +09:00
Qiu Chaofan	5c18d11366	[SPE] Disable strict-fp for SPE by default As discussed in PR50385, strict-fp on PowerPC SPE has not been handled well. This patch disables it by default for SPE. Reviewed By: nemanjai, vit9696, jhibbits Differential Revision: https://reviews.llvm.org/D103235	2021-05-28 00:14:35 +08:00
thomasraoux	b44007bec2	[mlir][gpu] Relax restriction on MMA store op to allow chain of mma ops. In order to allow large matmul operations using the MMA ops we need to chain operations this is not possible unless "DOp" and "COp" type have matching layout so remove the "DOp" layout and force accumulator and result type to match. Added a test for the case where the MMA value is accumulated. Differential Revision: https://reviews.llvm.org/D103023	2021-05-27 09:13:51 -07:00
Yaxun (Sam) Liu	6d2c095020	[HIP] Check compatibility of -fgpu-sanitize with offload arch -fgpu-sanitize is incompatible with offload arch containing xnack-. This patch checks that. Reviewed by: Artem Belevich Differential Revision: https://reviews.llvm.org/D102975	2021-05-27 12:06:42 -04:00
Fraser Cormack	6f4794feb6	[RISCV] Add a test case showing incorrect call-conv lowering @HsiangKai helped find a bug in the lowering of indirect split scalable-vector types in our calling convention. An imminent patch will fix this.	2021-05-27 16:55:48 +01:00
Matt Arsenault	e892705d74	GlobalISel: Do not change register types in lowerLoad Adjusting the load register type is a widenScalar type action, not a lowering. lowerLoad should be reserved for operations that change the memory access size, such as unaligned load decomposition. With this trying to adjust the register type, it was hard to avoid infinite loops in the legalizer. Adds a bandaid to avoid regressing a few AArch64 tests, but I'm not sure what the exact condition is and there's probably a cleaner way to do this. For AMDGPU this regresses handling of some cases for unaligned loads, but the way this is currently working is a pretty ugly hack.	2021-05-27 11:49:37 -04:00
jasonliu	7922ff6010	[AIX] Add -lc++abi and -lunwind for linking Summary: We are going to have libc++abi.a and libunwind.a on AIX. Add the necessary linking command to pick the libraries up. Reviewed By: daltenty Differential Revision: https://reviews.llvm.org/D102813	2021-05-27 15:48:53 +00:00
Aaron Puchert	cf0b337c1b	Thread safety analysis: Allow exlusive/shared joins for managed and asserted capabilities Similar to how we allow managed and asserted locks to be held and not held in joining branches, we also allow them to be held shared and exclusive. The scoped lock should restore the original state at the end of the scope in any event, and asserted locks need not be released. We should probably only allow asserted locks to be subsumed by managed, not by (directly) acquired locks, but that's for another change. Reviewed By: delesley Differential Revision: https://reviews.llvm.org/D102026	2021-05-27 17:46:04 +02:00

1 2 3 4 5 ...

389685 Commits All Branches Search

389685 Commits

All Branches