llvm-project

Commit Graph

Author	SHA1	Message	Date
Joe Loser	67964fc4b2	[libc++][NFC] Replace tab with whitespace in comment There is a stray tab character in a comment block. Replace the tab character with a space for consistency with other comments.	2021-10-10 12:53:35 -04:00
Kazu Hirata	0e9373a6a6	[Basic] Use llvm::is_contained (NFC)	2021-10-10 08:52:14 -07:00
Sanjay Patel	05281d95f2	[InstCombine] move fold for "(X-Y) == 0"; NFC This consolidates related folds that all have a similar use restriction that may not be necessary.	2021-10-10 11:26:03 -04:00
Sanjay Patel	cbd8041b0b	[InstCombine] add tests for (X - Y) == 0; NFC	2021-10-10 11:13:46 -04:00
Sanjay Patel	da210f5d34	[InstCombine] canonicalize "(C2 - Y) > C" as (Y + ~C2) < ~C The test diffs show that we have better analysis/folds for 'add' (although we should at least have the simplifications independently, so we don't have the one-use restriction). This is related to solving regressions that would appear in transforms related to D111410, and that is part of a series of enhancements that may eventually helpi solve PR34047. https://alive2.llvm.org/ce/z/3tB9KG define i1 @src(i8 %x, i8 %C, i8 %C2) { %sub = sub nuw i8 %C2, %x %r = icmp slt i8 %sub, %C ret i1 %r } define i1 @tgt(i8 %x, i8 %C, i8 %C2) { %Cnot = xor i8 %C, -1 %C2not = xor i8 %C2, -1 %add = add nuw i8 %x, %C2not %r = icmp sgt i8 %add, %Cnot ret i1 %r }	2021-10-10 11:06:49 -04:00
Sanjay Patel	c00cab878a	[InstCombine] add test for or-of-icmps; NFC	2021-10-10 11:06:49 -04:00
Chen Zheng	4ead32d1cf	[PowerPC] update test case using the scripts; nfc	2021-10-10 14:39:20 +00:00
Mark de Wever	dcbfceffde	[libc++][nfc] Remove a duplicated include.	2021-10-10 14:21:01 +02:00
Dávid Bolvanský	e6ce86bb62	[NFC] Added tests for PR52056	2021-10-10 11:34:39 +02:00
william woodruff	e7fc254875	[BitcodeAnalyzer] allow a motivated user to dump BLOCKINFO This adds the `--dump-blockinfo` flag to `llvm-bcanalyzer`, allowing a sufficiently motivated user to dump (parts of) the `BLOCKINFO_BLOCK` block. The default behavior is unchanged, and `--dump-blockinfo` only takes effect in the same context as other flags that control dump behavior (i.e., requires that `--dump` is also passed). Reviewed By: tejohnson Differential Revision: https://reviews.llvm.org/D107536	2021-10-10 10:15:14 +05:30
Amara Emerson	f95d9c95bb	[GlobalISel] Fix the stores of truncates -> wide store combine for non-evenly dividing type sizes. If the wide store we'd generate is not a multiple of the memory type of the narrow stores (e.g. s48 and s32), we'd assert. Fix that.	2021-10-09 21:18:20 -07:00
william woodruff	451d0596d7	[clang] Fix JSON AST output when a filter is used Without this, the combination of `-ast-dump=json` and `-ast-dump-filter FILTER` produces invalid JSON: the first line is a string that says `Dumping $SOME_DECL_NAME: `. Reviewed By: aaron.ballman Differential Revision: https://reviews.llvm.org/D108441	2021-10-10 07:46:17 +05:30
Med Ismail Bennani	c26e53e129	[lldb/test] Disable 'TestScriptedProcess.py' on macOS This is disabling 'TestScriptedProcess.py' on macOS since it fails on Green Dragon: https://green.lab.llvm.org/green/view/LLDB/job/lldb-cmake/35974 Signed-off-by: Med Ismail Bennani <medismail.bennani@gmail.com>	2021-10-10 03:28:36 +02:00
Joe Loser	903b30fea2	[libc++][test] Remove empty {ind.move.subsumption.compile.pass.cpp} `{ind.move.subsumption.compile.pass.cpp}` was accidentally commited in https://reviews.llvm.org/D102639. Per the conversation on Discord in	2021-10-09 17:20:19 -04:00
Amy Zhuang	5ce368cfe2	[mlir] Vectorize induction variables 1. Add support to vectorize induction variables of loops that are not mapped to any vector dimension in SuperVectorize pass. 2. Fix a bug in getForInductionVarOwner. Reviewed By: dcaballe Differential Revision: https://reviews.llvm.org/D111370	2021-10-09 12:40:24 -07:00
mydeveloperday	3019898e0d	[clang-format][NFC] improve the visual of the "clang-formatted %" NOTE: some files are being removed from those files that are clang-formatted which means some lack of formatting is slipping through the net on reviews	2021-10-09 19:37:03 +01:00
Mehdi Amini	dda810c332	Fix a comment at call-site to match the declared parameter (NFC) (clang-tidy warning)	2021-10-09 17:57:53 +00:00
Ron Lieberman	d022f39d9f	[libomptarget][amdgpu][NFC] tweak a comment	2021-10-09 12:51:53 -04:00
Kazu Hirata	3e1c787b31	[IR] Remove arg_operands and getNumArgOperands (NFC) The last uses were removed on Oct 8, 2021 in commit `46ef2e0bf9`. This is a relanding of `b2ee408dde`.	2021-10-09 09:38:15 -07:00
Sanjay Patel	acafde09a3	[InstCombine] enhance icmp with sub folds There were 2 related but over-specified folds for: C1 - X == C One allowed multi-use but was limited to equal constants. The other allowed different constants but disallowed multi-use. This combines the 2 folds into a more general match. The test diffs show the multi-use cases that were falling through the cracks. https://alive2.llvm.org/ce/z/4_hEt2 define i1 @src(i8 %x, i8 %subC, i8 %C) { %s = sub i8 %subC, %x %r = icmp eq i8 %s, %C ret i1 %r } define i1 @tgt(i8 %x, i8 %subC, i8 %C) { %newC = sub i8 %subC, %C %isneg = icmp eq i8 %x, %newC ret i1 %isneg }	2021-10-09 11:39:49 -04:00
Sanjay Patel	cd76fa79b0	[InstCombine] add tests for icmp of negated op; NFC	2021-10-09 11:39:49 -04:00
Sanjay Patel	38e3b30bd6	[InstCombine] add tests for (iN X s>> N-1) \| Y; NFC These are for a sibling fold suggested in D111410. The tests correspond to the 'and' tests added with: `a35673f4cf`	2021-10-09 11:39:49 -04:00
Dávid Bolvanský	943b304848	Fixed some errors detected by PVS Studio	2021-10-09 17:27:41 +02:00
Dávid Bolvanský	3649fb14d1	Fixed some errors detected by PVS Studio	2021-10-09 17:20:04 +02:00
Nikita Popov	ea12adc169	[CanonicalizeFreeze] Drop IVUsers.h include (NFC) Looking for users of IVUsers, this was a false positive. Only LSR uses IVUsers.	2021-10-09 17:01:26 +02:00
David Green	adec922361	[AArch64] Make -mcpu=generic schedule for an in-order core We would like to start pushing -mcpu=generic towards enabling the set of features that improves performance for some CPUs, without hurting any others. A blend of the performance options hopefully beneficial to all CPUs. The largest part of that is enabling in-order scheduling using the Cortex-A55 schedule model. This is similar to the Arm backend change from `eecb353d0e` which made -mcpu=generic perform in-order scheduling using the cortex-a8 schedule model. The idea is that in-order cpu's require the most help in instruction scheduling, whereas out-of-order cpus can for the most part out-of-order schedule around different codegen. Our benchmarking suggests that hypothesis holds. When running on an in-order core this improved performance by 3.8% geomean on a set of DSP workloads, 2% geomean on some other embedded benchmark and between 1% and 1.8% on a set of singlecore and multicore workloads, all running on a Cortex-A55 cluster. On an out-of-order cpu the results are a lot more noisy but show flat performance or an improvement. On the set of DSP and embedded benchmarks, run on a Cortex-A78 there was a very noisy 1% speed improvement. Using the most detailed results I could find, SPEC2006 runs on a Neoverse N1 show a small increase in instruction count (+0.127%), but a decrease in cycle counts (-0.155%, on average). The instruction count is very low noise, the cycle count is more noisy with a 0.15% decrease not being significant. SPEC2k17 shows a small decrease (-0.2%) in instruction count leading to a -0.296% decrease in cycle count. These results are within noise margins but tend to show a small improvement in general. When specifying an Apple target, clang will set "-target-cpu apple-a7" on the command line, so should not be affected by this change when running from clang. This also doesn't enable more runtime unrolling like -mcpu=cortex-a55 does, only changing the schedule used. A lot of existing tests have updated. This is a summary of the important differences: - Most changes are the same instructions in a different order. - Sometimes this leads to very minor inefficiencies, such as requiring an extra mov to move variables into r0/v0 for the return value of a test function. - misched-fusion.ll was no longer fusing the pairs of instructions it should, as per D110561. I've changed the schedule used in the test for now. - neon-mla-mls.ll now uses "mul; sub" as opposed to "neg; mla" due to the different latencies. This seems fine to me. - Some SVE tests do not always remove movprfx where they did before due to different register allocation giving different destructive forms. - The tests argument-blocks-array-of-struct.ll and arm64-windows-calls.ll produce two LDR where they previously produced an LDP due to store-pair-suppress kicking in. - arm64-ldp.ll and arm64-neon-copy.ll are missing pre/postinc on LPD. - Some tests such as arm64-neon-mul-div.ll and ragreedy-local-interval-cost.ll have more, less or just different spilling. - In aarch64_generated_funcs.ll.generated.expected one part of the function is no longer outlined. Interestingly if I switch this to use any other scheduled even less is outlined. Some of these are expected to happen, such as differences in outlining or register spilling. There will be places where these result in worse codegen, places where they are better, with the SPEC instruction counts suggesting it is not a decrease overall, on average. Differential Revision: https://reviews.llvm.org/D110830	2021-10-09 15:58:31 +01:00
Nico Weber	e2a2e5475c	Revert "Reland "[gn build] (manually) port `6fe2beba7d` (ExceptionTests)"" This reverts commit `842035d8bd`. `1dba6b3` was reverted yet again in `04aff39504`.	2021-10-09 10:18:52 -04:00
Michał Górny	fefd0ca31d	[lldb] [DynamicRegisterInfo] Remove obsolete dwarf typedefs (NFC)	2021-10-09 15:42:34 +02:00
Raphael Isemann	b5ff511048	[lldb][NFC] Early-exit in DWARFASTParserClang::ParseSingleMember ParseSingleMember has two large ifs around the back of it's body: `if (!is_artificial)` and `if (member_type)`. This patch just converts those to early-exits. The patch is NFC. It even retains the curious fact that Objective-C properties that fail to parse are silently ignored, but now there is at least a FIXME that points this out.	2021-10-09 14:40:39 +02:00
Aaron Ballman	af971365a2	Fix a diagnoses-valid in C++20 with variadic macros C++20 and later allow you to pass no argument for the ... parameter in a variadic macro, whereas earlier language modes and C disallow it. We no longer diagnose in C++20 and later modes. This fixes PR51609.	2021-10-09 08:20:20 -04:00
Mark de Wever	a1f0f847ff	[NFC][libc++] Update back_insert_iterator style. As suggested in D110573 land the rename part separately.	2021-10-09 13:31:20 +02:00
Mark de Wever	b67a8a6513	[libc++][doc] Update format status. Updated based on recent commits.	2021-10-09 13:28:38 +02:00
mydeveloperday	3e553791ca	[clang-format][NFC] Fix spelling mistakes	2021-10-09 12:27:08 +01:00
Frederic Cambus	6417260a57	[Driver][OpenBSD] Use ToolChain reference instead of getToolChain(). Differential Revision: https://reviews.llvm.org/D111462	2021-10-09 13:21:39 +02:00
mydeveloperday	bbf4b3dbbe	[clang-format][NFC] Fix spelling mistake	2021-10-09 12:18:25 +01:00
mydeveloperday	a2a826d8b6	[clang-format][docs][NFC] correct the "first supported versions" of some of the clang-format options Some of the first supported version field were incorrectly attributed to a later branch. It wasn't possible to correctly determine the "introduced version" with my naive implementation using git blame alone, (especially if the type had been changed from a bool -> enum) I saw more things attributed to clang-format 13 than I remembered and reviewed those options to determine their introduced version. Reviewed By: HazardyKnusperkeks Differential Revision: https://reviews.llvm.org/D110803	2021-10-09 11:02:49 +01:00
Nikita Popov	a94002cd64	[Type] Avoid APFloat.h include (NFC) This is only used by a handful of methods working on fltSemantics, and having these defined inline in the header does not look particularly important.	2021-10-09 11:29:26 +02:00
Nikita Popov	55b9146848	[MCPseudoProbe] Clean up includes (NFC) This was including various things that don't appear to be used in the header at all.	2021-10-09 10:31:15 +02:00
luxufan	02ac5e5cf1	[Orc] Fix global variable destructor function support when --jit-kind=orc-lazy The bug was reported here https://bugs.llvm.org/show_bug.cgi?id=52030 This patch follows the idea that @lhames commented in the above webpage. Reviewed By: lhames Differential Revision: https://reviews.llvm.org/D110990	2021-10-09 15:58:21 +08:00
Max Kazantsev	4c0da23663	[LoopDeletion] Support selects when symbolically evaluating 1st iteration Adds support for selects for which we know value on the 1st iteration. Differential Revision: https://reviews.llvm.org/D104111 Reviewed By: nikic	2021-10-09 14:47:44 +07:00
Max Kazantsev	49ca01047f	[Test] Add commit justifying revert of D110922 Test by Arthur Eubanks!	2021-10-09 14:32:46 +07:00
luxufan	590326382d	[Orc] Support atexit in Orc(JITLink) There is a bug reported at https://bugs.llvm.org/show_bug.cgi?id=48938 After looking through the glibc, I found the `atexit(f)` is the same as `__cxa_atexit(f, NULL, NULL)`. In orc runtime, we identify different JITDylib by their dso_handle value, so that a NULL dso_handle is invalid. So in this patch, I added a `PlatformJDDSOHandle` to ELFNixRuntimeState, and functions which are registered by atexit will be registered at PlatformJD. Reviewed By: lhames Differential Revision: https://reviews.llvm.org/D111413	2021-10-09 12:25:47 +08:00
william woodruff	778bf73d7b	[BitcodeReader] fix a logic error in vector type element validation The current code checks whether the vector's element type is a valid structure element type, rather than a valid vector element type. The two have separate implementations and but only accept very slightly different sets of types, which is probably why this wasn't caught before. Differential Revision: https://reviews.llvm.org/D109655	2021-10-09 09:42:02 +05:30
Brad Smith	65df10f3cd	[OpenBSD] Use cortex-a8 as default CPU for ARMv7	2021-10-08 23:57:40 -04:00
hsmahesha	0481682996	[CFE][Codegen][In-progress] Remove CodeGenFunction::InitTempAlloca() CodeGenFunction::InitTempAlloca() inits the static alloca within the entry block which may not necessarily be correct always. For example, the current instruction insertion point (pointed by the instruction builder) could be a program point which is hit multiple times during the program execution, and it is expected that the static alloca is initialized every time the program point is hit. Hence remove CodeGenFunction::InitTempAlloca(), and initialize the static alloca where the instruction insertion point is at the moment. This patch, as a starting attempt, removes the calls to CodeGenFunction::InitTempAlloca() which do not have any side effect on the lit tests. Reviewed By: rjmccall Differential Revision: https://reviews.llvm.org/D111293	2021-10-09 09:23:14 +05:30
Michael Kruse	203c7fab73	[Polly] Fix test case fixing the colon. Commit `573531fb1f` fixed the colon at the end of a CHECK line (was a semicolon by mistake). With the check enabled, it turned out that it was failing. Check for the correct content. Also add the missing colon to the next CHECK line.	2021-10-08 22:46:55 -05:00
Qiu Chaofan	da0b62dfb3	Revert a LIT typo fix in a RUN line Commit `573531f` changes the behavior of the test, revert it back.	2021-10-09 11:29:44 +08:00
Mehdi Amini	8c9f506d8c	Disable mlir/test/mlir-cpu-runner/async-group.mlir with ASAN This test is crashing 9 out of 10 runs in CI, but I can't reproduce locally right now. Disabling to get the CI back to green and avoid backsliding with more ASAN issues that would go unnoticed.	2021-10-09 03:02:53 +00:00
Richard Smith	7eae8c6e62	Don't update the vptr at the start of the destructor of a final class. In this case, we know statically that we're destroying the most-derived class, so the vptr must already point to the current class and never needs to be updated.	2021-10-08 19:59:42 -07:00
Qiu Chaofan	85e565898f	[Clang] Enable _Complex __ibm128 type `fae0dfa` implemented the new __ibm128 type, this patch enables its complex form. Reviewed By: rjmccall Differential Revision: https://reviews.llvm.org/D109948	2021-10-09 10:48:44 +08:00

1 2 3 4 5 ...

401393 Commits All Branches Search

401393 Commits

All Branches