llvm-project

Commit Graph

Author	SHA1	Message	Date
Alexey Bataev	195c97e220	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. Summary: If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320510	2017-12-12 18:47:00 +00:00
Sanjoy Das	81a4a02cbc	Revert "[X86] Flag BroadWell scheduler model as complete" This reverts commit r320308. r320308 crashes LLC, please see the llvm-commits thread for a reproducer. llvm-svn: 320508	2017-12-12 18:40:58 +00:00
Craig Topper	712a209db9	[X86] Add a couple TODOs about missing coverage/features motivated by D40335 D40335 was wanting to add FMSUBADD support, but it discovered that there are two pieces of code to make FMADDSUB and only one of those is tested. So I've asked that review to implement the one path until we get tests that test the existing code. llvm-svn: 320507	2017-12-12 18:39:04 +00:00
Nirav Dave	674d053d18	[X86] Cleanup type conversion of 64-bit load-store pairs. Summary: Simplify and generalize chain handling and search for 64-bit load-store pairs. Nontemporal test now converts 64-bit integer load-store into f64 which it realizes directly instead of splitting into two i32 pairs. Reviewers: craig.topper, spatel Reviewed By: craig.topper Subscribers: hiraditya, llvm-commits Differential Revision: https://reviews.llvm.org/D40918 llvm-svn: 320505	2017-12-12 18:25:48 +00:00
Alexandre Ganea	757026dbe6	Test commit. llvm-svn: 320504	2017-12-12 18:00:43 +00:00
Geoff Berry	60c431022e	[MachineOperand][MIR] Add isRenamable to MachineOperand. Summary: Add isRenamable() predicate to MachineOperand. This predicate can be used by machine passes after register allocation to determine whether it is safe to rename a given register operand. Register operands that aren't marked as renamable may be required to be assigned their current register to satisfy constraints that are not captured by the machine IR (e.g. ABI or ISA constraints). Reviewers: qcolombet, MatzeB, hfinkel Subscribers: nemanjai, mcrosier, javed.absar, llvm-commits Differential Revision: https://reviews.llvm.org/D39400 llvm-svn: 320503	2017-12-12 17:53:59 +00:00
Alexey Bataev	6132a50d2a	Revert "[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast." This reverts commit r320499 again to resolve the problem with the sanitizers bbots. llvm-svn: 320501	2017-12-12 17:35:29 +00:00
Alexey Bataev	ca4c9a5246	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. Summary: If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320499	2017-12-12 17:19:15 +00:00
Alexey Bataev	d19dbe6791	Revert "[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast." This reverts commit r320496 to solve the problems with sanitizer buildbots. llvm-svn: 320498	2017-12-12 17:08:48 +00:00
Don Hinton	49777fa933	[cmake] Support moving debuginfo-tests to llvm/projects Differential Revision: https://reviews.llvm.org/D40972 llvm-svn: 320497	2017-12-12 17:06:08 +00:00
Alexey Bataev	d0c3aeb200	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. Summary: If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320496	2017-12-12 16:58:48 +00:00
Simon Pilgrim	68f9accf51	[X86] Remove CompleteModel tags from CPU targets until we have better error checking (PR35636) The checks we have for complete models are not great and miss many cases - e.g. in PR35636 it failed to recognise that only the first output (of 2) was actually tagged by the InstRW Raised PR35639 and PR35643 as examples llvm-svn: 320492	2017-12-12 16:12:53 +00:00
Alex Bradbury	c01db1ce8f	[RISCV][NFC] Formatting fix in RISCVInstrInfo.td llvm-svn: 320491	2017-12-12 16:10:21 +00:00
Alexey Bataev	c9f1d2e4a0	Revert "[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast." This reverts commit r320488 because of the failed asan buildbots.. llvm-svn: 320490	2017-12-12 16:05:52 +00:00
Alexey Bataev	fb68c48a82	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. Summary: If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320488	2017-12-12 15:54:49 +00:00
Alex Bradbury	9ed84c8ae8	[RISCV] Implement assembler pseudo instructions for RV32I and RV64I Adds the assembler pseudo instructions of RV32I and RV64I which can be mapped to a single canonical instruction. The missing pseudo instructions (e.g., call, tail, ...) are marked as TODO. Other things, like for example PCREL_LO, have to be implemented first. Currently, alias emission is disabled by default to keep the patch minimal. Alias emission by default will be enabled in a subsequent patch which also updates all affected tests. Note that this patch should actually break the floating point MC tests. However, the used FileCheck configuration is not tight enought to detect the breakage. Differential Revision: https://reviews.llvm.org/D40902 Patch by Mario Werner. llvm-svn: 320487	2017-12-12 15:46:15 +00:00
Alexey Bataev	ca2a8cea2f	Revert "[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast." This reverts commit r320483 because of the failed Windows buildbots. llvm-svn: 320485	2017-12-12 15:24:17 +00:00
Alex Bradbury	8bba6bfeef	[RISCV] MC layer support for the instructions added in the privileged spec Adds support for the instructions added in the RISC-V privileged ISA (https://content.riscv.org/wp-content/uploads/2017/05/riscv-privileged-v1.10.pdf): uret, sret, mret, wfi, and sfence.vma. Note from the committer: I made very minor formatting changes prior to commit, which didn't seem worth creating another review round-trip for. Differential Revision: https://reviews.llvm.org/D40383 Patch by David Craven. llvm-svn: 320484	2017-12-12 15:17:45 +00:00
Alexey Bataev	1daef8a667	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320483	2017-12-12 15:03:17 +00:00
Ayman Musa	c2eed926b0	[X86] Recognize constant arrays with special values and replace loads from it with subtract and shift instructions, which then will be replaced by X86 BZHI machine instruction. Recognize constant arrays with the following values: 0x0, 0x1, 0x3, 0x7, 0xF, 0x1F, .... , 2^(size - 1) -1 where //size// is the size of the array. the result of a load with index //idx// from this array is equivalent to the result of the following: (0xFFFFFFFF >> (sub 32, idx)) (assuming the array of type 32-bit integer). And the result of an 'AND' operation on the returned value of such a load and another input, is exactly equivalent to the X86 BZHI instruction behavior. See test cases in the LIT test for better understanding. Differential Revision: https://reviews.llvm.org/D34141 llvm-svn: 320481	2017-12-12 14:13:51 +00:00
Anna Thomas	2dd9835f35	[InstComineLoadStoreAlloca] Optimize stores to GEP off null base Summary: Currently, in InstCombineLoadStoreAlloca, we have simplification rules for the following cases: 1. load off a null 2. load off a GEP with null base 3. store to a null This patch adds support for the fourth case which is store into a GEP with null base. Since this is UB as well (and directly analogous to the load off a GEP with null base), we can substitute the stored val with undef in instcombine, so that SimplifyCFG can optimize this code into unreachable code. Note: Right now, simplifyCFG hasn't been taught about optimizing this to unreachable and adding an llvm.trap (this is already done for the above 3 cases). Reviewers: majnemer, hfinkel, sanjoy, davide Reviewed by: sanjoy, davide Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41026 llvm-svn: 320480	2017-12-12 14:12:33 +00:00
Nemanja Ivanovic	6479c72fcd	[PowerPC] Add branch flag on asm parser-only branch instructions This flag was missing but it wasn't an issue as nothing depended on it for these asm parser-only instructions. Now that LLDB support is slowly landing, it is important to get this right. Committing on behalf of Leonardo Bianconi. Differential revision: https://reviews.llvm.org/D40846 llvm-svn: 320475	2017-12-12 12:33:09 +00:00
Nemanja Ivanovic	b0783cccb7	[PowerPC] Follow-up to r318436 to get the missed CSE opportunities The last of the three patches that https://reviews.llvm.org/D40348 was broken up into. Canonicalize the materialization of constants so that they are more likely to be CSE'd regardless of the bit-width of the use. If a constant can be materialized using PPC::LI, materialize it the same way always. For example: li 4, -1 li 4, 255 li 4, 65535 are equivalent if the uses only use the low byte. Canonicalize it to the first form. Differential Revision: https://reviews.llvm.org/D40348 llvm-svn: 320473	2017-12-12 12:09:34 +00:00
Simon Pilgrim	0f8a5a41cf	Revert r320461 - causing ICE in windows buildss [X86] Use regular expressions more aggressively to reduce the number of scheduler entries needed for FMA3 instructions. When the scheduler tables are generated by tablegen, the instructions are divided up into groups based on their default scheduling information and how they are referenced by groups for each processor. For any set of instructions that are matched by a specific InstRW line, that group of instructions is guaranteed to not be in a group with any other instructions. So in general, the more InstRW class definitions are created, the more groups we end up with in the generated files. Particularly if a lot of the InstRW lines only match to single instructions, which is true of a large number of the Intel scheduler models. This change alone reduces the number of instructions groups from ~6000 to ~5500. And there's lots more we could do. llvm-svn: 320470	2017-12-12 11:34:25 +00:00
Jonas Devlieghere	f0945f48bd	[dsymutil] Accept line tables up to DWARFv5. This patch removes the hard-coded check for DWARFv2 line tables. Now dsymutil accepts line tables for DWARF versions 2 to 5 (inclusive). Differential revision: https://reviews.llvm.org/D41084 rdar://35968319 llvm-svn: 320469	2017-12-12 11:32:21 +00:00
Eugene Leviant	d53f3da772	Revert r320464 as it breaks gold plugin tests llvm-svn: 320467	2017-12-12 10:12:46 +00:00
Igor Laevsky	d63560b817	Revert r320049, r320014 and r319894 They were causing failures of the piglit OpenGL tests with AMD GPUs using the Mesa radeonsi driver. llvm-svn: 320466	2017-12-12 10:03:39 +00:00
Serguei Katkov	f4ceb77cd9	[NFC][SafepointIRVerifier] Add alias for set of available values Introduces usage of AvailableValueSet alias name instead of DenseSet<const Value *> for better reading. Patch Author: Daniil Suchkov Reviewers: mkazantsev, anna, apilipenko Reviewed By: anna Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41002 llvm-svn: 320465	2017-12-12 09:44:41 +00:00
Eugene Leviant	3695183395	[ThinLTO] Remove unused code from thinLTOInternalizeModule Differential revision: https://reviews.llvm.org/D40970 llvm-svn: 320464	2017-12-12 09:12:32 +00:00
Dorit Nuzman	927b31600e	[LV] Ignore the cost of values that will not appear in the vectorized loop VecValuesToIgnore holds values that will not appear in the vectorized loop. We should therefore ignore their cost when VF > 1. Differential Revision: https://reviews.llvm.org/D40883 llvm-svn: 320463	2017-12-12 08:57:43 +00:00
Craig Topper	c8e64ab539	[X86] Use regular expressions more aggressively to reduce the number of scheduler entries needed for FMA3 instructions. When the scheduler tables are generated by tablegen, the instructions are divided up into groups based on their default scheduling information and how they are referenced by groups for each processor. For any set of instructions that are matched by a specific InstRW line, that group of instructions is guaranteed to not be in a group with any other instructions. So in general, the more InstRW class definitions are created, the more groups we end up with in the generated files. Particularly if a lot of the InstRW lines only match to single instructions, which is true of a large number of the Intel scheduler models. This change alone reduces the number of instructions groups from ~6000 to ~5500. And there's lots more we could do. llvm-svn: 320461	2017-12-12 08:17:04 +00:00
Mikael Holmen	66cf383761	[CallSiteSplitting] Don't let debug intrinsics affect optimizations Summary: This solves PR35616. We don't want the compiler to generate different code when we compile with/without -g, so we now ignore debug intrinsics when determining if the optimization can trigger or not. Reviewers: junbuml Subscribers: davide, JDevlieghere, llvm-commits Differential Revision: https://reviews.llvm.org/D41068 llvm-svn: 320460	2017-12-12 07:29:57 +00:00
Craig Topper	468a813315	[X86] Use Ld scheduler classes for instructions with folded loads. llvm-svn: 320459	2017-12-12 07:06:35 +00:00
Craig Topper	c1e72c019d	[X86] Correct the FMA3 regular expressions in the znver1 scheduler model. llvm-svn: 320458	2017-12-12 07:06:32 +00:00
Tony Tye	a697880b38	[AMDGPU] Rename Bonaire target to be gfx704; remove gfx800 and make Iceland and Tonga both use gfx802; update target feature handling Correct committed version to match intended accepted review D40051 id=123417 - Rename Bonaire target to be gfx704. - Eliminate gfx800 and make Iceland and Tonga both use gfx802 as they use the same code. - List target features supported by each processor in the processor table together with the default value. - Add xnack flag to e_flags. - Remove xnack from kernel metadata and kernel descriptor since it is now a whole code object property. Differential Revision: https://reviews.llvm.org/D40051 llvm-svn: 320457	2017-12-12 05:47:00 +00:00
Vedant Kumar	7a911b5851	[llvm-cov] Simplify a test case. NFC. llvm-svn: 320439	2017-12-11 23:34:50 +00:00
Max Moroz	fe4d904917	[llvm-cov] Add an option for "export" command to emit only file summary data. Summary: That allows to get the same data as produced by "llvm-cov report", but in JSON format, which is better for further processing by end users. Reviewers: vsk Reviewed By: vsk Differential Revision: https://reviews.llvm.org/D41085 llvm-svn: 320435	2017-12-11 23:17:46 +00:00
Sam Clegg	f950b24a7a	Reland "[WebAssembly] Import the linear memory and function table." Original change: https://reviews.llvm.org/D40875 llvm-svn: 320432	2017-12-11 23:03:38 +00:00
Richard Trieu	efef032f02	Revert r318704 - [Sparc] efficient pattern for UINT_TO_FP conversion See bug https://bugs.llvm.org/show_bug.cgi?id=35631 r318704 is giving a fatal error on some code with unsigned to floating point conversions. llvm-svn: 320429	2017-12-11 22:25:04 +00:00
Matt Arsenault	3e268cc0dd	LSR: Check more intrinsic pointer operands llvm-svn: 320424	2017-12-11 21:38:43 +00:00
Hans Wennborg	27d1c00c01	Revert r320407 "[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast." The tests fail (opt asserts) on Windows. > Summary: > If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, > &V2)))), bitcast)`, but the load is used in other instructions, it leads > to looping in InstCombiner. Patch adds additional check that all users > of the load instructions are stores and then replaces all uses of load > instruction by the new one with new type. > > Reviewers: RKSimon, spatel, majnemer > > Subscribers: llvm-commits > > Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320421	2017-12-11 21:15:27 +00:00
Evandro Menezes	54be62df39	[CodeGen] Improve the consistency of instruction fusion* When either instruction in a fused pair has no other dependency, besides on the other instruction, make sure that other instructions do not get scheduled between them. Additionally, avoid fusing an instruction more than once along the same dependency chain. Differential revision: https://reviews.llvm.org/D36704 llvm-svn: 320420	2017-12-11 21:09:27 +00:00
Adrian Prantl	3c6c14d14b	ASAN: Provide reliable debug info for local variables at -O0. The function stack poisioner conditionally stores local variables either in an alloca or in malloc'ated memory, which has the unfortunate side-effect, that the actual address of the variable is only materialized when the variable is accessed, which means that those variables are mostly invisible to the debugger even when compiling without optimizations. This patch stores the address of the local stack base into an alloca, which can be referred to by the debug info and is available throughout the function. This adds one extra pointer-sized alloca to each stack frame (but mem2reg can optimize it away again when optimizations are enabled, yielding roughly the same debug info quality as before in optimized code). rdar://problem/30433661 Differential Revision: https://reviews.llvm.org/D41034 llvm-svn: 320415	2017-12-11 20:43:21 +00:00
Tony Jiang	3b49dc548f	[PowerPC] Partially enable the ISEL expansion pass. The pass to expand ISEL instructions into if-then-else sequences in patch D23630 is currently disabled. This patch partially enable it by always removing the unnecessary ISELs (all registers used by the ISELs are the same one) and folding the ISELs which have the same input registers into unconditional copies. Differential Revision: https://reviews.llvm.org/D40497 llvm-svn: 320414	2017-12-11 20:42:37 +00:00
Justin Bogner	b608076e56	[cmake] Pass TARGETS_TO_BUILD through to host tools build In r319620, the host build was changed to use Native for TARGETS_TO_BUILD because passing semicolons through add_custom_command is surprisingly difficult. However, Native really doesn't make any sense here, and it only works because we don't technically do any codegen in the host tools so pretty well anything will "work". The problem here is that passing something other than the correct value is very fragile - as evidence note how the llvm-config in the host tools acts differently than the target one now, and misreports the targets to build. Similarly, if there is any logic conditional on the targets in tablegen (now or in the future), it will do the wrong thing. To fix this, we need to escape the semicolons in the targets string and pass it through to the child cmake invocation. llvm-svn: 320413	2017-12-11 19:53:23 +00:00
George Burgess IV	8c5886b45f	Ensure moved-from container is cleared on move In all cases except for this optimistic attempt to reuse memory, the moved-from TinyPtrVector was left `empty()` at the end of this assignment. Though using a container after it's been moved from can be a bit sketchy, it's probably best to just be consistent here. llvm-svn: 320408	2017-12-11 19:22:59 +00:00
Alexey Bataev	ec128ace8a	[InstCombine] Fix PR35618: Instcombine hangs on single minmax load bitcast. Summary: If we have pattern `store (load(bitcast(select (cmp(V1, V2), &V1, &V2)))), bitcast)`, but the load is used in other instructions, it leads to looping in InstCombiner. Patch adds additional check that all users of the load instructions are stores and then replaces all uses of load instruction by the new one with new type. Reviewers: RKSimon, spatel, majnemer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D41072 llvm-svn: 320407	2017-12-11 19:11:16 +00:00
Krzysztof Parzyszek	a8ab1b75cb	[Hexagon] Add support for Hexagon V65 llvm-svn: 320404	2017-12-11 18:57:54 +00:00
Simon Pilgrim	e83876e31d	[X86] Add LODS schedule tests llvm-svn: 320403	2017-12-11 18:39:42 +00:00
Simon Pilgrim	e8715025f5	[X86] Add CMP/TEST schedule tests llvm-svn: 320402	2017-12-11 18:32:59 +00:00
Simon Pilgrim	5512525c5d	[X86] Add AND/OR/XOR schedule tests llvm-svn: 320400	2017-12-11 18:23:24 +00:00
Jonas Devlieghere	ba915897da	[dwarfdump] Fix off-by-one bug in accelerator table extractor. This fixes a bug where the verifier was complaining about empty accelerator tables. When the table is empty, its size is not a valid offset as it points after the end of the section. This patch also makes the extractor return llvm:Error instead of bool for better error reporting in the verifier. Differential revision: https://reviews.llvm.org/D41063 rdar://35932007 llvm-svn: 320399	2017-12-11 18:22:47 +00:00
Simon Pilgrim	9b2a5e1e0b	[X86] Add ADD/SUB schedule tests llvm-svn: 320397	2017-12-11 18:13:40 +00:00
Simon Pilgrim	dbe6c45fcd	[X86] Add ADC/SBB schedule tests llvm-svn: 320395	2017-12-11 17:59:05 +00:00
Simon Pilgrim	8c2d90a2f4	[X86] Add MOVSLQ schedule tests llvm-svn: 320392	2017-12-11 17:37:08 +00:00
Simon Pilgrim	6d89f407db	Normalize line endings. NFCI. llvm-svn: 320389	2017-12-11 17:01:21 +00:00
Amara Emerson	df9b529d42	[GlobalISel] Disable GISel for big endian. This is due to PR26161 needing to be resolved before we can fix big endian bugs like PR35359. The work to split aggregates into smaller LLTs instead of using one large scalar will take some time, so in the mean time we'll fall back to SDAG. Some ARM BE tests xfailed for now as a result. Differential Revision: https://reviews.llvm.org/D40789 llvm-svn: 320388	2017-12-11 16:58:29 +00:00
Simon Pilgrim	fabe354b42	[X86] Add LWP schedule tests Tag LWP instructions as WriteSystem llvm-svn: 320387	2017-12-11 16:47:21 +00:00
Simon Pilgrim	67644be692	[X86] Add INT/INTO schedule tests llvm-svn: 320386	2017-12-11 16:32:58 +00:00
Simon Pilgrim	1fe82016a2	[X86] Add IN/OUT schedule tests llvm-svn: 320385	2017-12-11 16:16:40 +00:00
Simon Pilgrim	d0ce975528	[X86] Add IDIV schedule tests llvm-svn: 320384	2017-12-11 16:08:21 +00:00
Simon Pilgrim	6c29962f2e	[X86] Add CMPXCHG schedule tests llvm-svn: 320383	2017-12-11 16:04:08 +00:00
Simon Pilgrim	1c83cd18ae	[X86] Add CLZERO schedule test llvm-svn: 320382	2017-12-11 15:53:12 +00:00
Alexander Potapenko	3c934e4864	[MSan] Hotfix compilation For some reason the override directives got removed in r320373. I suspect this to be an unwanted effect of clang-format. llvm-svn: 320381	2017-12-11 15:48:56 +00:00
Simon Pilgrim	d9d37f8c3c	[X86] Add ADCX/ADOX/XADD/XLAT schedule tests llvm-svn: 320380	2017-12-11 15:41:52 +00:00
Nirav Dave	e830b758b8	[X86] Modify Nontemporal tests to avoid deadstore optimization. llvm-svn: 320379	2017-12-11 15:35:40 +00:00
Tony Tye	31105cc997	[AMDGPU] Rename Bonaire target to be gfx704; update target feature handling - Rename Bonaire target to be gfx704. - Eliminate gfx800 and make Iceland and Tonga both use gfx802 as they use the same code. - List target features supported by each processor in the processor table together with the default value. - Add xnack flag to e_flags. - Remove xnack from kernel metadata and kernel descriptor since it is now a whole code object property. Differential Revision: https://reviews.llvm.org/D40051 llvm-svn: 320378	2017-12-11 15:35:27 +00:00
Simon Pilgrim	4f2c415a13	[X86] Add SETCC/STC/STD/UD2 schedule tests llvm-svn: 320376	2017-12-11 15:25:31 +00:00
Dmitry Preobrazhensky	ac2b02643b	[AMDGPU][MC][GFX9] Corrected encoding of ttmp registers, disabled tba/tma See bugs 35494 and 35559: https://bugs.llvm.org/show_bug.cgi?id=35494 https://bugs.llvm.org/show_bug.cgi?id=35559 Reviewers: vpykhtin, artem.tamazov, arsenm Differential Revision: https://reviews.llvm.org/D41007 llvm-svn: 320375	2017-12-11 15:23:20 +00:00
Sanjay Patel	f3436d7dab	[DAGCombiner] protect against an infinite loop between shl <--> mul (PR35579) At first, I tried to thread the x86 needle and use a target hook (isVectorShiftByScalarCheap()) to disable the transform only for non-splat pow-of-2 constants, but not AVX2, but only some element types, but...it's difficult. Here we just avoid the loop with the x86 vector transform that conflicts with the general DAG combine and preserve all of the existing behavior AFAICT otherwise. Some tests that will probably fail if someone does try to restrict this in a more targeted way for x86-only may be found in: test/CodeGen/X86/combine-mul.ll test/CodeGen/X86/vector-mul.ll test/CodeGen/X86/widen_arith-5.ll This should prevent the infinite looping seen with: https://bugs.llvm.org/show_bug.cgi?id=35579 Differential Revision: https://reviews.llvm.org/D41040 llvm-svn: 320374	2017-12-11 15:19:31 +00:00
Alexander Potapenko	c07e6a0eff	[MSan] introduce getShadowOriginPtr(). NFC. This patch introduces getShadowOriginPtr(), a method that obtains both the shadow and origin pointers for an address as a Value pair. The existing callers of getShadowPtr() and getOriginPtr() are updated to use getShadowOriginPtr(). The rationale for this change is to simplify KMSAN instrumentation implementation. In KMSAN origins tracking is always enabled, and there's no direct mapping between the app memory and the shadow/origin pages. Both the shadow and the origin pointer for a given address are obtained by calling a single runtime hook from the instrumentation, therefore it's easier to work with those pointers together. Reviewed at https://reviews.llvm.org/D40835. llvm-svn: 320373	2017-12-11 15:05:22 +00:00
Simon Pilgrim	5154d249a8	[X86] Add SAR/SHL/SHR schedule tests llvm-svn: 320371	2017-12-11 14:56:44 +00:00
Simon Pilgrim	426add6915	[X86] Add RCL/RCR schedule tests llvm-svn: 320370	2017-12-11 14:46:42 +00:00
Krzysztof Parzyszek	152414595b	[Hexagon] Crash in instruction selection for insert_vector_elt for HVX A wrong type was passed to insertVector, causing an out-of-bounds value to be added an an operand to HexagonISD::INSERT. This later failed in instruction selection. llvm-svn: 320369	2017-12-11 14:46:06 +00:00
Nemanja Ivanovic	50d37a1129	[PowerPC] Sign-extend negative constant stores Second part of https://reviews.llvm.org/D40348. Revision r318436 has extended all constants feeding a store to 64 bits to allow for CSE on the SDAG. However, negative constants were zero extended which made the constant being loaded appear to be a positive value larger than 16 bits. This resulted in long sequences to materialize such constants rather than simply a "load immediate". This patch just sign-extends those updated constants so that they remain 16-bit signed immediates if they started out that way. llvm-svn: 320368	2017-12-11 14:35:48 +00:00
Nemanja Ivanovic	25d9af0cb5	[DAGCombiner] Add combined indexed load to the work list This commit is the first part of https://reviews.llvm.org/D40348. In order to allow target combines to be performed on newly combined indexed loads, add them back to the worklist. The remainder of the above patch will be committed in subsequent revisions and will use this. Test cases will be included with those follow-up commits. llvm-svn: 320365	2017-12-11 14:16:02 +00:00
Diana Picus	291e8d924f	[ARM GlobalISel] Add test for a MOVTi16 pattern. NFC Add test for matching an OR with 0xFFFF0000 to a MOVTi16. llvm-svn: 320362	2017-12-11 13:28:45 +00:00
Simon Pilgrim	969850f514	[X86] Add fsgsbase schedule tests. llvm-svn: 320361	2017-12-11 13:25:02 +00:00
Alex Bradbury	dc31c61b18	[RISCV] Add custom CC_RISCV calling convention and improved call support The TableGen-based calling convention definitions are inflexible, while writing a function to implement the calling convention is very straight-forward, and allows difficult cases to be handled more easily. With this patch adds support for: * Passing large scalars according to the RV32I calling convention * Byval arguments * Passing values on the stack when the argument registers are exhausted The custom CC_RISCV calling convention is also used for returns. This patch also documents the ABI lowering that a language frontend is expected to perform. I would like to work to simplify these requirements over time, but this will require further discussion within the LLVM community. We add PendingArgFlags CCState, as a companion to PendingLocs. The PendingLocs vector is used by a number of backends to handle arguments that are split during legalisation. However CCValAssign doesn't keep track of the original argument alignment. Therefore, add a PendingArgFlags vector which can be used to keep track of the ISD::ArgFlagsTy for every value added to PendingLocs. Differential Revision: https://reviews.llvm.org/D39898 llvm-svn: 320359	2017-12-11 12:49:02 +00:00
Alex Bradbury	bfb00d4c1c	[RISCV] Allow lowering of dynamic_stackalloc, stacksave, stackrestore llvm-svn: 320358	2017-12-11 12:38:17 +00:00
Alex Bradbury	b014e3de52	[RISCV] Implement prolog and epilog insertion As frame pointer elimination isn't implemented until a later patch and we make extensive use of update_llc_test_checks.py, this changes touches a lot of the RISC-V tests. Differential Revision: https://reviews.llvm.org/D39849 llvm-svn: 320357	2017-12-11 12:34:11 +00:00
Simon Pilgrim	220b1c13bf	[X86] Regenerate fsgsbase intrinsic tests. NFCI. llvm-svn: 320356	2017-12-11 12:22:15 +00:00
Roger Ferrer Ibanez	5ea0f2501f	[ARM] Use ADDCARRY / SUBCARRY This is a preparatory step for D34515. This change: - makes nodes ISD::ADDCARRY and ISD::SUBCARRY legal for i32 - lowering is done by first converting the boolean value into the carry flag using (_, C) ← (ARMISD::ADDC R, -1) and converted back to an integer value using (R, _) ← (ARMISD::ADDE 0, 0, C). An ARMISD::ADDE between the two operations does the actual addition. - for subtraction, given that ISD::SUBCARRY second result is actually a borrow, we need to invert the value of the second operand and result before and after using ARMISD::SUBE. We need to invert the carry result of ARMISD::SUBE to preserve the semantics. - given that the generic combiner may lower ISD::ADDCARRY and ISD::SUBCARRYinto ISD::UADDO and ISD::USUBO we need to update their lowering as well otherwise i64 operations now would require branches. This implies updating the corresponding test for unsigned. - add new combiner to remove the redundant conversions from/to carry flags to/from boolean values (ARMISD::ADDC (ARMISD::ADDE 0, 0, C), -1) → C - fixes PR34045 - fixes PR34564 - fixes PR35103 Differential Revision: https://reviews.llvm.org/D35192 llvm-svn: 320355	2017-12-11 12:13:45 +00:00
Alex Bradbury	660bcceccf	[RISCV] Support lowering FrameIndex Introduces the AddrFI "addressing mode", which is necessary simply because it's not possible to write a pattern that directly matches a frameindex. Ensure callee-saved registers are accessed relative to the stackpointer. This is necessary as callee-saved register spills are performed before the frame pointer is set. Move HexagonDAGToDAGISel::isOrEquivalentToAdd to SelectionDAGISel, so we can make use of it in the RISC-V backend. Differential Revision: https://reviews.llvm.org/D39848 llvm-svn: 320353	2017-12-11 11:53:54 +00:00
Diana Picus	775bb74379	[ARM GlobalISel] Add tests for PKHBT and PKHTB Test (some of) the patterns for selecting PKHBT and PKHTB. The others are just very similar to the ones we're testing and there would be little value in covering them as well. llvm-svn: 320352	2017-12-11 11:44:23 +00:00
Aleksandar Beserminji	d6dada17ff	[mips] Removal of microMIPS64R6 All files and parts of files related to microMIPS4R6 are removed. When target is microMIPS4R6, errors are printed. This is LLVM part of patch. Differential Revision: https://reviews.llvm.org/D35625 llvm-svn: 320350	2017-12-11 11:21:40 +00:00
Dylan McKay	2124bcf805	[AVR] Implement some missing code paths This has been broken since r320009. llvm-svn: 320348	2017-12-11 11:01:27 +00:00
Dylan McKay	ab6204b1e5	[AVR] Fix incorrectly-calculated AVRMCExpr evaluations This has been broken since r320009. llvm-svn: 320347	2017-12-11 11:01:19 +00:00
Craig Topper	ad45bf5895	[DAGCombiner] Support folding (mulhs/u X, 0)->0 for vectors. We should probably also fold (mulhs/u X, 1) for vectors, but that's harder. llvm-svn: 320344	2017-12-11 08:33:20 +00:00
Craig Topper	65ed4d4492	[DAGCombiner] Reuse existing SDLoc variable instead of creating a new one. NFC llvm-svn: 320343	2017-12-11 08:33:19 +00:00
Craig Topper	0bea09b737	[X86] Regenerate test with update_llc_test_checks.py llvm-svn: 320342	2017-12-11 06:16:26 +00:00
Craig Topper	1e83485613	[X86] Add a test case for masked scatter where the index needs to be legalized from v2i32 while other types are legal. llvm-svn: 320340	2017-12-11 01:48:10 +00:00
Simon Pilgrim	6b1f532ccf	[X86] Add ROL/ROR schedule tests llvm-svn: 320334	2017-12-10 22:11:56 +00:00
Simon Pilgrim	a6564e2358	[X86] Add DIV/MUL/NEG/NOP/NOT/PAUSE schedule tests llvm-svn: 320333	2017-12-10 21:56:24 +00:00
Simon Pilgrim	8e6d0fcbac	[X86] Add DEC/INC schedule tests Include i686 (non-REX) variant tests as well llvm-svn: 320332	2017-12-10 21:28:00 +00:00
Simon Pilgrim	f1c51d187a	[X86] Add INS/OUTS schedule tests llvm-svn: 320331	2017-12-10 21:10:28 +00:00
Simon Pilgrim	07ebbd53f0	[X86] Add CMPS/MOVS/SCAS/STOS schedule tests llvm-svn: 320330	2017-12-10 20:58:22 +00:00
Simon Pilgrim	f65831d731	[X86] Add CMOV schedule tests llvm-svn: 320329	2017-12-10 20:46:57 +00:00
Simon Pilgrim	4a431edddc	[X86] Add BT/BTC/BTR/BTS schedule tests llvm-svn: 320328	2017-12-10 20:22:47 +00:00
Craig Topper	c6a4a97260	[X86] Add VCOMISDZrr, VCOMISSZrr, VUCOMISDZrr, and VUCOMISSZrr to the skylake server sheduler model llvm-svn: 320326	2017-12-10 19:47:57 +00:00
Craig Topper	a0be5a06c1	[X86] Rename some instructions that start with Int_ to have the _Int at the end. This matches AVX512 version and is more consistent overall. And improves our scheduler models. In some cases this adds _Int to instructions that didn't have any Int_ before. It's a side effect of the adjustments made to some of the multiclasses. llvm-svn: 320325	2017-12-10 19:47:56 +00:00
Simon Pilgrim	c493d4f5b9	[X86][X87] Fix typo in znver1 FIST/FISTT schedule patterns llvm-svn: 320322	2017-12-10 19:19:22 +00:00
Simon Pilgrim	930e435937	[X86][X87] Add missing x87 scheduler tests Split off some 'n' instruction versions to make it clearer when WAIT is being inserted llvm-svn: 320321	2017-12-10 18:53:15 +00:00
Craig Topper	1de942b2d1	[X86] Rename some instructions from 'rb' to 'rrb' to make 'b' a proper suffix. Fix the scheduling information for some of them. Some of the scheduling information was only present for the 'rb' version' and not the 'rr' version. Now we match 'rr(b?)' llvm-svn: 320320	2017-12-10 17:42:44 +00:00
Craig Topper	c7445f2cdc	[X86] Add VCVTQQ2PS to the skylake server scheduler models. llvm-svn: 320319	2017-12-10 17:42:43 +00:00
Craig Topper	c268527b2f	[X86] Add VPMULLWZ256 to the skylake server scheduler model llvm-svn: 320318	2017-12-10 17:42:42 +00:00
Craig Topper	4ec397cbd3	[X86] Add 256/512-bit EVEX VPSADBW instructions to skylake server scheduler model. llvm-svn: 320317	2017-12-10 17:42:41 +00:00
Craig Topper	aa904d5ab6	[X86] Fix a few instructions that were named Z512 instead of just Z. This makes things consistent with our normal instruction naming. llvm-svn: 320316	2017-12-10 17:42:39 +00:00
Craig Topper	7c89de1760	[X86] Add VPSRLWZrr to skylake server scheduler model. llvm-svn: 320315	2017-12-10 17:42:38 +00:00
Craig Topper	1d7760db49	[X86] Add VPUNPCKLWDZrr to skylake server scheduler model. llvm-svn: 320314	2017-12-10 17:42:37 +00:00
Craig Topper	57c2815cbe	[X86] Adjust tablegen includes so we can use Instructions in scheduler models instead of just instregexs. This separates the CPU specific scheduler model includes to occur after the instructions. Moves the instruction includes between the basic scheduler information and the CPU specific scheduler models. llvm-svn: 320313	2017-12-10 17:42:36 +00:00
Sanjay Patel	b23e148114	[SimplifyLibCalls] propagate FMF when folding pow(x, -1.0) call Follow-up for a bug that's similar to: https://bugs.llvm.org/show_bug.cgi?id=35601 llvm-svn: 320312	2017-12-10 17:25:54 +00:00
Sanjay Patel	ac9cbd6c56	[InstCombine] add test for pow(x, -1.0) with FMF; NFC llvm-svn: 320311	2017-12-10 17:21:51 +00:00
Sanjay Patel	09ec34349a	[SimplifyLibCalls] propagate FMF when folding pow(x, 2.0) call (PR35601) This should fix the larger problem with sqrt shown in: https://bugs.llvm.org/show_bug.cgi?id=35601 llvm-svn: 320310	2017-12-10 16:52:26 +00:00
Sanjay Patel	719bc64ba5	[InstCombine] add test for pow(x, 2.0) with FMF; NFC llvm-svn: 320309	2017-12-10 16:43:34 +00:00
Simon Pilgrim	1f8cfba0bb	[X86] Flag BroadWell scheduler model as complete Locally tag COPY as WriteMove, which has caused some reg-reg + reg-mem instruction tests to reorder. llvm-svn: 320308	2017-12-10 13:49:51 +00:00
Simon Pilgrim	4ff43d8120	Regenerate some AVX2+ scheduling tests that got missed llvm-svn: 320307	2017-12-10 13:41:29 +00:00
Simon Pilgrim	49c74934dd	Strip trailing whitespace. NFCI. llvm-svn: 320306	2017-12-10 13:00:37 +00:00
Simon Pilgrim	af35b76bda	Regenerate some scheduling tests that got missed llvm-svn: 320305	2017-12-10 12:59:55 +00:00
Simon Pilgrim	320996576d	[X86] Flag ZNVER1 scheduler model as complete We just have to locally tag COPY as WriteMove llvm-svn: 320304	2017-12-10 12:43:53 +00:00
Simon Pilgrim	8547645948	[X86] Flag SLM scheduler model as complete We just have to locally tag COPY as WriteMove llvm-svn: 320303	2017-12-10 12:36:29 +00:00
Simon Pilgrim	91c159d841	[X86][AVX[ Tag VZEROALL/VZEROUPPER instructions scheduler classes llvm-svn: 320302	2017-12-10 12:26:35 +00:00
Simon Pilgrim	6de94a1adc	[X86] Tag SSE4A instructions as SSE INTALU scheduler classes llvm-svn: 320301	2017-12-10 12:08:04 +00:00
Simon Pilgrim	cd58171110	[X86] Flag BTVER2 scheduler model as complete We just have to locally tag COPY as WriteMove llvm-svn: 320300	2017-12-10 11:51:29 +00:00
Simon Pilgrim	b7fb2e2fa1	[X86] Tag ADJSTACK instructions as INTALU scheduler class llvm-svn: 320299	2017-12-10 11:34:08 +00:00
Dorit Nuzman	5809e70540	[SCEV] Fix wrong Equal predicate created in getAddRecForPhiWithCasts CreateAddRecFromPHIWithCastsImpl() adds an IncrementNUSW overflow predicate which allows the PSCEV rewriter to rewrite this scev expression: (zext i8 {0, + , (trunc i32 step to i8)} to i32) into {0, +, (sext i8 (trunc i32 step to i8) to i32)} But then it adds the wrong Equal predicate: %step == (zext i8 (trunc i32 %step to i8) to i32). instead of: %step == (sext i8 (trunc i32 %step to i8) to i32) This is fixed here. Differential Revision: https://reviews.llvm.org/D40641 llvm-svn: 320298	2017-12-10 11:13:35 +00:00
Simon Pilgrim	1a030016a6	[X86] Tag MORESTACK instructions as ret scheduler class llvm-svn: 320296	2017-12-10 10:08:21 +00:00
Craig Topper	253562eb81	[X86] Fix duplicate entries in skylake server scheduler model by changing Z128 to Z256 Based on the fact that the 'Y' version of the instruction is next to this, I assume Z256 is the intended value. llvm-svn: 320295	2017-12-10 09:14:45 +00:00
Craig Topper	90c9c15936	[X86] Add MOVQI2PQIrm, MOVSDmr, and MOVSDrm to scheduler information The VEX versions were present but not the legacy SSE versions. llvm-svn: 320294	2017-12-10 09:14:44 +00:00
Craig Topper	28e55386ac	[X86] Add LEA64_32r to scheduler models for Sandybridge,Haswell,Broadwell,Skylake llvm-svn: 320293	2017-12-10 09:14:42 +00:00
Craig Topper	8ade4640f3	[X86] Add IN16/OUT16 to scheduling information for Haswell,Broadwell,Skylake Sandy Bridge is also missing it, but it has other issues. See PR35590. llvm-svn: 320292	2017-12-10 09:14:41 +00:00
Craig Topper	1a88c50fd7	[X86] Fix scheduler models to support ADD32ri in addition to ADD32ri8. Similar for all sizes of AND/OR/XOR/SUB/ADC/SBB/CMP. llvm-svn: 320291	2017-12-10 09:14:39 +00:00
Craig Topper	c89e282f7d	[X86] Rename some instructions so that 'b' is added as a suffix instead of replacing an 'r' llvm-svn: 320290	2017-12-10 09:14:38 +00:00
Craig Topper	6c65910160	[X86] Add CMPSDrr/rm to the scheduler models. Somehow CMPSSrr/rm was there and the VEX version was there, but this was consistently missing. llvm-svn: 320289	2017-12-10 09:14:37 +00:00
Craig Topper	d435a1950f	[Docs] Fix typo in scheduler model documentation. enumemation->enumeration llvm-svn: 320288	2017-12-10 09:14:35 +00:00
Tim Northover	cf4701bb89	PowerPC: support external pid instructions in MC layer. This adds assembly & disassembly support for the e500mc "external pid" instructions. See https://reviews.llvm.org/D39249. Patch by vit9696 <vit9696@avp.su> llvm-svn: 320287	2017-12-10 08:43:19 +00:00
Xinliang David Li	fa3f1a15b2	[PGO] change arg type to uint64_t to match member field type llvm-svn: 320285	2017-12-10 07:39:53 +00:00
Craig Topper	da7e78e18c	[X86] Rename the rb form of scalar ADD/SUB/MUL/DIV to include _Int since they can only be selected by intrinsics. llvm-svn: 320283	2017-12-10 04:07:28 +00:00
Craig Topper	4e57776fb2	[X86] Correct the _Int part of more scheduler model instrexes. Put _b in the correct order relative to _Int llvm-svn: 320282	2017-12-10 03:16:38 +00:00
Craig Topper	a2f5528084	[X86] Remove ReadAfterLd from several several rb instructions This affects CVTSD2SS, FMA, RCP28, RSQRT28, and SQRT scalar instructions 'b' here refers to 'sae' not broadcast. These aren't memory instructions. llvm-svn: 320281	2017-12-10 03:16:36 +00:00
Craig Topper	29868dcbaa	[X86] Fix test case I failed ot update in r320279. llvm-svn: 320280	2017-12-10 01:27:54 +00:00
Craig Topper	391c6f9507	[X86] Fix bad regular expressions in the scheduler models. Question marks should be outside of multicharacter parenthesized expressions If the question mark is inside the parentheses it only applies to the single character proceeding it. I had to make a few additional cleanups to fix some duplicate warnings that were exposed by fixing this. llvm-svn: 320279	2017-12-10 01:24:08 +00:00
Craig Topper	8ee98d0b51	[X86] Make the _Int part of some instregex sheduler patterns optional llvm-svn: 320278	2017-12-10 01:24:06 +00:00
Craig Topper	5ffe80103e	[X86] Add the commutable floating point min/max pseudo instructions to sandybridge,haswell,broadwell,skylakeclient scheduler models. llvm-svn: 320277	2017-12-10 01:24:05 +00:00
Simon Pilgrim	6655eef1b4	[X86] Tag PIC setup instruction as jump scheduler class llvm-svn: 320276	2017-12-10 00:40:37 +00:00
Simon Pilgrim	5d74949e5f	[X86] Tag ACQUIRE/RELEASE atomic instructions as microcoded scheduler classes Note: We may be too pessimistic here and should possibly use something closer to the LOCK arithmetic instructions llvm-svn: 320275	2017-12-10 00:30:57 +00:00
Simon Pilgrim	dcbe723d28	[X86] Tag TLS instructions as system scheduler classes llvm-svn: 320274	2017-12-10 00:12:57 +00:00
Simon Pilgrim	3508a09455	[X86] Tag ALLOCA/VAARG instructions as system scheduler classes llvm-svn: 320273	2017-12-10 00:03:16 +00:00
Joel Jones	5cc21e83ce	[AArch64] Improve loop unrolling performance on Cavium T99 This patch improves performance on Cavium T99 as shown here (libquantum 0.2.4): https://docs.google.com/spreadsheets/d/1Lo1o2E1NjrpkwS7DvYYWsiVvPdd93h7KBaqeptMrZPY/edit?usp=sharing By increasing the LoopMicroOpsBufferSize in the Cavium T99 Scheduler file, loop unrolling becomes more aggressive. This helps performance on T99. Test case included. Patch by Stefan Teleman Differential Revision: https://reviews.llvm.org/D40695 llvm-svn: 320272	2017-12-09 23:59:55 +00:00
Simon Pilgrim	a42a54258e	[InstCombine] Fix SimplifyDemandedUseBits SHL handling (PR35515) Don't assume that the pattern matched SRL can be cast to an Instruction (might be ConstExpr etc.) llvm-svn: 320270	2017-12-09 23:42:56 +00:00
Simon Dardis	70dbd5fbd0	Infer lowest bits of an integer Multiply when the low bits of the operands are known When the lowest bits of the operands to an integer multiply are known, the low bits of the result are deducible. Code to deduce known-zero bottom bits already existed, but this change improves on that by deducing known-ones. Patch by: Pedro Ferreira Reviewers: craig.topper, sanjoy, efriedma Differential Revision: https://reviews.llvm.org/D34029 llvm-svn: 320269	2017-12-09 23:25:57 +00:00
Craig Topper	f4e3044db9	[X86] Use KMOV instructions to zero upper bits of vectors when possible. llvm-svn: 320268	2017-12-09 23:10:59 +00:00
Craig Topper	5ac75d5628	[X86] Improve lowering of vXi1 insert_subvectors to better utilize (insert_subvector zero, vec, 0) for zeroing upper bits. This can be better recognized during isel when the producer already zeroed the upper bits. llvm-svn: 320267	2017-12-09 22:44:42 +00:00
Simon Pilgrim	e049038692	[X86] Tag LOCK/REX64/DATA16/DATA32 instruction prefix scheduler classes llvm-svn: 320266	2017-12-09 21:27:03 +00:00
Simon Pilgrim	b2b93f6204	Strip trailing whitespace. NFCI. llvm-svn: 320265	2017-12-09 20:44:51 +00:00
Simon Pilgrim	7e636cc419	[X86] Tag FS/GS BASE R/W instruction scheduler classes llvm-svn: 320264	2017-12-09 20:42:27 +00:00
Simon Pilgrim	231fab072f	[X86] Tag REP/REPNE prefix instructions as microcoded scheduler classes llvm-svn: 320263	2017-12-09 20:16:37 +00:00
Simon Pilgrim	2e7314eb2f	[X86] Tag missing EH pseudo instruction scheduler classes llvm-svn: 320262	2017-12-09 20:04:02 +00:00
Simon Pilgrim	cb71e72707	[X86] Tag frame pointer XORs instruction scheduler classes llvm-svn: 320261	2017-12-09 19:56:39 +00:00
Craig Topper	504534514c	[X86] Don't use getTargetConstant for all 0s and all 1s mask vector. llvm-svn: 320260	2017-12-09 19:18:30 +00:00
Adrian Prantl	844f8f21fd	Remove duplicate option from documentation. llvm-svn: 320258	2017-12-09 19:09:59 +00:00
Simon Pilgrim	df702104d3	[X86] Tag segment prefixes as NOP instruction scheduling classes llvm-svn: 320257	2017-12-09 16:58:34 +00:00
Simon Pilgrim	d3e21c6b79	[X86][AVX512] Drop a default NoItinerary argument that isn't used any more. NFCI. Requires re-ordering of AVX512_maskable_custom arguments. llvm-svn: 320255	2017-12-09 16:20:54 +00:00
Simon Pilgrim	a335e1e29d	Fix 'enumeral and non-enumeral type in conditional expression' gcc warning. NFCI. llvm-svn: 320254	2017-12-09 16:19:18 +00:00
Simon Pilgrim	3d0be4f507	Fix signed/unsigned gcc warning. NFCI. llvm-svn: 320253	2017-12-09 16:04:57 +00:00
Florian Hahn	c5bebffe4f	[InlineFunction] Set debug loc for call to forward varargs. Reviewers: aprantl, dblaikie, rnk Reviewed By: rnk Subscribers: eraman, llvm-commits, JDevlieghere Differential Revision: https://reviews.llvm.org/D40432 llvm-svn: 320252	2017-12-09 14:25:33 +00:00
Craig Topper	6504a8f888	[X86] When inserting into the upper bits of a vXi1 vector, make sure we shift enough bits if we widened the vector. We may need to widen the vector to make the shifts legal, but if we do that we need to make sure we shift left/right after accounting for the new size. If not we can't guarantee we are shifting in zeros. The test cases affected actually show cases where we should move the shifts all together, but that's another problem. llvm-svn: 320248	2017-12-09 08:19:07 +00:00
Dylan McKay	ba23343a45	Revert and accidentally committed revert commit This reverts commit r320245. llvm-svn: 320247	2017-12-09 08:01:28 +00:00
Dylan McKay	f7e8ec1348	[AVR] Fix two CodeGen tests These were broken because of various printing format changes. llvm-svn: 320246	2017-12-09 07:51:43 +00:00
Dylan McKay	f5422afdf0	Revert "[AVR] Override ParseDirective" This reverts commit 57c16f9267969ebb09d6448607999b4a9f40c418. llvm-svn: 320245	2017-12-09 07:51:37 +00:00
Craig Topper	b3e14ce90c	[X86] Improve lowering of concats of mask vectors to better optimize zero vector inputs. We were previously using kunpck with zero inputs unnecessarily. And we had cases where we would insert into a zero vector and then insert into larger zero vector incurring two sets of shifts. llvm-svn: 320244	2017-12-09 07:02:19 +00:00
Dylan McKay	80463fe64d	Relax unaligned access assertion when type is byte aligned Summary: This relaxes an assertion inside SelectionDAGBuilder which is overly restrictive on targets which have no concept of alignment (such as AVR). In these architectures, all types are aligned to 8-bits. After this, LLVM will only assert that accesses are aligned on targets which actually require alignment. This patch follows from a discussion on llvm-dev a few months ago http://llvm.1065342.n5.nabble.com/llvm-dev-Unaligned-atomic-load-store-td112815.html Reviewers: bogner, nemanjai, joerg, efriedma Reviewed By: efriedma Subscribers: efriedma, cactus, llvm-commits Differential Revision: https://reviews.llvm.org/D39946 llvm-svn: 320243	2017-12-09 06:45:36 +00:00
Jessica Paquette	a249c4f513	[MachineOutliner] Outline calls The outliner previously would never outline calls. Calls are pretty common in files, so it makes sense to outline them. In fact, in the LLVM test suite, if you count the number of instructions that the outliner misses when you outline calls vs when you don't, it turns out that, on average, around 6% of the instructions encountered are calls. So, if we outline calls, we can find more candidates, and thus save some more space. This commit adds that functionality and updates the mir test to reflect that. llvm-svn: 320229	2017-12-09 00:43:49 +00:00
Wolfgang Pieb	8b1a175be6	[NFC] Change the string offsets table tests to generate the object on the fly which enables us to remove the test scripts and object files from the repository. https://reviews.llvm.org/D40914 llvm-svn: 320227	2017-12-09 00:39:53 +00:00
Kamil Rytarowski	3d3f91e832	Register NetBSD/x86_64 in MemorySanitizer.cpp Summary: Reuse the Linux new mapping as it is. Sponsored by <The NetBSD Foundation> Reviewers: joerg, eugenis, vitalybuka Reviewed By: vitalybuka Subscribers: llvm-commits, #sanitizers Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D41022 llvm-svn: 320219	2017-12-09 00:32:09 +00:00
Evgeniy Stepanov	c667c1f47a	Hardware-assisted AddressSanitizer (llvm part). Summary: This is LLVM instrumentation for the new HWASan tool. It is basically a stripped down copy of ASan at this point, w/o stack or global support. Instrumenation adds a global constructor + runtime callbacks for every load and store. HWASan comes with its own IR attribute. A brief design document can be found in clang/docs/HardwareAssistedAddressSanitizerDesign.rst (submitted earlier). Reviewers: kcc, pcc, alekseyshl Subscribers: srhines, mehdi_amini, mgorny, javed.absar, eraman, llvm-commits, hiraditya Differential Revision: https://reviews.llvm.org/D40932 llvm-svn: 320217	2017-12-09 00:21:41 +00:00
Paul Robinson	8bd9d6ad83	Fix out-of-order stepping behavior in programs with sunk instructions. MachineSink attempts to place instructions near the basic blocks where they are needed. Once an instruction has been sunk, its location relative to other instructions no longer is consistent with the original source code. In order to ensure correct stepping in the debugger, the debug location for sunk instructions is either merged with the insertion point or erased if the target successor block is empty. Originally submitted as r318679, revised to fix sanitizer failure and improve testing. Patch by Matthew Voss! Differential Revision: https://reviews.llvm.org/D39933 llvm-svn: 320216	2017-12-09 00:17:01 +00:00
Adrian Prantl	01fb31cc89	dwarfdump: Add support for the --diff option. --diff Emit the output in a diff-friendly way by omitting offsets and addresses. <rdar://problem/34502625> llvm-svn: 320214	2017-12-08 23:32:47 +00:00
Craig Topper	e29f50da4d	[X86][Mips] Remove unused method declaration from the X86 and Mips AsmPrinters. Both had a declaration of EmitXRayTable, but there is no method defined in either with that name. There is a emitXRayTable in the base class with a lower case 'e' and they both call that. llvm-svn: 320213	2017-12-08 23:30:03 +00:00
Francis Visoiu Mistrih	440f69c95a	[CodeGen] Move printing MO_Immediate operands to MachineOperand::print Work towards the unification of MIR and debug output by refactoring the interfaces. Add support for operand subreg index as an immediate to debug printing and use ::print in the MIRPrinter. Differential Review: https://reviews.llvm.org/D40965 llvm-svn: 320209	2017-12-08 22:53:21 +00:00
Duncan P. N. Exon Smith	9b8caf5bd7	Revert part of "Cleanup some GraphTraits iteration code" This reverts part of r300656, which caused a regression in propagateMassToSuccessors by counting edges n^2 times, where n is the number of edges from the source basic block to the same successor basic block. The result was both incorrect and very slow to compute for large values of n (e.g. switches with multiple cases that go to the same basic block). Patch by Andrew Scheidecker! llvm-svn: 320208	2017-12-08 22:42:43 +00:00
Richard Smith	8a3adc3abb	Avoid constructing an out-of-range value for an enumeration (which results in UB). llvm-svn: 320206	2017-12-08 22:32:35 +00:00
Abderrazek Zaafrani	5a2583f026	[AArch64] Rename AArch64VecorByElementOpt.cpp into AArch64SIMDInstrOpt.cpp to reflect the recently added features. The name change is dicsussed in https://reviews.llvm.org/D38196 llvm-svn: 320204	2017-12-08 22:04:13 +00:00
Adrian Prantl	d13170174c	Generalize llvm::replaceDbgDeclare and actually support the use-case that is mentioned in the documentation (inserting a deref before the plus_uconst). llvm-svn: 320203	2017-12-08 21:58:18 +00:00
Vedant Kumar	195dfd10a6	[Debugify] Add a pass to test debug info preservation The Debugify pass synthesizes debug info for IR. It's paired with a CheckDebugify pass which determines how much of the original debug info is preserved. These passes make it easier to create targeted tests for debug info preservation. Here is the Debugify algorithm: NextLine = 1 for (Instruction &I : M) attach DebugLoc(NextLine++) to I NextVar = 1 for (Instruction &I : M) if (canAttachDebugValue(I)) attach dbg.value(NextVar++) to I The CheckDebugify pass expects contiguous ranges of DILocations and DILocalVariables. If it fails to find all of the expected debug info, it prints a specific error to stderr which can be FileChecked. This was discussed on llvm-dev in the thread: "Passes to add/validate synthetic debug info" Differential Revision: https://reviews.llvm.org/D40512 llvm-svn: 320202	2017-12-08 21:57:28 +00:00
Florian Hahn	e5089e2e94	[CodeExtractor] Add debug locations for new call and branch instrs. Summary: If a partially inlined function has debug info, we have to add debug locations to the call instruction calling the outlined function. We use the debug location of the first instruction in the outlined function, as the introduced call transfers control to this statement and there is no other equivalent line in the source code. We also use the same debug location for the branch instruction added to jump from artificial entry block for the outlined function, which just jumps to the first actual basic block of the outlined function. Reviewers: davide, aprantl, rriddle, dblaikie, danielcdh, wmi Reviewed By: aprantl, rriddle, danielcdh Subscribers: eraman, JDevlieghere, llvm-commits Differential Revision: https://reviews.llvm.org/D40413 llvm-svn: 320199	2017-12-08 21:49:03 +00:00
Dan Gohman	3a762bf9df	[WebAssembly] Reapply r319186: "Support bitcasted function addresses with varargs." This puts the functionality under control of a command-line option which is off by default to avoid breaking existing setups. llvm-svn: 320197	2017-12-08 21:27:00 +00:00
Dan Gohman	6736f59078	[WebAssemby] Re-apply r320041: "Support main functions with alternate signatures." This includes a fix so that it doesn't transform declarations, and it puts the functionality under control of a command-line option which is off by default to avoid breaking existing setups. llvm-svn: 320196	2017-12-08 21:18:21 +00:00
Evandro Menezes	5d7a9e6e54	[AArch64] Add Exynos to host detection Differential revision: https://reviews.llvm.org/D40985 llvm-svn: 320195	2017-12-08 21:09:59 +00:00
Konstantin Zhuravlyov	c40d9f2e5d	AMDGPU/GCN: Bring processors in sync with AMDGPUUsage - Add gfx704 - Change bonaire to gfx704 - Remove gfx804 - Remove gfx901 - Remove gfx903 Differential Revision: https://reviews.llvm.org/D40046 llvm-svn: 320194	2017-12-08 20:52:28 +00:00
Simon Pilgrim	5f7fcb2ea9	[X86] CMOV pseudo instructions shouldn't need scheduling info as they should be lowered early llvm-svn: 320193	2017-12-08 20:42:35 +00:00
Simon Pilgrim	f621dcf8d7	[X86][X87] Tag x87 load/store instructions scheduler classes llvm-svn: 320192	2017-12-08 20:31:48 +00:00
Craig Topper	7f0d456ef8	[X86] Teach lowering to only let through (insert_subvector (vXi1 zeros), subvec, 0) for vector sizes that have native KSHIFT support. For narrow sizes we'll widen the zero vector and widen the insert. Then do an extract_subvector to get back down to correct size. This allows us to remove some patterns from the isel table that had to COPY_TO_REGCLASS to an oversized register, do the shift and then COPY_TO_REGCLASS back to the narrow register. Now this is represented explicitly in the DAG. This seems to have perturbed the register allocation in one of the tests, but the number of instructions didn't change. llvm-svn: 320190	2017-12-08 20:10:33 +00:00
Simon Pilgrim	6415f56c79	[X86][X87] Tag x87 float compare instructions scheduler classes llvm-svn: 320189	2017-12-08 20:10:31 +00:00
Matt Arsenault	73ce93b08b	AMDGPU: Set IntrReadMem on memtime intrinsics llvm-svn: 320188	2017-12-08 20:01:02 +00:00
Matt Arsenault	856777d8c9	AMDGPU: image_getlod and image_getresinfo do not read memory llvm-svn: 320187	2017-12-08 20:00:57 +00:00
Matt Arsenault	ecad0d5364	AMDGPU: Preserve MMO in adjustWritemask Follow up to r319705. Currently the MMO is produced after this in the custom inserter, so this doesn't change anything yet. llvm-svn: 320186	2017-12-08 20:00:45 +00:00
Shoaib Meenai	d9073510b7	[llvm] Add install-distribution-stripped This is identical to the install-distribution target, except that it strips the installed binaries. Differential Revision: https://reviews.llvm.org/D40689 llvm-svn: 320184	2017-12-08 19:44:45 +00:00
Shoaib Meenai	038fd0056a	[cmake] Only pass CMAKE_SYSROOT if non-empty In my build environment (cmake 3.6.1 and gcc 4.8.5 on CentOS 7), having an empty CMAKE_SYSROOT in the cache results in --sysroot="" being passed to all compile commands, and then the compiler errors out because of the empty sysroot. Only set CMAKE_SYSROOT if non-empty to avoid this. Differential Revision: https://reviews.llvm.org/D40934 llvm-svn: 320183	2017-12-08 19:42:47 +00:00
Shoaib Meenai	e8828d49d0	[runtimes] Add install--stripped targets These should be the only remaining missing install--stripped targets. They're modeled after the existing install targets. Differential Revision: https://reviews.llvm.org/D40927 llvm-svn: 320182	2017-12-08 19:42:46 +00:00
Xinliang David Li	d91057bf52	Revert r320104: infinite loop profiling bug fix Causes unexpected memory issue with New PM this time. The new PM invalidates BPI but not BFI, leaving the reference to BPI from BFI invalid. Abandon this patch. There is a more general solution which also handles runtime infinite loop (but not statically). llvm-svn: 320180	2017-12-08 19:38:07 +00:00
Brian M. Rzycki	0eae123d9e	[JumpThreading] Minor comment cleanup. NFC. (test commit) llvm-svn: 320179	2017-12-08 19:36:32 +00:00
Simon Pilgrim	2db2851378	[X86][MPX] Tag TSX/HLE/SGX instructions scheduler classes Currently tagged these as system instructions. llvm-svn: 320177	2017-12-08 19:26:22 +00:00
Konstantin Zhuravlyov	e30f88f3a9	AMDGPU: Report Arg's Value name in metadata if kernel_arg_name metadata is not available Differential Revision: https://reviews.llvm.org/D40924 llvm-svn: 320176	2017-12-08 19:22:12 +00:00
Michael Trent	ad840d2206	Reverting r320166 to fix test failures. llvm-svn: 320174	2017-12-08 19:09:26 +00:00
Simon Pilgrim	42fcda9a6c	[X86][MPX] Tag MPX instructions scheduler classes Currently tagged these as system instructions, once we have uses for them (ASAN?) and they are faster we will need to improve on this. llvm-svn: 320173	2017-12-08 19:03:42 +00:00
Sanjay Patel	d4468912b0	[x86] use hasAVX2() rather than hasInt256(); NFC These are aliases, but the thing we're checking here is that the target has vpsllv*, not that the data type is 256-bit. Those instructions exist for 128-bit vectors too...but sadly, not for all element sizes. llvm-svn: 320170	2017-12-08 18:35:51 +00:00
Simon Pilgrim	8e39dc36b8	[X86] Tag move immediate instructions scheduler classes llvm-svn: 320169	2017-12-08 18:35:40 +00:00
Michael Trent	de5209bdbd	Updated llvm-objdump to display local relocations in Mach-O binaries Summary: llvm-objdump's Mach-O parser was updated in r306037 to display external relocations for MH_KEXT_BUNDLE file types. This change extends the Macho-O parser to display local relocations for MH_PRELOAD files. When used with the -macho option relocations will be displayed in a historical format. rdar://35778019 Reviewers: enderby Reviewed By: enderby Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D40867 llvm-svn: 320166	2017-12-08 17:51:04 +00:00
Davide Italiano	b5a62cc81a	[DebugInfo] Use llc instead of llc_dwarf to fix this test. We work around the fact that some platforms add a triple when they expand llc_dwarf in lit. llvm-svn: 320164	2017-12-08 17:15:50 +00:00
Simon Pilgrim	19d460b066	[X86][SHA] Tag SHA instructions scheduler classes Put these under VecIMul itinerary classes for now - seems to be a good average value llvm-svn: 320161	2017-12-08 16:38:41 +00:00
Simon Pilgrim	4ba3314d55	[X86] Tag VIA PadLock crypto instructions scheduler classes llvm-svn: 320159	2017-12-08 16:06:40 +00:00
Simon Pilgrim	1ddcae665e	[X86] Tag PKU/INVPCID/RDPID/SMAP/SMX/PTWRITE system instructions scheduler classes llvm-svn: 320158	2017-12-08 15:48:37 +00:00
Alexey Bataev	ec95c6cc0a	[InstCombine] PR35354: Convert store(bitcast, load bitcast (select (Cond, &V1, &V2)) --> store (, load (select(Cond, load &V1, load &V2))) Summary: If we have the code like this: ``` float a, b; a = std::max(a ,b); ``` it is converted into something like this: ``` %call = call dereferenceable(4) float* @_ZSt3maxIfERKT_S2_S2_(float* nonnull dereferenceable(4) %a.addr, float* nonnull dereferenceable(4) %b.addr) %1 = bitcast float* %call to i32* %2 = load i32, i32* %1, align 4 %3 = bitcast float* %a.addr to i32* store i32 %2, i32* %3, align 4 ``` After inlinning this code is converted to the next: ``` %1 = load float, float* %a.addr %2 = load float, float* %b.addr %cmp.i = fcmp fast olt float %1, %2 %__b.__a.i = select i1 %cmp.i, float* %a.addr, float* %b.addr %3 = bitcast float* %__b.__a.i to i32* %4 = load i32, i32* %3, align 4 %5 = bitcast float* %arrayidx to i32* store i32 %4, i32* %5, align 4 ``` This pattern is not recognized as minmax pattern. Patch solves this problem by converting sequence ``` store (bitcast, (load bitcast (select ((cmp V1, V2), &V1, &V2)))) ``` to a sequence ``` store (,load (select((cmp V1, V2), &V1, &V2))) ``` After this the code is recognized as minmax pattern. Reviewers: RKSimon, spatel Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D40304 llvm-svn: 320157	2017-12-08 15:32:10 +00:00
Simon Pilgrim	83708cabc0	[X86][AVX512] Tag CLWB instruction to CLFLUSH/PREFETCH scheduler class llvm-svn: 320156	2017-12-08 15:19:10 +00:00
Alexey Bataev	ad1d023d94	[PatternMatch] Add matcher for LoadInst, NFC. llvm-svn: 320155	2017-12-08 15:17:37 +00:00
Simon Pilgrim	26f106fda4	[X86][AVX512] Tag AVX512_512_SEXT_MASK_* instructions scheduler classes Match VPTERNLOG which these pseudos will eventually alias to llvm-svn: 320154	2017-12-08 15:17:32 +00:00
Tim Renouf	cead41d42f	[AMDGPU] add labels to +DumpCode output Summary: +DumpCode is a hack to embed disassembly in the ELF file. This commit fixes it to include labels, to make it slightly more useful. Reviewers: arsenm, kzhuravl Subscribers: nhaehnle, timcorringham, dstuttard, llvm-commits, t-tye, yaxunl, wdng, kzhuravl Differential Revision: https://reviews.llvm.org/D40169 llvm-svn: 320146	2017-12-08 14:09:34 +00:00
Max Kazantsev	63a3de057e	[NFC] Rename variable from Cond to Pred to make it more sound llvm-svn: 320144	2017-12-08 12:54:32 +00:00
Max Kazantsev	9c08b7a053	[SCEV] Fix predicate usage in computeExitLimitFromICmp In this method, we invoke `SimplifyICmpOperands` which takes the `Cond` predicate by reference and may change it along with `LHS` and `RHS` SCEVs. But then we invoke `computeShiftCompareExitLimit` with Values from which the SCEVs have been derived, these Values have not been modified while `Cond` could be. One of possible outcomes of this is that we may falsely prove that an infinite loop ends within some finite number of iterations. In this patch, we save the original `Cond` and pass it along with original operands. This logic may be removed in future once `computeShiftCompareExitLimit` works with SCEVs instead of value operands. Reviewed By: sanjoy Differential Revision: https://reviews.llvm.org/D40953 llvm-svn: 320142	2017-12-08 12:19:45 +00:00
Francis Visoiu Mistrih	f4bd295576	[CodeGen] Move printing MO_MachineBasicBlock operands to MachineOperand::print Work towards the unification of MIR and debug output by refactoring the interfaces. llvm-svn: 320141	2017-12-08 11:48:02 +00:00
Francis Visoiu Mistrih	6c4ca713f1	[CodeGen] Move printing MO_CImmediate operands to MachineOperand::print Work towards the unification of MIR and debug output by refactoring the interfaces. llvm-svn: 320140	2017-12-08 11:40:06 +00:00
Pavel Labath	f5f0fffea5	[cmake] Make setting of CMAKE_C(XX)_COMPILER flags overridable for cross-builds Summary: r319898 made it possible to override these variables via the CROSS_TOOLCHAIN_FLAGS setting, but this only worked if one explicitly specifies these variables there. If, instead, one uses CROSS_TOOLCHAIN_FLAGS to specify a toolchain file (as our internal builds do, to point cmake to a checked-in toolchain), the CMAKE_C(XX)_COMPILER flags would still win over the ones specified by the toolchain file. To fix is to make the mere presence of these flags overridable. I do this by putting them as a default value for the CROSS_TOOLCHAIN_FLAGS setting, so they can be overridden at cmake configuration time. Reviewers: hintonda, beanz Subscribers: bogner, llvm-commits, mgorny Differential Revision: https://reviews.llvm.org/D40947 llvm-svn: 320138	2017-12-08 09:59:48 +00:00
Gadi Haber	2cf601f28f	[X86][Haswell]: Updating the scheduling information for the Haswell subtarget. Updated the scheduling information for the Haswell subtarget with the following changes: Regrouped the instructions after adding appropriate load + store latencies. Added scheduling for missing instructions such as the GATHER instrs. The changes were made after revisiting the latencies impact of all memory uOps. Reviewers: RKSimon, zvi, craig.topper, apilipenko Differential Revision: https://reviews.llvm.org/D40021 Change-Id: Iaf6c1f5169add1552845a8a566af4e5a359217a7 llvm-svn: 320137	2017-12-08 09:48:44 +00:00
Igor Laevsky	76b36d3a7f	[FuzzMutate] Correctly insert sinks and sources around invoke instructions Differential Revision: https://reviews.llvm.org/D40840 llvm-svn: 320136	2017-12-08 08:53:16 +00:00
Craig Topper	037115c29f	[X86] Always consider inserting a vXi1 vector into the lsbs of a zero vector to be legal during lowering. Add isel patterns to emit shifts. Previously we only allowed these through if the subvector came from a compare or test instruction which we would again check for during isel. With this change we only check for the compare and test instructions during isel and have fallback patterns that emit the shifts if needed. I noticed that in a lot of cases we don't actually see the compare during lowering and rely on an odd legalization of concat_vectors with a zero vector as the second argument. This keeps the concat_vectors around long enough for a later dag combine to expose the compare then we re-legalize the concat_vectors and catch the compare. llvm-svn: 320134	2017-12-08 08:10:58 +00:00
Abderrazek Zaafrani	2c80e4c7c3	[AArch64] Avoid SIMD interleaved store instruction for Exynos. Replace interleaved store instructions by equivalent and more efficient instructions based on latency cost model. Https://reviews.llvm.org/D38196 llvm-svn: 320123	2017-12-08 00:58:49 +00:00
Derek Schuff	9e1baeda74	Revert "[WebAssemby] Support main functions with alternate signatures." This reverts commit 959e37e669b0c3cfad4cb9f1f7c9261ce9f5e9ae. That commit doesn't handle the case where main is declared rather than defined, in particular the even-more special case where main is a prototypeless declaration (which is of course the one actually used by musl currently). llvm-svn: 320121	2017-12-08 00:39:54 +00:00
Craig Topper	323ba39f10	[X86] Handle alls version of vXi1 insert_vector_elt with a constant index without falling back to shuffles. We previously only supported inserting to the LSB or MSB where it was easy to zero to perform an OR to insert. This change effectively extracts the old value and the new value, xors them together and then xors that single bit with the correct location in the original vector. This will cancel out the old value in the first xor leaving the new value in the position. The way I've implemented this uses 3 shifts and two xors and uses an additional register. We can avoid the additional register at the cost of another shift. llvm-svn: 320120	2017-12-08 00:16:09 +00:00
Craig Topper	fd86b3cf22	[X86] Fix indentation. NFC llvm-svn: 320119	2017-12-08 00:15:57 +00:00
Lang Hames	2f0c5bbc4b	[ORC] Mark SymbolStringPool methods as inline to avoid linkage errors, add a less-than comparison to SymbolStringPtr and a corresponding unit test. llvm-svn: 320116	2017-12-07 23:32:11 +00:00
Don Hinton	25e64a1b15	[dump] Make LLVM_ENABLE_DUMP independent, and move to llvm-config.h Summary: Make LLVM_ENABLE_DUMP independent LLVM_ENABLE_ASSERTIONS, move it to llvm-config.h, and update description. Differential Revision: https://reviews.llvm.org/D38406 llvm-svn: 320111	2017-12-07 22:55:40 +00:00
Bill Seurer	957a076cce	[PowerPC][asan] Update asan to handle changed memory layouts in newer kernels In more recent Linux kernels with 47 bit VMAs the layout of virtual memory for powerpc64 changed causing the address sanitizer to not work properly. This patch adds support for 47 bit VMA kernels for powerpc64 and fixes up test cases. https://reviews.llvm.org/D40907 There is an associated patch for compiler-rt. Tested on several 4.x and 3.x kernel releases. llvm-svn: 320109	2017-12-07 22:53:33 +00:00
Zachary Turner	ecd2684ed7	[DebugInfo] Fix register variables not showing up in pdb. Previously, when linking against libcmt from the MSVC runtime, lld-link /verbose would show "Ignoring unknown symbol record with kind 0x1006". It turns out this was because TypeIndexDiscovery did not handle S_REGISTER records, so these records were not getting properly remapped. Patch by: Alexnadre Ganea Differential Revision: https://reviews.llvm.org/D40919 llvm-svn: 320108	2017-12-07 22:51:16 +00:00
Alina Sbirlea	193429f0c8	[ModRefInfo] Make enum ModRefInfo an enum class [NFC]. Summary: Make enum ModRefInfo an enum class. Changes to ModRefInfo values should be done using inline wrappers. This should prevent future bit-wise opearations from being added, which can be more error-prone. Reviewers: sanjoy, dberlin, hfinkel, george.burgess.iv Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D40933 llvm-svn: 320107	2017-12-07 22:41:34 +00:00
Eric Christopher	a469acac03	Temporarily revert "[PowerPC] Allow tail calls of fastcc functions from C CallingConv functions." It is causing sanitizer failures on llvm tests in a bootstrapped compiler. No bot link since it's currently down, but following up to get the bot up. This reverts commit r319218. llvm-svn: 320106	2017-12-07 22:26:19 +00:00
Xinliang David Li	4b0027f671	[PGO] detect infinite loop and form MST properly Differential Revision: http://reviews.llvm.org/D40873 llvm-svn: 320104	2017-12-07 22:23:28 +00:00
Jessica Paquette	59948666fb	[MachineOutliner] Fix offset overflow check The offset overflow check before was incorrect. It would always give the correct result, but it was comparing the SCALED potential fixed-up offset against an UNSCALED minimum/maximum. As a result, the outliner was missing a bunch of frame setup/destroy instructions that ought to have been safe to outline. This fixes that, and adds an instruction to the .mir test that failed the old test. llvm-svn: 320090	2017-12-07 21:51:43 +00:00
Mark Searles	095d4ea4bf	[AMDGPU] Fix typo in Kernel Descriptor for GFX6-GFX9 Differential Revision: https://reviews.llvm.org/D40981 llvm-svn: 320087	2017-12-07 21:24:27 +00:00
Mark Searles	9ebdbb433a	[AMDGPU] Revert "[AMDGPU] Add options for waitcnt pass debugging; add instr count in debug output." Patch caused a buildbot failure; http://lab.llvm.org:8011/builders/lld-x86_64-darwin13/builds/15733/steps/build_Lld/logs/stdio : lib/Target/AMDGPU/SIInsertWaitcnts.cpp:396:11: error: private field 'InstCnt' is not used [-Werror,-Wunused-private-field] int32_t InstCnt = 0; ^ 1 error generated. " This reverts commit 71627f79010aafe74fdcba901bba28dd7caa0869. llvm-svn: 320086	2017-12-07 21:14:41 +00:00
Mark Searles	a84d23489a	[AMDGPU] Add options for waitcnt pass debugging; add instr count in debug output. -amdgpu-waitcnt-forcezero={1\|0} Force all waitcnt instrs to be emitted as s_waitcnt vmcnt(0) expcnt(0) lgkmcnt(0) -amdgpu-waitcnt-forceexp=<n> Force emit a s_waitcnt expcnt(0) before the first <n> instrs -amdgpu-waitcnt-forcelgkm=<n> Force emit a s_waitcnt lgkmcnt(0) before the first <n> instrs -amdgpu-waitcnt-forcevm=<n> Force emit a s_waitcnt vmcnt(0) before the first <n> instrs Differential Revision: https://reviews.llvm.org/D40091 llvm-svn: 320084	2017-12-07 20:36:39 +00:00
Mark Searles	d29f24acfb	[AMDGPU] Add GCNHazardRecognizer::checkInlineAsmHazards() and GCNHazardRecognizer::checkVALUHazardsHelper(). checkInlineAsmHazards() checks INLINEASM for hazards that we particularly care about (so not exhaustive); this patch adds a check for INLINEASM that defs vregs that hold data-to-be stored by immediately preceding store of more than 8 bytes. If the instr were not within an INLINEASM, this scenario would be handled by checkVALUHazard(). Add checkVALUHazardsHelper(), which will be called by both checkVALUHazards() and checkInlineAsmHazards(). Differential Revision: https://reviews.llvm.org/D40098 llvm-svn: 320083	2017-12-07 20:34:25 +00:00
Craig Topper	dfc79c7c33	[X86] Fix InsertBitToMaskVector to only issue KSHIFTS of native size so that upper bits are properly zeroed. There's no v2i1 or v4i1 kshift, and v8i1 is only supported with AVXDQ. Isel has fake patterns to extend these types to native shifts, but makes no guarantees about the value of any bits shifted in when shifting right. This patch promotes the vector to a type that supports a native shift first and only allows inserting into the msb of a native sized shift. I've constructed this in a way that doesn't do the promotion if we're going to fallback to using a xmm/ymm/zmm shuffle. I think I have a plan to remove the shuffle fall back entirely. In which case we this can be simplified, but I wanted to fix the correctness issue first. llvm-svn: 320081	2017-12-07 20:10:04 +00:00
Craig Topper	7b8fa5f782	[X86] Fix typo in variable name. NFC llvm-svn: 320080	2017-12-07 20:10:01 +00:00
Craig Topper	b67e5da89b	[X86] Make a couple helper lowering methods static. llvm-svn: 320079	2017-12-07 20:09:55 +00:00
Sanjay Patel	6cfc136870	[InstCombine] add tests for abs using bit hackery; NFC llvm-svn: 320068	2017-12-07 18:13:33 +00:00
Simon Pilgrim	6d9ac1b1eb	[X86] Replace tabs with spaces. NFCI. llvm-svn: 320065	2017-12-07 17:55:19 +00:00
Simon Pilgrim	386b23f1fa	[X86] Tag BMI/BMI2/TBM instructions scheduler classes Put these under UNARY/BINOP ALU itinerary classes for now - seems to be a good average value llvm-svn: 320064	2017-12-07 17:37:39 +00:00
Krzysztof Parzyszek	039d4d9286	[Hexagon] Generate HVX code for basic arithmetic operations Handle and, or, xor, add, sub, mul for vectors of i8, i16, and i32. llvm-svn: 320063	2017-12-07 17:37:28 +00:00
Simon Pilgrim	d2e93e76b8	[X86][TBM] Add TBM scheduling tests llvm-svn: 320062	2017-12-07 17:23:00 +00:00
Francis Visoiu Mistrih	e6fc3ce470	[CodeGen] Fix index when printing tied machine operands llvm-svn: 320061	2017-12-07 17:12:30 +00:00
Craig Topper	5db260fca4	[X86] Rename function in recently added test case to not be 'main' returning 'void'. NFC llvm-svn: 320059	2017-12-07 17:02:49 +00:00
Davide Italiano	f6e180d523	[DebugInfo] Move this test to X86/ now that it specifies a triple. Should bring back the arm/arm64 bots. Reported by Yvan Roux. llvm-svn: 320057	2017-12-07 16:10:39 +00:00
Simon Pilgrim	2983b46973	[X86] Tag SALC instructions scheduler class Treat these the same as LAHF/SAHF (although its not a x86_64 instruction) llvm-svn: 320055	2017-12-07 16:07:06 +00:00
Simon Pilgrim	ffce0d8fbc	[X86] Add LAHF/SAHF scheduling test llvm-svn: 320054	2017-12-07 16:04:20 +00:00
Simon Pilgrim	a13271bcba	[X86][VMX] Tag VMX instructions scheduler classes Tagged all as system instructions llvm-svn: 320053	2017-12-07 15:57:32 +00:00
Simon Pilgrim	a383f84233	[X86] Add SALC scheduling test llvm-svn: 320052	2017-12-07 15:46:58 +00:00
Simon Pilgrim	f1d599adb2	[X86] Tag LZCNT/TZCNT instructions scheduler classes Tagged as IMUL instructions for a reasonable approximation (ALU tends to be a lot faster) - POPCNT is currently tagged as FAdd which I think should be replaced with IMUL as well llvm-svn: 320051	2017-12-07 15:24:14 +00:00
Sanjay Patel	9012391af1	[DAGCombiner] eliminate shuffle of insert element I noticed this pattern in D38316 / D38388. We failed to combine a shuffle that is either repeating a scalar insertion at the same position in a vector or translated to a different element index. Like the earlier patch, this could be an instcombine too, but since we opted to make this a DAG transform earlier, I've made this one a DAG patch too. We do not need any legality checking because the new insert is identical to the existing insert except that it may have a different constant insertion operand. The constant insertion test in test/CodeGen/X86/vector-shuffle-combining.ll was the motivation for D38756. Differential Revision: https://reviews.llvm.org/D40209 llvm-svn: 320050	2017-12-07 15:17:58 +00:00
Igor Laevsky	4a4f2e8c67	[InstCombine] Don't crash on out of bounds index in the insertelement Differential Revision: https://reviews.llvm.org/D40390 llvm-svn: 320049	2017-12-07 15:00:52 +00:00
Simon Pilgrim	ff5212091a	[X86][FMA] Regenerate fma schedule tests llvm-svn: 320048	2017-12-07 14:51:47 +00:00
Simon Pilgrim	6b7cd86ca7	[X86][SVM] Tag SVM instructions scheduler classes Tagged all as system instructions llvm-svn: 320047	2017-12-07 14:35:17 +00:00
Francis Visoiu Mistrih	567611ef23	[CodeGen] Use more getMFIfAvailable llvm-svn: 320046	2017-12-07 14:32:15 +00:00
Simon Pilgrim	60411d9a8c	[X86] Tag RDRAND/RDSEED instruction scheduler classes llvm-svn: 320045	2017-12-07 14:18:48 +00:00
Simon Pilgrim	bd5f7455a2	[X86][X87] X87 math binop pseudo instructions don't need scheduling info llvm-svn: 320044	2017-12-07 14:07:18 +00:00
Simon Pilgrim	ca63dcce7f	[X86][SSE42] SSE42 string pseudo instructions don't need scheduling info llvm-svn: 320043	2017-12-07 13:52:07 +00:00
Simon Pilgrim	9a2898ed22	[X86] Regenerate RDTSC codegen tests llvm-svn: 320042	2017-12-07 13:50:29 +00:00
Dan Gohman	cdaa87dd2e	[WebAssemby] Support main functions with alternate signatures. WebAssembly requires caller and callee signatures to match, so the usual C runtime trick of calling main and having it just work regardless of whether main is defined as '()' or '(int argc, char *argv[])' doesn't work. Extend the FixFunctionBitcasts pass to rewrite main to use the latter form. llvm-svn: 320041	2017-12-07 13:49:27 +00:00
Simon Pilgrim	439679c085	[X86][RDSEED] Add rdseed scheduling tests llvm-svn: 320040	2017-12-07 13:47:17 +00:00
Simon Pilgrim	eb87fe62ec	[X86][RDRAND] Add rdrand scheduling tests llvm-svn: 320039	2017-12-07 13:46:47 +00:00
Alex Bradbury	f8f4b90544	[RISCV] MC layer support for the jump/branch instructions of the RVC extension Differential Revision: https://reviews.llvm.org/D40002 Patch by Shiva Chen. llvm-svn: 320038	2017-12-07 13:19:57 +00:00
Alex Bradbury	9f6aec4b7a	[RISCV] MC layer support for load/store instructions of the C (compressed) extension Differential Revision: https://reviews.llvm.org/D40001 Patch by Shiva Chen. llvm-svn: 320037	2017-12-07 12:50:32 +00:00
Alex Bradbury	87a54d6110	[RISCV][NFC] Use TargetRegisterClass::hasSubClassEq in storeRegToStackSlot/loadReadFromStackSlot Simply checking for register class equality will break once additional register classes are added (as is done for the RVC instruction set extension). llvm-svn: 320036	2017-12-07 12:45:05 +00:00
Nikolai Bozhenov	1cf9c54e5c	[Nios2] final infrastructure to provide compilation of a return from a function This patch includes all missing functionality needed to provide first compilation of a simple program that just returns from a function. I've added a test case that checks for "ret" instruction printed in assembly output. Patch by Andrei Grischenko (andrei.l.grischenko@intel.com) Differential revision: https://reviews.llvm.org/D39688 llvm-svn: 320035	2017-12-07 12:35:02 +00:00
Andrew V. Tischenko	44cfc51415	Add proper BTVER2 sched support for MOV instr. Differential Revision: https://reviews.llvm.org/D40345 llvm-svn: 320034	2017-12-07 11:19:49 +00:00
Jonas Devlieghere	e385d00960	[dsymutil] Add -verify option to run DWARF verifier after linking. This patch adds support for running the DWARF verifier on the linked debug info files. If the -verify options is specified and verification fails, dsymutil exists with abort with non-zero exit code. This behavior is not enabled by default. Differential revision: https://reviews.llvm.org/D40777 llvm-svn: 320033	2017-12-07 11:17:19 +00:00
Igor Laevsky	e8a3475b89	[FuzzMutate] Allow only sized pointers for the GEP instruction Differential Revision: https://reviews.llvm.org/D40837 llvm-svn: 320032	2017-12-07 11:10:11 +00:00
Alex Bradbury	72281a228b	[RISCV] Add missed tests for RV64D MC layer support Add tests missed in r320029. llvm-svn: 320031	2017-12-07 11:05:38 +00:00
Alex Bradbury	ee8950efd5	[RISCV] MC layer support for the standard RV64D instruction set extension llvm-svn: 320029	2017-12-07 11:04:18 +00:00
Alex Bradbury	4dd94e0ccd	[RISCV] MC layer support for the standard RV64F instruction set extension llvm-svn: 320028	2017-12-07 11:02:55 +00:00
Alex Bradbury	48f95a655d	[RISCV] MC layer support for the standard RV64A instruction set extension llvm-svn: 320027	2017-12-07 10:59:12 +00:00
Alex Bradbury	81def7224e	[RISCV] MC layer support for the standard RV64M instruction set extension llvm-svn: 320026	2017-12-07 10:56:07 +00:00
Pavel Labath	e8354fe606	[Testing/Support] Make matchers work with Expected<T&> Summary: This did not work because the ExpectedHolder was trying to hold the value in an Optional<T*>. Instead of trying to mimic the behavior of Expected and try to make ExpectedHolder work with references and non-references, I simply store the reference to the Expected object in the holder. I also add a bunch of tests for these matchers, which have helped me flesh out some problems in my initial implementation of this patch, and uncovered the fact that we are not consistent in quoting our values in the matcher output (which I also fix). Reviewers: zturner, chandlerc Subscribers: mgorny, llvm-commits Differential Revision: https://reviews.llvm.org/D40904 llvm-svn: 320025	2017-12-07 10:54:23 +00:00
Alex Bradbury	a6e6248307	[RISCV] MC layer support for the standard RV64I instructions llvm-svn: 320024	2017-12-07 10:53:48 +00:00
Alex Bradbury	7bc2a95bb9	[RISCV] MC layer support for the standard RV32D instruction set extension As the FPR32 and FPR64 registers have the same names, use validateTargetOperandClass in RISCVAsmParser to coerce a parsed FPR32 to an FPR64 when necessary. The rest of this patch is very similar to the RV32F patch. Differential Revision: https://reviews.llvm.org/D39895 llvm-svn: 320023	2017-12-07 10:46:23 +00:00
Francis Visoiu Mistrih	a8a83d150f	[CodeGen] Use MachineOperand::print in the MIRPrinter for MO_Register. Work towards the unification of MIR and debug output by refactoring the interfaces. For MachineOperand::print, keep a simple version that can be easily called from `dump()`, and a more complex one which will be called from both the MIRPrinter and MachineInstr::print. Add extra checks inside MachineOperand for detached operands (operands with getParent() == nullptr). https://reviews.llvm.org/D40836 * find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/kill: ([^ ]+) ([^ ]+)<def> ([^ ]+)/kill: \1 def \2 \3/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/kill: ([^ ]+) ([^ ]+) ([^ ]+)<def>/kill: \1 \2 def \3/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/kill: def ([^ ]+) ([^ ]+) ([^ ]+)<def>/kill: def \1 \2 def \3/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/<def>//g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<kill>/killed \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<imp-use,kill>/implicit killed \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<dead>/dead \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<def[ ],[ ]dead>/dead \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<imp-def[ ],[ ]dead>/implicit-def dead \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<imp-def>/implicit-def \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<imp-use>/implicit \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name ".s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<internal>/internal \1/g' find . $ -name ".mir" -o -name ".cpp" -o -name ".h" -o -name ".ll" -o -name "*.s" $ -type f -print0 \| xargs -0 sed -i '' -E 's/([^ ]+)<undef>/undef \1/g' llvm-svn: 320022	2017-12-07 10:40:31 +00:00
Alex Bradbury	0d6cf90663	[RISCV] MC layer support for the standard RV32F instruction set extension The most interesting part of this patch is probably the handling of rounding mode arguments. Sadly, the RISC-V assembler handles floating point rounding modes as a special "argument" when it would be more consistent to handle them like the atomics, opcode suffixes. This patch supports parsing this optional parameter, using InstAlias to allow parsing these floating point instructions when no rounding mode is specified. Differential Revision: https://reviews.llvm.org/D39893 llvm-svn: 320020	2017-12-07 10:26:05 +00:00
Alex Bradbury	d590c85753	[TableGen] Give the option of tolerating duplicate register names A number of architectures re-use the same register names (e.g. for both 32-bit FPRs and 64-bit FPRs). They are currently unable to use the tablegen'erated MatchRegisterName and MatchRegisterAltName, as tablegen (when built with asserts enabled) will fail. When the AllowDuplicateRegisterNames in AsmParser is set, duplicated register names will be tolerated. A backend can then coerce registers to the desired register class by (for instance) implementing validateTargetOperandClass. At least the in-tree Sparc backend could benefit from this, as does RISC-V (single and double precision floating point registers). Differential Revision: https://reviews.llvm.org/D39845 llvm-svn: 320018	2017-12-07 09:51:55 +00:00
Gadi Haber	dd62ac49cb	[X86][FMA][FMA4]: Adding full coverage of MC encoding for the FMA, FMA4 isa sets.<NFC> NFC. Adding MC regressions tests to cover the FMA and FMA4 ISA sets. This patch is part of a larger task to cover MC encoding of all X86 ISA Sets starting revision https://reviews.llvm.org/D39952 Reviewers: craig.topper, RKSimon, zvi Differential Revision: https://reviews.llvm.org/D40880 Change-Id: Ie39c0edce69ad647076b3d4e816948b2b6e1a9e4 llvm-svn: 320016	2017-12-07 09:16:34 +00:00
Gadi Haber	e33a0cb8a8	[X86][X87]: Adding full coverage of MC encoding for all X87 ISA Sets.<NFC> NFC. Currently, not all the X86 ISA Sets are covered by the MC regressions tests for X86. A full coverage needs to be added for each ISA set and for both 32bit and 64bit instructions + registers. This patch includes MC assembly tests for the X87 32bit and 64bit. Reviewers: craigt, RKSimon, zvi Differential Revision: https://reviews.llvm.org/D39952 Change-Id: I55e1719c09a70644a6a4073c720cb5341c80fee9 llvm-svn: 320015	2017-12-07 09:00:19 +00:00
Igor Laevsky	54d1ff0a58	[InstSimplify] Add tests for the rL319894 Differential Revision: https://reviews.llvm.org/D40650 llvm-svn: 320014	2017-12-07 08:52:24 +00:00
Craig Topper	dfecd45f37	[SelectionDAG] In SplitVecOp_EXTRACT_VECTOR_ELT, simplify the code that makes the type byte addressable. We can just extend the original vector to vXi1 and trust that the legalization process will revisit it. llvm-svn: 320013	2017-12-07 08:04:34 +00:00
Craig Topper	26ed8d1263	[SelectionDAG] Use TLI.getVectorIdxTy to determine type for an EXTRACT_VECTOR_ELT index instead of hardcoding MVT::i8. llvm-svn: 320012	2017-12-07 08:04:33 +00:00
Mikael Holmen	b5deac444d	Skip DBG instr in OptimizePHIs when looking for dead PHI cycles Summary: Changed use_instructions() to use_nodbg_instructions() when building an instruction set. We don't want the presence of debug info to affect the code we generate. Reviewers: dblaikie, Eugene.Zelenko, chandlerc, aprantl Reviewed By: aprantl Subscribers: aprantl, llvm-commits Differential Revision: https://reviews.llvm.org/D40882 llvm-svn: 320010	2017-12-07 07:01:21 +00:00
Leslie Zhai	8543d53fd9	[AVR] Override ParseDirective Reviewers: dylanmckay, kparzysz Reviewed By: dylanmckay Differential Revision: https://reviews.llvm.org/D38029 llvm-svn: 320009	2017-12-07 06:56:09 +00:00
Sam Clegg	8460b26403	Revert "[WebAssembly] Import the linear memory and function table." We need to a little time to prepare and lld-side change that supports this. Original change: https://reviews.llvm.org/D40875 llvm-svn: 320003	2017-12-07 03:05:45 +00:00
Sam Clegg	e1694f9bf8	[WebAssembly] section kind can be code Currently, when creating a named section, the Wasm frontend forces it to use `SectionKind::Data`, whereas in fact C++ does generate code sections with custom names. Patch by Nicholas Wilson Differential Revision: https://reviews.llvm.org/D40906 llvm-svn: 320002	2017-12-07 02:55:51 +00:00
Evgeniy Stepanov	cdf1abc365	Update BitCodeFormat. Add 2 recently added attributes to list of well-known attributes in BitCodeFormat.rst. llvm-svn: 319999	2017-12-07 01:38:20 +00:00
Davide Italiano	7495c762d1	[DebugInfo] Explicitly pass a triple to this test. As we emit different linetables format on different operating systems, this currently fails on linux. Speculative commit to fix the bots. llvm-svn: 319997	2017-12-07 01:22:10 +00:00
Davide Italiano	23b3f6da14	[MC/Dwarf] Use the older DWARF linetables format on Darwin. dsymutil doesn't yet understand the new format and the change, among others, breaks a large fraction of the debugger tests on mac OS. rdar://problem/35856354 llvm-svn: 319995	2017-12-07 00:57:25 +00:00
Alina Sbirlea	d6037ebeeb	[ModRefInfo] Replace remaining bit-wise operations with wrappers. llvm-svn: 319993	2017-12-07 00:43:19 +00:00
Dan Gohman	5cf6473903	[WebAssembly] Don't try to emit size information for unsized types Patch by John Sully! Fixes PR35164. Differential Revision: https://reviews.llvm.org/D39519 llvm-svn: 319991	2017-12-07 00:14:30 +00:00
Vedant Kumar	337b0db100	[Coverage] Scan ahead for the most-recent completed count (PR35495) This extends r319391. It teaches the segment builder to emit the right completed segment when more than one region ends at the same location. Fixes PR35495. llvm-svn: 319990	2017-12-07 00:01:15 +00:00
Dan Gohman	96d22e12a2	[WebAssembly] Import the linear memory and function table. Instead of having .o files contain linear-memory and function table definitions, use imports. This is more consistent with the stack pointer being imported, and it's consistent with the linker being the one to decide whether linear memory and function table are imported or defined in the linked output. This implements tool-conventions #23. Differential Revision: https://reviews.llvm.org/D40875 llvm-svn: 319989	2017-12-06 23:57:11 +00:00
Matt Morehouse	492a5b4830	[CMake] Use PRIVATE when linking LLVM fuzzers. More fuzzers missed by r319840. llvm-svn: 319987	2017-12-06 23:32:46 +00:00
Alina Sbirlea	9c26546d61	[ModRefInfo] Use ModRefInfo wrappers in FunctionModRefBehavior when testing for info found only in ModRefInfo [NFC]. llvm-svn: 319985	2017-12-06 23:12:43 +00:00
Florian Hahn	5d6a4e43ba	[AArch64] Add patterns to replace fsub fmul with fma fneg. Summary: This patch adds MachineCombiner patterns for transforming (fsub (fmul x y) z) into (fma x y (fneg z)). This has a lower latency on micro architectures where fneg is cheap. Patch based on work by George Steed. Reviewers: rengolin, joelkevinjones, joel_k_jones, evandro, efriedma Reviewed By: evandro Subscribers: aemerson, javed.absar, llvm-commits, kristof.beyls Differential Revision: https://reviews.llvm.org/D40306 llvm-svn: 319980	2017-12-06 22:48:36 +00:00
Adam Nemet	a502ee73c4	[LV] Interleaved access vectorization: fix computing new alias info As a new access is generated spanning across multiple fields, we need to propagate alias info from all the fields to form the most generic alias info. rdar://35602528 Differential Revision: https://reviews.llvm.org/D40617 llvm-svn: 319979	2017-12-06 22:42:24 +00:00
Krzysztof Parzyszek	d2967868be	[Hexagon] Recognize vdealb, vdealh, vshuffb and vshuffh specifically llvm-svn: 319978	2017-12-06 22:41:49 +00:00
Krzysztof Parzyszek	64533cf630	[Hexagon] Handle perfect shuffles on single vectors llvm-svn: 319965	2017-12-06 21:25:03 +00:00
Sanjay Patel	b6404a8ca6	[InstCombine] canonicalize constant-minus-boolean to select-of-constants This restores the half of: https://reviews.llvm.org/rL75531 that was reverted at: https://reviews.llvm.org/rL159230 For the x86 case mentioned there, we now produce: leal 1(%rdi), %eax subl %esi, %eax We have target hooks to invert this in DAGCombiner (and x86 is enabled) with: https://reviews.llvm.org/rL296977 https://reviews.llvm.org/rL311731 AArch64 and possibly other targets would probably benefit from enabling those hooks too. See PR30327: https://bugs.llvm.org/show_bug.cgi?id=30327#c2 Differential Revision: https://reviews.llvm.org/D40612 llvm-svn: 319964	2017-12-06 21:22:57 +00:00
Matthew Simpson	e363d2cebb	[PGO] Make indirect call promotion a utility This patch factors out the main code transformation utilities in the pgo-driven indirect call promotion pass and places them in Transforms/Utils. The change is intended to be a non-functional change, letting non-pgo-driven passes share a common implementation with the existing pgo-driven pass. The common utilities are used to conditionally promote indirect call sites to direct call sites. They perform the underlying transformation, and do not consider profile information. The pgo-specific details (e.g., the computation of branch weight metadata) have been left in the indirect call promotion pass. Differential Revision: https://reviews.llvm.org/D40658 llvm-svn: 319963	2017-12-06 21:22:54 +00:00
Dan Gohman	7ae3f46539	[WebAssembly] Commit a file I accidentally omitted from r319956. llvm-svn: 319962	2017-12-06 21:16:04 +00:00
Dan Gohman	ad19047d83	[WebAssembly] Remove WASM_STACK_POINTER. WASM_STACK_POINTER and the .stack_pointer directive are no longer needed now that the stack pointer global is an import. llvm-svn: 319956	2017-12-06 20:56:40 +00:00
Florian Hahn	001c3dd202	[MachineCombiner] Add up latencies of all instructions in new pattern. Summary: When calculating the RootLatency, we add up all the latencies of the deleted instructions. But for NewRootLatency we only add the latency of the new root instructions, ignoring the latencies of the other instructions inserted. This leads the combiner to underestimate the cost of patterns which add multiple instructions. This patch fixes that by summing up the latencies of all new instructions. For NewRootNode, the more complex getLatency function is used. Note that we may be slightly more precise than just summing up all latencies. For example, consider a pattern like r1 = INS1 .. r2 = INS2 .. r3 = INS3 r1, r2 I think in some other places, the total latency of the pattern would be estimated as lat(INS3) + max(lat(INS1), lat(INS2)). If you consider that worth changing, I think it would be best to do in a follow-up patch. Reviewers: Gerolf, sebpop, spop, fhahn Reviewed By: fhahn Subscribers: evandro, llvm-commits Differential Revision: https://reviews.llvm.org/D40307 llvm-svn: 319951	2017-12-06 20:27:33 +00:00
Alina Sbirlea	18fea013de	[ModRefInfo] Do not use ModRefInfo result in if conditions as this makes assumptions about the values in the enum. Replace with wrapper returning bool [NFC]. llvm-svn: 319949	2017-12-06 19:56:37 +00:00
Florian Hahn	115d99162c	[InlineFunction] Only replace call if there are VarArgs to forward. Summary: There is no need to replace the original call instruction if no VarArgs need to be forwarded. Reviewers: davide, rnk, majnemer, efriedma Reviewed By: efriedma Subscribers: eraman, llvm-commits Differential Revision: https://reviews.llvm.org/D40412 llvm-svn: 319947	2017-12-06 19:47:24 +00:00
Sanjay Patel	3e069f5724	[LoopUtils] simplify createTargetReduction(); NFCI llvm-svn: 319946	2017-12-06 19:37:00 +00:00
Simon Pilgrim	9afbe77a91	[X86][AVX512] Tag mask reg op instruction scheduler classes llvm-svn: 319945	2017-12-06 19:36:00 +00:00
Tim Shen	b684b1aa35	[Hexagon] Suppress more warnings on unused variables defined for asserts. llvm-svn: 319944	2017-12-06 19:33:42 +00:00
Alina Sbirlea	5beb1838bb	[ModRefInfo] Use createModRefInfo wrapper to create a ModRefInfo from FunctionModRefBehavior. llvm-svn: 319941	2017-12-06 19:23:03 +00:00
Tim Shen	7654ed03e3	[Hexagon] Suppress warnings on unused variables defind for asserts. llvm-svn: 319940	2017-12-06 19:22:19 +00:00
Rui Ueyama	efb5024e57	[COFF] Ignore semicolons in module definition identifiers Patch by David Major. The NSS project's .def files make heavy use of semicolons in a frightening attempt at portability: https://hg.mozilla.org/projects/nss/raw-file/tip/lib/ckfw/capi/nsscapi.def lld-link was treating the semicolon as part of the export name, resulting in unresolved symbols. This patch includes ';' in the list of characters to split on. Differential Revision: https://reviews.llvm.org/D39968 llvm-svn: 319933	2017-12-06 19:18:24 +00:00
Sanjay Patel	1ea7b6f7a1	[LoopUtils] fix variable name to match FMF vocabulary; NFC llvm-svn: 319928	2017-12-06 19:11:23 +00:00
Zachary Turner	c221dc71b1	Update obj2yaml and yaml2obj for .debug$H section. Differential Revision: https://reviews.llvm.org/D40842 llvm-svn: 319925	2017-12-06 18:58:48 +00:00
Davide Italiano	9c60c7dcf4	[Target] dumpr() is defined only in debug builds. This fixes the clang build on macOS. llvm-svn: 319923	2017-12-06 18:54:17 +00:00
Simon Pilgrim	7724b03cde	[X86][SSE] Regenerate vpmovm2/vpmov2m avx512 schedule tests llvm-svn: 319921	2017-12-06 18:47:37 +00:00
Simon Pilgrim	d255a6201d	[X86][AVX512] Tag scalar insert/extract instruction scheduler classes Classes don't look great but match what we're doing on SSE/AVX llvm-svn: 319920	2017-12-06 18:46:06 +00:00
Craig Topper	8b0f185c31	[X86] Simplify the TTI code for getInterleavedMemoryOpCost around for AVX512BW. NFCI Previously the lambda for AVX512 passed out a flag that indicated whether AVX512BW was required and that was checked against the AVX512BW subtarget flag outside. This patch changes the interface to pass the AVX512BW subtarget bit in and return its value if we detect 16 or 8 bit types. llvm-svn: 319919	2017-12-06 18:40:46 +00:00
Shoaib Meenai	6aa13adf0e	[cmake] Remove unnecessary header include in atomics check The header include was required to work around PR19898, as noted in that comment. That PR has since been marked resolved fixed, and the configuration check passes without the header inclusion both when compiling on Windows with cl and when cross-compiling on Linux using clang-cl. I noticed this because the inclusion was cased incorrectly (Intrin.h instead of intrin.h), which when cross-compiling on a case sensitive file system would cause the intrin.h from the Windows SDK to be included (which LLVM can't handle) instead of the one from clang's resource directory, making the check fail. This is the same issue as r309980. Correcting the case of the inclusion makes the check pass when cross compiling, but it seems better to get rid of the inclusion entirely, since it appears to be unnecessary now. Differential Revision: https://reviews.llvm.org/D40910 llvm-svn: 319917	2017-12-06 18:33:07 +00:00
Simon Pilgrim	809c024b3d	[X86][AVX2] Tag MASKMOV instruction scheduler classes llvm-svn: 319915	2017-12-06 18:24:48 +00:00
Craig Topper	fa172a5251	[X86] Regenerate test for r319778 llvm-svn: 319914	2017-12-06 18:04:39 +00:00
Simon Pilgrim	df05251921	[X86][AVX512] Tag aligned/unaligned move instruction scheduler classes llvm-svn: 319913	2017-12-06 17:59:26 +00:00
Simon Pilgrim	3ee91ade9b	[X86][AVX] Regenerate vpmovm2/vpmov2m avx512 schedule tests llvm-svn: 319912	2017-12-06 17:57:18 +00:00
Craig Topper	c3c3ebf29d	[X86] Attempt to fix a ubsan failure in the autoupgrade of kunpck intrinsics. llvm-svn: 319911	2017-12-06 17:54:07 +00:00
Zvi Rackover	2e6e88f689	InstructionSimplify: 'extractelement' with an undef index is undef Summary: An undef extract index can be arbitrarily chosen to be an out-of-range index value, which would result in the instruction being undef. This change closes a gap identified while working on lowering vector permute intrinsics with variable index vectors to pure LLVM IR. Reviewers: arsenm, spatel, majnemer Reviewed By: arsenm, spatel Subscribers: fhahn, nhaehnle, wdng, llvm-commits Differential Revision: https://reviews.llvm.org/D40231 llvm-svn: 319910	2017-12-06 17:51:46 +00:00
Artem Belevich	a659d2590e	[NVPTX,CUDA] Added llvm.nvvm.fns intrinsic and matching __nvvm_fns builtin in clang. Differential Revision: https://reviews.llvm.org/D40872 llvm-svn: 319909	2017-12-06 17:50:05 +00:00
Zvi Rackover	ffaed72089	AMDGPU Tests: Change a case to be run with -O0 D40231 requires to run case with -O0 to prevent InstructionSimplify from transforming an extractelement with undef index. llvm-svn: 319907	2017-12-06 17:40:09 +00:00
Jonas Paulsson	a74ff71a37	[SystemZ] Add IntrWriteMem flag to int_s390_tabort intrinsic Tabort (transaction abort) does not load from memory. mayLoad flag removed from corresponding TABORT machine instruction. Review: Ulrich Weigand llvm-svn: 319905	2017-12-06 17:01:08 +00:00
Adam Nemet	9e5e51aeed	[opt-viewer] Suppress noisy Swift remarks Most likely, this is not how we want to handle this in the long term. This code should probably be in the Swift repo and somehow plugged into the opt-viewer. This is still however very experimental at this point so I don't want to over-engineer it at this point. llvm-svn: 319902	2017-12-06 16:50:50 +00:00
Krzysztof Parzyszek	7d37dd8902	[Hexagon] Generate HVX code for vector construction and access Support for: - build vector, - extract vector element, subvector, - insert vector element, subvector, - shuffle. llvm-svn: 319901	2017-12-06 16:40:37 +00:00
Simon Pilgrim	aa902be158	[X86][AVX512] Tag BROADCAST instruction scheduler classes llvm-svn: 319900	2017-12-06 15:48:40 +00:00
Nirav Dave	7d8f3e0c93	[ARM][AArch64][DAG] Reenable post-legalize store merge Reenable post-legalize stores with constant merging computation and corresponding test case. * Properly truncate store merge constants * Disable merging of truncated stores floating points * Ensure merges of constant stores into a single vector are constructed from legal elements. Reviewers: eastig, efriedma Reviewed By: eastig Subscribers: spatel, rengolin, aemerson, javed.absar, kristof.beyls, hiraditya, llvm-commits Differential Revision: https://reviews.llvm.org/D40701 llvm-svn: 319899	2017-12-06 15:30:13 +00:00
Don Hinton	2e004b3ddb	[cmake] Move CMAKE_(C\|CXX)_COMPILER variables before CROSS_TOOLCHAIN_FLAGS so they can be overridden when cross compiling. Summary: Since CROSS_TOOLCHAN_FLAGS can set CMAKE_(C\|CXX)_COMPILER variables, move the compiler variables up front so they can be overridden. This is a followup to https://reviews.llvm.org/D40229 committed in rL319620. Thanks to Pavel Labath for reporting this issue. Reviewers: labath, beanz Subscribers: mgorny, llvm-commits Differential Revision: https://reviews.llvm.org/D40896 llvm-svn: 319898	2017-12-06 15:25:14 +00:00
Simon Pilgrim	a9282309e5	[X86][AVX512] Regenerate vpmovm2/vpmov2m avx512 schedule tests llvm-svn: 319895	2017-12-06 14:07:38 +00:00
Igor Laevsky	03655c7636	[InstSimplify] Fold insertelement into undef if index is out of bounds Differential Revision: https://reviews.llvm.org/D40650 llvm-svn: 319894	2017-12-06 14:04:45 +00:00
Jonas Paulsson	19380bae05	[SystemZ] Bugfix in expandRxSBG() Csmith discovered a program that caused wrong code generation with -O0: When handling a SIGN_EXTEND in expandRxSBG(), RxSBG.BitSize may be less than the Input width (if a truncate was previously traversed), so maskMatters() should be called with a masked based on the width of the sign extend result instead. Review: Ulrich Weigand llvm-svn: 319892	2017-12-06 13:53:24 +00:00
Benjamin Kramer	1e9bf765a1	[X86] Avoid unused variable warning in Release builds. NFCI. llvm-svn: 319891	2017-12-06 13:32:36 +00:00
Simon Pilgrim	07dc6d6975	[X86][AVX512] Drop default NoItinerary arguments that aren't needed Requires reordering of AVX512_maskable_common arguments, but helps track what is still missing itinerary tags llvm-svn: 319890	2017-12-06 13:14:44 +00:00
Max Kazantsev	d4f5987c58	[SCEV][NFC] Check NoWrap flags before lexicographical comparison of SCEVs Lexicographical comparison of SCEV trees is potentially expensive for big expression trees. We can define ordering between them for AddRecs and N-ary operations by SCEV NoWrap flags to make non-equality check cheaper. This change does not prevent grouping eqivalent SCEVs together and is not supposed to have any meaningful impact on behavior of any transforms. Reviewed By: sanjoy Differential Revision: https://reviews.llvm.org/D40645 llvm-svn: 319889	2017-12-06 12:44:56 +00:00

... 5 6 7 8 9 ...

158121 Commits