llvm-project

Commit Graph

Author	SHA1	Message	Date
Artur Pilipenko	c43b20a43b	[DAGCombine] NFC. MatchLoadCombine extract MemoryByteOffset lambda helper This refactoring will simplify the upcoming change to fix the bug in folding patterns with non-zero offsets on BE targets. llvm-svn: 296332	2017-02-27 11:42:54 +00:00
Artur Pilipenko	f2c26e0bf2	[DAGCombine] NFC. MatchLoadCombine remember the first byte provider, not the load node This refactoring will simplify the upcoming change to fix a bug in folding patterns with non-zero offsets on BE targets. llvm-svn: 296331	2017-02-27 11:40:14 +00:00
Sjoerd Meijer	6d171006f4	AArch64AsmParser: don't try to parse “[1]” for non-vector register operands There are no instructions that have "[1]" as part of the assembly string; FMOVXDhighr is out of date. This removes dead code. Differential Revision: https://reviews.llvm.org/D30165 llvm-svn: 296327	2017-02-27 10:51:11 +00:00
Konstantin Zhuravlyov	972948b36e	[AMDGPU] Runtime metadata fixes: - Verify that runtime metadata is actually valid runtime metadata when assembling, otherwise we could accept the following when assembling, but ocl runtime will reject it: .amdgpu_runtime_metadata { amd.MDVersion: [ 2, 1 ], amd.RandomUnknownKey, amd.IsaInfo: ... - Make IsaInfo optional, and always emit it. Differential Revision: https://reviews.llvm.org/D30349 llvm-svn: 296324	2017-02-27 07:55:17 +00:00
Brian Cain	50aa37b96c	llvm-mc-fuzzer: add support for assembly This creates an llvm-mc-disassemble-fuzzer from the existing llvm-mc-fuzzer and finishing the assemble support in llvm-mc-assemble-fuzzer. llvm-svn: 296323	2017-02-27 06:22:17 +00:00
Craig Topper	9ef28ba53c	[APInt] Use UINT64_MAX instead of ~integerPart(0). NFC llvm-svn: 296322	2017-02-27 06:05:33 +00:00
Craig Topper	ed0101a0b9	[X86] Check for less than 0 rather than explicit compare with -1. NFC llvm-svn: 296321	2017-02-27 06:05:30 +00:00
Amaury Sechet	681472cd0f	Do full codegen for various tests. NFC llvm-svn: 296305	2017-02-27 01:15:57 +00:00
Craig Topper	5d25d03844	[APInt] Use UINT64_MAX instead of ~uint64_t(0ULL). NFC llvm-svn: 296301	2017-02-26 21:15:18 +00:00
Craig Topper	7d7b6d767d	[APInt] Use UINT64_MAX instead of ~0ULL. NFC llvm-svn: 296300	2017-02-26 19:28:48 +00:00
Craig Topper	a8b26b8715	[APInt] Remove unnecessary early out from getLowBitsSet. The same case is handled equally well by the next check. llvm-svn: 296299	2017-02-26 19:28:45 +00:00
Xin Tong	3ca169f3a9	Update comments. NFCI llvm-svn: 296298	2017-02-26 19:08:44 +00:00
Daniel Jasper	3ca4525612	Revert "[CGP] Split some critical edges coming out of indirect branches" This reverts commit r296149 as it leads to crashes when compiling for PPC. llvm-svn: 296295	2017-02-26 11:09:12 +00:00
Davide Italiano	49a0aac1aa	[LoopDeletion] Modernize and simplify a bit. NFCI. llvm-svn: 296294	2017-02-26 07:08:20 +00:00
Craig Topper	6028584d8c	[X86] Fix execution domain for cmpss/sd instructions. llvm-svn: 296293	2017-02-26 06:45:59 +00:00
Craig Topper	036693302b	[AVX-512] Fix execution domain for scalar commutable min/max instructions. llvm-svn: 296292	2017-02-26 06:45:56 +00:00
Craig Topper	e70231be51	[AVX-512] Fix execution domain for vmovhpd/lpd/hps/lps. llvm-svn: 296291	2017-02-26 06:45:54 +00:00
Craig Topper	fe25988c68	[AVX-512] Fix the execution domain for AVX-512 integer broadcasts. llvm-svn: 296290	2017-02-26 06:45:51 +00:00
Craig Topper	49ba3f5406	[AVX-512] Disable the redundant patterns in the VPBROADCASTBr_Alt and VPBROADCASTWr_Alt instructions. NFC llvm-svn: 296289	2017-02-26 06:45:48 +00:00
Craig Topper	6bf9b809ce	[AVX-512] Fix execution domain for VPMADD52 instructions. llvm-svn: 296288	2017-02-26 06:45:45 +00:00
Craig Topper	9ef7e44d2f	[AVX-512] Use update_llc_test_checks.py to regenerate a test. llvm-svn: 296287	2017-02-26 06:45:43 +00:00
Craig Topper	aa8e903150	[AVX-512] Fix the execution domain for VSCALEF instructions. llvm-svn: 296286	2017-02-26 06:45:40 +00:00
Craig Topper	cac5d698df	[AVX-512] Fix execution domain of scalar VRANGE/REDUCE/GETMANT with sae. llvm-svn: 296285	2017-02-26 06:45:37 +00:00
Craig Topper	ed64904c74	[X86] Fix the execution domain for scalar SQRT intrinsic instruction. llvm-svn: 296284	2017-02-26 06:45:35 +00:00
Craig Topper	a87b40051d	[X86] Add an additional CHECK prefix to a test. Some of the cases used it, but it wasn't on the FileCheck command lines. llvm-svn: 296283	2017-02-26 06:45:32 +00:00
Xin Tong	42ef2177af	[SCCP] Remove manual folding of terminator instructions. Summary: BranchInst, SwitchInst (with non-default case) with Undef as input is not possible at this point. As we always default-fold terminator to one target in ResolvedUndefsIn and set the input accordingly. So we should only have constantint/blockaddress here. If ConstantFoldTerminator fails, that could mean 2 things. 1. ConstantFoldTerminator is doing something unexpected, i.e. not folding on constantint or blockaddress and not making blocks that should be dead dead. 2. This is not a terminator on constantint or blockaddress. Its on a constant or overdefined, then this block should not be dead. In both cases, we should assert. Reviewers: davide, efriedma, sanjoy Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30381 llvm-svn: 296281	2017-02-26 02:11:24 +00:00
David L. Jones	d95c34abe1	[X86] Clean up test/CodeGen/X86/2006-03-02-InstrSchedBug.ll Summary: Migrated from grep to FileCheck. Re-indented code, removed boilerplate comments. Added 'entry' label at beginning of basic block. Patch by Jorge Gorbe! Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30320 llvm-svn: 296280	2017-02-26 01:32:35 +00:00
Nirav Dave	73cd0194cf	Revert "In visitSTORE, always use FindBetterChain, rather than only when UseAA is enabled." This reverts commit r296252 until 256-bit operations are more efficiently generated in X86. llvm-svn: 296279	2017-02-26 01:27:32 +00:00
Eric Christopher	4a8208c266	vec perm can go down either pipeline on P8. No observable changes, spotted while looking at the scheduling description. llvm-svn: 296277	2017-02-26 00:11:58 +00:00
Sanjoy Das	4897cea4ed	Fix signed-unsigned comparison warning llvm-svn: 296274	2017-02-25 22:25:48 +00:00
Sanjoy Das	39a684d117	[ValueTracking] Don't do an unchecked shift in ComputeNumSignBits Summary: Previously we used to return a bogus result, 0, for IR like `ashr %val, -1`. I've also added an assert checking that `ComputeNumSignBits` at least returns 1. That assert found an already checked in test case where we were returning a bad result for `ashr %val, -1`. Fixes PR32045. Reviewers: spatel, majnemer Reviewed By: spatel, majnemer Subscribers: efriedma, mcrosier, llvm-commits Differential Revision: https://reviews.llvm.org/D30311 llvm-svn: 296273	2017-02-25 20:30:45 +00:00
Simon Pilgrim	0f5fb5f549	[APInt] Add APInt::extractBits() method to extract APInt subrange (reapplied) The current pattern for extract bits in range is typically: Mask.lshr(BitOffset).trunc(SubSizeInBits); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation of memory for the temporary variable. This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::extractBits() helper method which avoids the temporary memory allocation. Differential Revision: https://reviews.llvm.org/D30336 llvm-svn: 296272	2017-02-25 20:01:58 +00:00
Craig Topper	2caa97c891	[AVX-512] Fix the execution domain for scalar FMA instructions. llvm-svn: 296271	2017-02-25 19:36:28 +00:00
Craig Topper	176f3310b6	[AVX-512] Fix the execution domain on some instructions. llvm-svn: 296270	2017-02-25 19:18:11 +00:00
Craig Topper	0524035fb4	[AVX-512] Add an additional test case to show the execution domain for vrqsrtsd is wrong. llvm-svn: 296269	2017-02-25 19:18:08 +00:00
Craig Topper	9b43d459bf	[AVX-512] Use update_llc_test_checks.py to regenerate the avx512er intrinsic test. llvm-svn: 296268	2017-02-25 19:18:04 +00:00
Nirav Dave	4a20711826	reenable accidentally disabled test NFC. llvm-svn: 296266	2017-02-25 19:11:53 +00:00
Craig Topper	d2011e3612	[AVX-512] Remove unnecessary masked versions of VCVTSS2SD and VCVTSD2SS using the scalar register class. We only have patterns for the masked intrinsics. llvm-svn: 296264	2017-02-25 18:43:42 +00:00
Craig Topper	3b8aca2ecf	[ExecutionDepsFix] Don't make copies of LiveReg objects when collecting operands for soft instructions Summary: While collecting operands we make copies of the LiveReg objects which are stored in the LiveRegs array. If the instruction uses the same register multiple times we end up with multiple copies. Later we iterate through the collected list of LiveReg objects and merge DomainValues. In the process of doing this the merge function can change the contents of the original LiveReg object in the LiveRegs array, but not the copies that have been made. So when we get to the second usage of the register we end up seeing a stale copy of the LiveReg object. To fix this I've stopped copying and now just store a pointer to the original LiveReg object. Another option might be to avoid adding the same register to the Regs array twice, but this approach seemed simpler. The included test case exposes this bug due to an AVX-512 masked OR instruction using the same register for the passthru operand and one of the inputs to the OR operation. Fixes PR30284. Reviewers: RKSimon, stoklund, MatzeB, spatel, myatsina Reviewed By: RKSimon Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30242 llvm-svn: 296260	2017-02-25 18:12:25 +00:00
Artyom Skrobov	ac56719231	No need to copy the variable [NFC] llvm-svn: 296259	2017-02-25 17:18:09 +00:00
NAKAMURA Takumi	05a75e40da	Revert r296215, "[PDB] General improvements to Stream library." and followings. r296215, "[PDB] General improvements to Stream library." r296217, "Disable BinaryStreamTest.StreamReaderObject temporarily." r296220, "Re-enable BinaryStreamTest.StreamReaderObject." r296244, "[PDB] Disable some tests that are breaking bots." r296249, "Add static_cast to silence -Wc++11-narrowing." std::errc::no_buffer_space should be used for OS-oriented errors for socket transmission. (Seek discussions around llvm/xray.) I could substitute s/no_buffer_space/others/g, but I revert whole them ATM. Could we define and use LLVM errors there? llvm-svn: 296258	2017-02-25 17:04:23 +00:00
Amaury Sechet	09ecd3117e	Update various test's codegen. NFC llvm-svn: 296257	2017-02-25 16:46:47 +00:00
Amaury Sechet	7bda8cdef2	Add test for known bits in uaddo and saddo. llvm-svn: 296255	2017-02-25 15:58:34 +00:00
Artyom Skrobov	2716910caf	The automatic CHECK: to CHECK-LABEL: conversion, back in 2013, had missed most labels in this test because they didn't end with a colon. llvm-svn: 296254	2017-02-25 15:17:16 +00:00
Victor Leschuk	96d9981ec6	[DebugInfo] Skip implicit_const attributes when dumping .debug_info. NFC. When dumping .debug_info section we loop through all attributes mentioned in .debug_abbrev section and dump values using DWARFFormValue::extractValue(). We need to skip implicit_const attributes here as their values are not really located in .debug_info but directly in .debug_abbrev. This patch fixes triggered assert() in DWARFFormValue::extractValue() caused by trying to access implicit_const values from .debug_info. llvm-svn: 296253	2017-02-25 13:15:57 +00:00
Nirav Dave	beabf456df	In visitSTORE, always use FindBetterChain, rather than only when UseAA is enabled. Recommiting after fixup of 32-bit aliasing sign offset bug in DAGCombiner. * Simplify Consecutive Merge Store Candidate Search Now that address aliasing is much less conservative, push through simplified store merging search and chain alias analysis which only checks for parallel stores through the chain subgraph. This is cleaner as the separation of non-interfering loads/stores from the store-merging logic. When merging stores search up the chain through a single load, and finds all possible stores by looking down from through a load and a TokenFactor to all stores visited. This improves the quality of the output SelectionDAG and the output Codegen (save perhaps for some ARM cases where we correctly constructs wider loads, but then promotes them to float operations which appear but requires more expensive constant generation). Some minor peephole optimizations to deal with improved SubDAG shapes (listed below) Additional Minor Changes: 1. Finishes removing unused AliasLoad code 2. Unifies the chain aggregation in the merged stores across code paths 3. Re-add the Store node to the worklist after calling SimplifyDemandedBits. 4. Increase GatherAllAliasesMaxDepth from 6 to 18. That number is arbitrary, but seems sufficient to not cause regressions in tests. 5. Remove Chain dependencies of Memory operations on CopyfromReg nodes as these are captured by data dependence 6. Forward loads-store values through tokenfactors containing {CopyToReg,CopyFromReg} Values. 7. Peephole to convert buildvector of extract_vector_elt to extract_subvector if possible (see CodeGen/AArch64/store-merge.ll) 8. Store merging for the ARM target is restricted to 32-bit as some in some contexts invalid 64-bit operations are being generated. This can be removed once appropriate checks are added. This finishes the change Matt Arsenault started in r246307 and jyknight's original patch. Many tests required some changes as memory operations are now reorderable, improving load-store forwarding. One test in particular is worth noting: CodeGen/PowerPC/ppc64-align-long-double.ll - Improved load-store forwarding converts a load-store pair into a parallel store and a memory-realized bitcast of the same value. However, because we lose the sharing of the explicit and implicit store values we must create another local store. A similar transformation happens before SelectionDAG as well. Reviewers: arsenm, hfinkel, tstellarAMD, jyknight, nhaehnle llvm-svn: 296252	2017-02-25 11:43:58 +00:00
Piotr Padlewski	4810772905	[Doc] Modernize programmers manual Summary: Fixed bunch of for loops to range based for loop and bunch of rendundat types with auto. Reviewers: echristo, silvas, chandlerc Subscribers: mehdi_amini, llvm-commits Differential Revision: https://reviews.llvm.org/D30338 llvm-svn: 296251	2017-02-25 10:33:37 +00:00
Xin Tong	b529c66ef3	Empty line. NFCI llvm-svn: 296250	2017-02-25 08:10:28 +00:00
Daniel Jasper	6249d4d337	Add static_cast to silence -Wc++11-narrowing. llvm-svn: 296249	2017-02-25 07:53:36 +00:00
Zachary Turner	4284ce1348	[PDB] Disable some tests that are breaking bots. This has to do with big endian, but I can't fix it until Monday. The code itself is fine, just the tests are wrong. Disabling 3 tests for now. llvm-svn: 296244	2017-02-25 05:57:57 +00:00

1 2 3 4 5 ...

145462 Commits