llvm-project

Commit Graph

Author	SHA1	Message	Date
Duncan P. N. Exon Smith	e07f13ae35	Support: Add dwarf::getAttributeEncoding() llvm-svn: 228470	2015-02-06 23:46:49 +00:00
Duncan P. N. Exon Smith	dd563dd341	Support: Rewrite AttributeEncodingString(), NFC llvm-svn: 228469	2015-02-06 23:45:37 +00:00
Duncan P. N. Exon Smith	8d4eeb53f6	Support: Stop stringifying DW_ATE_{lo,hi}_user llvm-svn: 228468	2015-02-06 23:44:24 +00:00
Hal Finkel	12b607ae1b	[PowerPC] Fixup incomplete revert of test/CodeGen/PowerPC/tls-pic.ll llvm-svn: 228467	2015-02-06 23:30:06 +00:00
Lang Hames	e1d10711d1	[Orc] Add a Kaleidoscope/Orc tutorial demonstrating lazy-irgen. This tutorial builds on the lazy_codegen kaleidoscope/orc tutorial by making a small set of changes (~75 lines diff) to defer ir-generation for function definitions until functions are actually referenced. llvm-svn: 228466	2015-02-06 23:26:33 +00:00
Kevin Enderby	74b43cb403	Add code to llvm-objdump so the -section option with -macho will dump literal sections with the Mach-O S_{4,8,16}BYTE_LITERALS section types. llvm-svn: 228465	2015-02-06 23:25:38 +00:00
Ahmed Bougacha	df956a2e78	[AArch64] Use the source location of the IR branch when creating Bcc from a conditional branch fed by an add/sub/mul-with-overflow node. We previously used the SDLoc of the overflow node, for no good reason. In some cases, this led to the Bcc and B terminators having different source orders, and DBG_VALUEs being inserted between them. The real issue is with the code that can't handle DBG_VALUEs between terminators: the few places affected by this will be fixed soon. In the meantime, fixing the SDLoc is a positive change no matter what. No tests, as I have no idea how to get .loc emitted for branches? rdar://19347133 llvm-svn: 228463	2015-02-06 23:15:39 +00:00
Simon Pilgrim	76cb85a6c7	[X86][AVX2] Begun adding AVX2 integer stack folding tests. llvm-svn: 228462	2015-02-06 23:12:15 +00:00
Hal Finkel	0d2a1515d5	Revert "r227976 - [PowerPC] Yet another approach to __tls_get_addr" and related fixups Unfortunately, even with the workaround of disabling the linker TLS optimizations in Clang restored (which has already been done), this still breaks self-hosting on my P7 machine (-O3 -DNDEBUG -mcpu=native). Bill is currently working on an alternate implementation to address the TLS issue in a way that also fully elides the linker bug (which, unfortunately, this approach did not fully), so I'm reverting this now. llvm-svn: 228460	2015-02-06 23:07:40 +00:00
Lang Hames	ed218a0da0	[Orc] Add a Kaleidoscope/Orc tutorial demonstrating lazy-codegen. This tutorial builds on the initial kaleidoscope/orc tutorial by adding a LazyEmittingLayer to the custom stack. This extra layer defers compilation of modules in the JIT until they are statically referenced. llvm-svn: 228459	2015-02-06 23:04:53 +00:00
Duncan P. N. Exon Smith	d40af00e66	Support: Add dwarf::getLanguage() llvm-svn: 228458	2015-02-06 22:55:13 +00:00
Duncan P. N. Exon Smith	0317944b1a	Support: Rewrite dwarf::LanguageString(), NFC llvm-svn: 228457	2015-02-06 22:53:19 +00:00
Lang Hames	d855e4515e	[Orc] Add a Kaleidoscope tutorial for Orc demonstrating eager compilation. This tutorial demonstrates a very basic custom Orc JIT stack that performs eager compilation: All modules are CodeGen'd immediately upon being added to the JIT. llvm-svn: 228456	2015-02-06 22:52:04 +00:00
Duncan P. N. Exon Smith	af677ebb41	IR: Allow 32-bits for lines in debug location Remove unnecessary restriction of 24-bits for line numbers in `MDLocation`. The rest of the debug info schema (with the exception of local variables) uses 32-bits for line numbers. As I introduce the specialized nodes, it makes sense to canonicalize on one size or the other. llvm-svn: 228455	2015-02-06 22:50:13 +00:00
Lang Hames	acf8288669	[Orc] Add more missing headers. llvm-svn: 228454	2015-02-06 22:48:43 +00:00
Sanjay Patel	3d982214b0	use local variables; NFC llvm-svn: 228452	2015-02-06 22:43:52 +00:00
Duncan P. N. Exon Smith	4031beb10a	Support: Stop stringifying DW_LANG_{lo,hi}_user llvm-svn: 228451	2015-02-06 22:34:48 +00:00
Duncan P. N. Exon Smith	a81f7e1a63	AsmParser: Use DW_TAG_hi_user instead of magic constant, NFC llvm-svn: 228448	2015-02-06 22:29:35 +00:00
Duncan P. N. Exon Smith	c6da7f1c74	AsmWriter: Extract writeTag(), NFC llvm-svn: 228447	2015-02-06 22:28:05 +00:00
Duncan P. N. Exon Smith	61a0933007	AsmWriter: Extract writeMetadataAsOperand(), NFC llvm-svn: 228446	2015-02-06 22:27:22 +00:00
Evgeniy Stepanov	4e12057760	[msan] Fix "missing origin" in atomic store. An atomic store always make the target location fully initialized (in the current implementation). It should not store origin. Initialized memory can't have meaningful origin, and, due to origin granularity (4 bytes) there is a chance that this extra store would overwrite meaningfull origin for an adjacent location. llvm-svn: 228444	2015-02-06 21:47:39 +00:00
Cameron Esfahani	f0bf1947b7	Test commit to see if it triggers an email to llvm-commits. No change. llvm-svn: 228442	2015-02-06 21:33:08 +00:00
Zachary Turner	8022f24fe8	Try to fix Makefile build for LLVMDebugInfoPDB. llvm-svn: 228437	2015-02-06 20:42:03 +00:00
Zachary Turner	0e9e66330a	Resubmit "Create lib/DebugInfo/PDB" (r228428) This change resubmits the patch that broke the build, this time without unittests. The unittests will be submitted separately after the problem has been addressed: --Original Commit Message-- Create lib/DebugInfo/PDB. This patch creates a platform-independent interface to a PDB reader. There is currently no implementation of this interface, which will be provided in future patches. This defines the basic object model which any implementation must conform to. Reviewed by: David Blaikie Differential Revision: http://reviews.llvm.org/D7356 llvm-svn: 228435	2015-02-06 20:30:52 +00:00
Michael Zolotukhin	7af83c1f39	Use estimated number of optimized insns in unroll-threshold computation. If complete-unroll could help us to optimize away N% of instructions, we might want to do this even if the final size would exceed loop-unroll threshold. However, we don't want to unroll huge loop, and we are add AbsoluteThreshold to avoid that - this threshold will never be crossed, even if we expect to optimize 99% instructions after that. llvm-svn: 228434	2015-02-06 20:20:40 +00:00
Michael Zolotukhin	4e8598eee3	[InstSimplify] Add SimplifyFPBinOp function. It is a variation of SimplifyBinOp, but it takes into account FastMathFlags. It is needed in inliner and loop-unroller to accurately predict the transformation's outcome (previously we dropped the flags and were too conservative in some cases). Example: float foo(float a, float b) { float r; if (a[1] b) r = /* a lot of expensive computations /; else r = 1; return r; } float boo(float a) { return foo(a, 0.0); } Without this patch, we don't inline 'foo' into 'boo'. llvm-svn: 228432	2015-02-06 20:02:51 +00:00
Zachary Turner	8c89a82c88	Revert "Create lib/DebugInfo/PDB." This reverts commit 21028, as it is causing failures in LLVMConfig. llvm-svn: 228431	2015-02-06 20:00:18 +00:00
Kostya Serebryany	db4d645714	[fuzzer] move default sanitizer options to a separate file llvm-svn: 228429	2015-02-06 19:52:07 +00:00
Zachary Turner	079ba9224f	Create lib/DebugInfo/PDB. This patch creates a platform-independent interface to a PDB reader. There is currently no implementation of this interface, which will be provided in future patches. This defines the basic object model which any implementation must conform to. Reviewed by: David Blaikie Differential Revision: http://reviews.llvm.org/D7356 llvm-svn: 228428	2015-02-06 19:44:09 +00:00
Lang Hames	8a6f35546b	[Orc] Move SectionMemoryManager's implementation from MCJIT to ExecutionEngine. This is a more sensible home for SectionMemoryManager, and allows the implementation to be shared between Orc and MCJIT. llvm-svn: 228427	2015-02-06 19:36:40 +00:00
Lang Hames	510e2f48f0	[Orc] Add some missing headers. llvm-svn: 228426	2015-02-06 19:34:40 +00:00
Lang Hames	569720629c	[Orc] Fix syntax error in LazyEmittingLayer::removeModuleSet. This was a trivial think-o, but it's in a method of a templated class and doesn't have any callers yet, so the compiler let it pass. I hope to add a unit test to cover this soon. llvm-svn: 228425	2015-02-06 19:34:04 +00:00
Quentin Colombet	a8cb36ea43	[LiveIntervalAnalysis] Speed up creation of live ranges for physical registers by using a segment set. The patch addresses a compile-time performance regression in the LiveIntervals analysis pass (see http://llvm.org/bugs/show_bug.cgi?id=18580). This regression is especially critical when compiling long functions. Our analysis had shown that the most of time is taken for generation of live intervals for physical registers. Insertions in the middle of the array of live ranges cause quadratic algorithmic complexity, which is apparently the main reason for the slow-down. Overview of changes: - The patch introduces an additional std::set<Segment>* member in LiveRange for storing segments in the phase of initial creation. The set is used if this member is not NULL, otherwise everything works the old way. - The set of operations on LiveRange used during initial creation (i.e. used by createDeadDefs and extendToUses) have been reimplemented to use the segment set if it is available. - After a live range is created the contents of the set are flushed to the segment vector, because the set is not as efficient as the vector for the later uses of the live range. After the flushing, the set is deleted and cannot be used again. - The set is only for live ranges computed in LiveIntervalAnalysis::computeLiveInRegUnits() and getRegUnit() but not in computeVirtRegs(), because I did not bring any performance benefits to computeVirtRegs() and for some examples even brought a slow down. Patch by Vaidas Gasiunas <vaidas.gasiunas@sap.com> Differential Revision: http://reviews.llvm.org/D6013 llvm-svn: 228421	2015-02-06 18:42:41 +00:00
Adam Nemet	7206d7a5d2	[LV] Move addRuntimeCheck to LoopAccessAnalysis This will allow it to be shared with the new Loop Distribution pass. getFirstInst is currently duplicated across LoopVectorize.cpp and LoopAccessAnalysis.cpp. This is a short-term work-around until we figure out a better solution. NFC. (The code moved is adjusted a bit for the name of the Loop member and that PtrRtCheck is now a reference rather than a pointer.) llvm-svn: 228418	2015-02-06 18:31:04 +00:00
Reid Kleckner	526ec29370	Don't dllexport declarations Fixes PR22488 llvm-svn: 228411	2015-02-06 17:59:49 +00:00
Benjamin Kramer	970eac40bf	Make helper functions/classes/globals static. NFC. llvm-svn: 228410	2015-02-06 17:51:54 +00:00
Matthias Braun	2e404597f4	InstCombine: Combine select sequences into a single select Normalize select(C0, select(C1, a, b), b) -> select((C0 & C1), a, b) select(C0, a, select(C1, a, b)) -> select((C0 \| C1), a, b) This normal form may enable further combines on the And/Or and shortens paths for the values. Many targets prefer the other but can go back easily in CodeGen. Differential Revision: http://reviews.llvm.org/D7399 llvm-svn: 228409	2015-02-06 17:49:36 +00:00
Matthias Braun	7dc96dea0a	LiveInterval: Fix SubRange memory leak. llvm-svn: 228405	2015-02-06 17:28:47 +00:00
Daniel Sanders	808141c2af	[mips] Fix FileCheck prefixes with whitespace between 'CHECK' and ':' llvm-svn: 228403	2015-02-06 16:37:30 +00:00
Benjamin Kramer	0b9234baef	Value: Remove superfluous typedefs and deprecated method. NFC. llvm-svn: 228400	2015-02-06 14:44:02 +00:00
Benjamin Kramer	69b4ad29af	AArch64PromoteConstant: Modernize and resolve some Use<->User confusion. NFC. llvm-svn: 228399	2015-02-06 14:43:55 +00:00
Benjamin Kramer	39f76acb5c	IRCE: Demote template to ArrayRef and SmallVector to array. NFC. llvm-svn: 228398	2015-02-06 14:43:49 +00:00
Chad Rosier	92c1f363f4	Whitespace. llvm-svn: 228397	2015-02-06 14:14:41 +00:00
Rafael Espindola	9992d61727	Correcting keyword highlighting in llvm-mode.el. llvm-mode was previously confused when variable names contained keywords. This changes ensures that keywords are only highlighted when they're standalone. Patch by Wilfred Hughes! llvm-svn: 228396	2015-02-06 13:57:58 +00:00
Peter Zotov	7b9b9d8c85	[OCaml] Add Llvm.build_empty_phi. llvm-svn: 228395	2015-02-06 13:42:03 +00:00
Arnaud A. de Grandmaison	77f62d45ef	[PBQP] Fix comment wording. NFC llvm-svn: 228390	2015-02-06 11:28:16 +00:00
Craig Topper	d193763f1c	[X86] Add assembler and disassembler test cases for clflushopt, clwb, pcommit, xsaves, xrstors, xsavec llvm-svn: 228385	2015-02-06 06:19:28 +00:00
Craig Topper	bc7a86b98d	[X86] Remove a ton of duplicate test cases for the assembler. llvm-svn: 228383	2015-02-06 05:50:50 +00:00
Michel Danzer	a786077253	R600/SI: Amend a test to ensure WQM is enabled for LDS in pixel shaders Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228374	2015-02-06 02:51:29 +00:00
Michel Danzer	d89a557f33	R600/SI: Don't enable WQM for V_INTERP_* instructions v2 Doesn't seem necessary anymore. I think this was mostly compensating for not enabling WQM for texture sampling instructions. v2: Add test coverage Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228373	2015-02-06 02:51:25 +00:00
Michel Danzer	494391ba47	R600/SI: Also enable WQM for image opcodes which calculate LOD v3 If whole quad mode isn't enabled for these, the level of detail is calculated incorrectly for pixels along diagonal triangle edges, causing artifacts. v2: Use a TSFlag instead of lots of switch cases v3: Add test coverage Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=88642 Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228372	2015-02-06 02:51:20 +00:00
Ramkumar Ramachandra	8378ac3684	Introduce print-memderefs to test isDereferenceablePointer Since testing the function indirectly is tricky, introduce a direct print-memderefs pass, in the same spirit as print-memdeps, which prints dereferenceability information matched by FileCheck. Differential Revision: http://reviews.llvm.org/D7075 llvm-svn: 228369	2015-02-06 01:46:42 +00:00
Matthias Braun	8ab869010f	AArch64: Make test more robust. Avoid the creation of select instructions which can result in different scheduling of the selects. I also added a bunch of additional store volatiles. Those avoid A CodeGen problem (bug?) where normalizes and denomarlizing the control moves all shift instructions into the first block where ISel can't match them together with the cmps. llvm-svn: 228362	2015-02-05 23:52:14 +00:00
Matthias Braun	5cfba2f573	X86: Test cleanup Use FileCheck, make it more consistent and do not rely on unoptimized or(cmp,cmp) getting combined for max to be matched. llvm-svn: 228361	2015-02-05 23:52:12 +00:00
Daniel Jasper	4bb224decb	Small cleanup of MachineLICM.cpp Specifically: - Calculate the loop pre-header once at the stat of HoistOutOfLoop, so: - We don't-DFS walk the MachineDomTree if we aren't going to do anything - Don't call getCurPreheader for each Scope - Don't needlessly use a do-while loop - Use early exit for Scopes.size() == 0 No functional changes intended. llvm-svn: 228350	2015-02-05 22:39:46 +00:00
Colin LeMahieu	6e3e62fd13	[Hexagon] Renaming v4 compare-and-jump instructions. llvm-svn: 228349	2015-02-05 22:03:32 +00:00
Colin LeMahieu	6b3aac8ad9	[Hexagon] Deleting unused patterns. llvm-svn: 228348	2015-02-05 21:43:56 +00:00
Colin LeMahieu	de68b66b4d	[Hexagon] Simplifying and formatting several patterns. Changing a pattern multiply to be expanded. llvm-svn: 228347	2015-02-05 21:13:25 +00:00
Ahmed Bougacha	a7e33112f0	[BasicAA] Add datalayouts to make some tests more useful. NFC. Fixes PR22462: two of the tests have regressed for a while, but were using CHECK-NOT to match "May:". The actual output was changed to "MayAlias:" at some point, which made the tests useless. Two others return MayAlias only because of a lack of analysis; BasicAA returns PartialAlias in those cases, when a datalayout is present. llvm-svn: 228346	2015-02-05 21:10:14 +00:00
Colin LeMahieu	99c5ce1ce4	[Hexagon] Factoring a class out of some store patterns, deleting unused definitions and reformatting some patterns. llvm-svn: 228345	2015-02-05 20:38:58 +00:00
Colin LeMahieu	d40bb5353d	[Hexagon] Factoring out a class for immediate transfers and cleaning up formatting. llvm-svn: 228343	2015-02-05 20:08:52 +00:00
Justin Bogner	de3dec5d16	InstrProf: Avoid using std::to_string Apparently std::to_string doesn't exist in mingw32: http://lab.llvm.org:8011/builders/clang-native-mingw32-win7/builds/7990 https://gcc.gnu.org/bugzilla/show_bug.cgi?id=52015 llvm-svn: 228340	2015-02-05 19:54:27 +00:00
Alexey Samsonov	19763c48df	[ASan] Enable -asan-stack-dynamic-alloca by default. By default, store all local variables in dynamic alloca instead of static one. It reduces the stack space usage in use-after-return mode (dynamic alloca will not be called if the local variables are stored in a fake stack), and improves the debug info quality for local variables (they will not be described relatively to %rbp/%rsp, which are assumed to be clobbered by function calls). llvm-svn: 228336	2015-02-05 19:39:20 +00:00
Eric Christopher	24f3f65196	Remove the use of getSubtarget in the creation of the X86 PassManager instance. In one case we can make the determination from the Triple, in the other (execution dependency pass) the pass will avoid running if we don't have any code that uses that register class so go ahead and add it to the pipeline. llvm-svn: 228334	2015-02-05 19:27:04 +00:00
Eric Christopher	d361ff8282	Use cached subtargets inside X86FixupLEAs. llvm-svn: 228333	2015-02-05 19:27:01 +00:00
Eric Christopher	d7dec666cc	Migrate the X86 AsmPrinter away from using the subtarget when dealing with module level emission. Currently this is using the Triple to determine, but eventually the logic should probably migrate to TLOF. llvm-svn: 228332	2015-02-05 19:06:45 +00:00
Sylvestre Ledru	9be0b77f3f	Fix an incorrect identifier Summary: EIEIO is not a correct declaration and breaks the build under Debian HURD. Instead, E_IEIO is used. // http://www.gnu.org/software/libc/manual/html_node/Reserved-Names.html Some additional classes of identifier names are reserved for future extensions to the C language or the POSIX.1 environment. While using these names for your own purposes right now might not cause a problem, they do raise the possibility of conflict with future versions of the C or POSIX standards, so you should avoid these names. ... Names beginning with a capital ‘E’ followed a digit or uppercase letter may be used for additional error code names. See Error Reporting.// Reported here: https://bugs.debian.org/cgi-bin/bugreport.cgi?bug=776965 And patch wrote by Svante Signell With this patch, LLVM, Clang & LLDB build under Debian HURD: https://buildd.debian.org/status/fetch.php?pkg=llvm-toolchain-3.6&arch=hurd-i386&ver=1%3A3.6~%2Brc2-2&stamp=1423040039 Reviewers: hfinkel Reviewed By: hfinkel Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7437 llvm-svn: 228331	2015-02-05 18:57:02 +00:00
Colin LeMahieu	b882f2b5cf	[Hexagon] Renaming Y2_barrier. Fixing issues where doubleword variants of instructions can't be newvalue producers. llvm-svn: 228330	2015-02-05 18:56:28 +00:00
Hal Finkel	c9dd02066c	[PowerPC] Prepare loops for pre-increment loads/stores PowerPC supports pre-increment load/store instructions (except for Altivec/VSX vector load/stores). Using these on embedded cores can be very important, but most loops are not naturally set up to use them. We can often change that, however, by placing loops into a non-canonical form. Generically, this means transforming loops like this: for (int i = 0; i < n; ++i) array[i] = c; to look like this: T p = array[-1]; for (int i = 0; i < n; ++i) ++p = c; the key point is that addresses accessed are pulled into dedicated PHIs and "pre-decremented" in the loop preheader. This allows the use of pre-increment load/store instructions without loop peeling. A target-specific late IR-level pass (running post-LSR), PPCLoopPreIncPrep, is introduced to perform this transformation. I've used this code out-of-tree for generating code for the PPC A2 for over a year. Somewhat to my surprise, running the test suite + externals on a P7 with this transformation enabled showed no performance regressions, and one speedup: External/SPEC/CINT2006/483.xalancbmk/483.xalancbmk -2.32514% +/- 1.03736% So I'm going to enable it on everything for now. I was surprised by this because, on the POWER cores, these pre-increment load/store instructions are cracked (and, thus, harder to schedule effectively). But seeing no regressions, and feeling that it is generally easier to split instructions apart late than it is to combine them late, this might be the better approach regardless. In the future, we might want to integrate this functionality into LSR (but currently LSR does not create new PHI nodes, so (for that and other reasons) significant work would need to be done). llvm-svn: 228328	2015-02-05 18:43:00 +00:00
Hal Finkel	65d1cbf9df	[PowerPC] Generate pre-increment floating-point ld/st instructions PowerPC supports pre-increment floating-point load/store instructions, both r+r and r+i, and we had patterns for them, but they were not marked as legal. Mark them as legal (and add a test case). llvm-svn: 228327	2015-02-05 18:42:53 +00:00
Colin LeMahieu	27d50073b3	[Hexagon] Renaming A2_subri, A2_andir, A2_orir. Fixing formatting. llvm-svn: 228326	2015-02-05 18:38:08 +00:00
Ahmed Bougacha	e892d13d90	[CodeGen] Add hook/combine to form vector extloads, enabled on X86. The combine that forms extloads used to be disabled on vector types, because "None of the supported targets knows how to perform load and sign extend on vectors in one instruction." That's not entirely true, since at least SSE4.1 X86 knows how to do those sextloads/zextloads (with PMOVS/ZX). But there are several aspects to getting this right. First, vector extloads are controlled by a profitability callback. For instance, on ARM, several instructions have folded extload forms, so it's not always beneficial to create an extload node (and trying to match extloads is a whole 'nother can of worms). The interesting optimization enables folding of s/zextloads to illegal (splittable) vector types, expanding them into smaller legal extloads. It's not ideal (it introduces some legalization-like behavior in the combine) but it's better than the obvious alternative: form illegal extloads, and later try to split them up. If you do that, you might generate extloads that can't be split up, but have a valid ext+load expansion. At vector-op legalization time, it's too late to generate this kind of code, so you end up forced to scalarize. It's better to just avoid creating egregiously illegal nodes. This optimization is enabled unconditionally on X86. Note that the splitting combine is happy with "custom" extloads. As is, this bypasses the actual custom lowering, and just unrolls the extload. But from what I've seen, this is still much better than the current custom lowering, which does some kind of unrolling at the end anyway (see for instance load_sext_4i8_to_4i64 on SSE2, and the added FIXME). Also note that the existing combine that forms extloads is now also enabled on legal vectors. This doesn't have a big effect on X86 (because sext+load is usually combined to sext_inreg+aextload). On ARM it fires on some rare occasions; that's for a separate commit. Differential Revision: http://reviews.llvm.org/D6904 llvm-svn: 228325	2015-02-05 18:31:02 +00:00
Ahmed Bougacha	db1da7a54c	[CodeGen] Add isLoadExtLegalOrCustom helper to TargetLowering. llvm-svn: 228322	2015-02-05 18:15:59 +00:00
Andrew Trick	7fc4583eda	X86 ABI fix for return values > 24 bytes. The return value's address must be returned in %rax. i.e. the callee needs to copy the sret argument (%rdi) into the return value (%rax). This probably won't manifest as a bug when the caller is LLVM-compiled code. But it is an ABI guarantee and tools expect it. llvm-svn: 228321	2015-02-05 18:09:05 +00:00
Colin LeMahieu	f297dbed48	[Hexagon] Renaming A2_addi and formatting. llvm-svn: 228318	2015-02-05 17:49:13 +00:00
Sanjay Patel	67667bcdc4	move fold comments to the corresponding fold; NFC llvm-svn: 228317	2015-02-05 17:33:59 +00:00
Colin LeMahieu	a66cf6f2df	[Hexagon] Since decoding conflicts have been resolved, isCodeGenOnly = 0 by default and remove explicitly setting it. llvm-svn: 228316	2015-02-05 17:32:17 +00:00
Sylvestre Ledru	648cced01c	Identical code for different branches (CID 1254883) Reviewers: kledzik, rafael Reviewed By: rafael Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D6303 llvm-svn: 228313	2015-02-05 17:00:23 +00:00
Hans Wennborg	8b4dbdf15d	LowerSwitch: Use ConstantInt for CaseRange::{Low,High} Case values are always ConstantInt. This allows us to remove a bunch of casts. NFC. llvm-svn: 228312	2015-02-05 16:58:10 +00:00
Hans Wennborg	8c82fbcb73	LowerSwitch: remove default args from CaseRange ctor; NFC llvm-svn: 228311	2015-02-05 16:50:27 +00:00
Sylvestre Ledru	fe0c7ad852	revert 228308. The code has changed since the review llvm-svn: 228309	2015-02-05 16:35:44 +00:00
Sylvestre Ledru	d0ee6daffd	Identical code for different branches (CID 1254883) Reviewers: kledzik, rafael Reviewed By: rafael Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D6303 llvm-svn: 228308	2015-02-05 16:30:25 +00:00
Tom Stellard	eea3f70432	R600/SI: Fix bug in TTI loop unrolling preferences We should be setting UnrollingPreferences::MaxCount to MAX_UINT instead of UnrollingPreferences::Count. Count is a 'forced unrolling factor', while MaxCount sets an upper limit to the unrolling factor. Setting Count to MAX_UINT was causing the loop in the testcase to be unrolled 15 times, when it only had a maximum of 4 iterations. llvm-svn: 228303	2015-02-05 15:32:18 +00:00
Tom Stellard	0f29de78e6	R600/SI: Fix bug from insertion of llvm.SI.end.cf into loop headers The llvm.SI.end.cf intrinsic is used to mark the end of if-then blocks, if-then-else blocks, and loops. It is responsible for updating the exec mask to re-enable threads that had been masked during the preceding control flow block. For example: s_mov_b64 exec, 0x3 ; Initial exec mask s_mov_b64 s[0:1], exec ; Saved exec mask v_cmpx_gt_u32 exec, s[2:3], v0, 0 ; llvm.SI.if do_stuff() s_or_b64 exec, exec, s[0:1] ; llvm.SI.end.cf The bug fixed by this patch was one where the llvm.SI.end.cf intrinsic was being inserted into the header of loops. This would happen when an if block terminated in a loop header and we would end up with code like this: s_mov_b64 exec, 0x3 ; Initial exec mask s_mov_b64 s[0:1], exec ; Saved exec mask v_cmpx_gt_u32 exec, s[2:3], v0, 0 ; llvm.SI.if do_stuff() LOOP: ; Start of loop header s_or_b64 exec, exec, s[0:1] ; llvm.SI.end.cf <-BUG: The exec mask has the same value at the beginning of each loop iteration. do_stuff(); s_cbranch_execnz LOOP The fix is to create a new basic block before the loop and insert the llvm.SI.end.cf there. This way the exec mask is restored before the start of the loop instead of at the beginning of each iteration. llvm-svn: 228302	2015-02-05 15:32:15 +00:00
Bill Schmidt	433b1c3aae	[PowerPC] Implement the vclz instructions for PWR8 Patch by Kit Barton. Add the vector count leading zeros instruction for byte, halfword, word, and doubleword sizes. This is a fairly straightforward addition after the changes made for vpopcnt: 1. Add the correct definitions for the various instructions in PPCInstrAltivec.td 2. Make the CTLZ operation legal on vector types when using P8Altivec in PPCISelLowering.cpp Test Plan Created new test case in test/CodeGen/PowerPC/vec_clz.ll to check the instructions are being generated when the CTLZ operation is used in LLVM. Check the encoding and decoding in test/MC/PowerPC/ppc_encoding_vmx.s and test/Disassembler/PowerPC/ppc_encoding_vmx.txt respectively. llvm-svn: 228301	2015-02-05 15:24:47 +00:00
Rafael Espindola	f8d662aa4b	Add a FIXME. Thanks to Eric for the suggestion. llvm-svn: 228300	2015-02-05 14:57:47 +00:00
Aaron Ballman	94d4d33a38	Removing an unused variable warning I accidentally introduced with my last warning fix; NFC. llvm-svn: 228295	2015-02-05 13:52:42 +00:00
Aaron Ballman	1b072b340b	Silencing an MSVC warning about a switch statement with no cases; NFC. llvm-svn: 228294	2015-02-05 13:40:04 +00:00
Bruno Cardoso Lopes	ab9ae87623	[X86][MMX] Handle i32->mmx conversion using movd Implement a BITCAST dag combine to transform i32->mmx conversion patterns into a X86 specific node (MMX_MOVW2D) and guarantee that moves between i32 and x86mmx are better handled, i.e., don't use store-load to do the conversion.. llvm-svn: 228293	2015-02-05 13:23:07 +00:00
Bruno Cardoso Lopes	cc6089d2e0	[X86][MMX] Add several bitcast tests Avoid regression in previously supported MMX code by adding different combinations of tests which exercise MMX bitcasts. Small improvements to these patterns should come next. llvm-svn: 228292	2015-02-05 13:22:57 +00:00
Bruno Cardoso Lopes	e446aefcfe	[X86][MMX] Move MMX DAG node to proper file llvm-svn: 228291	2015-02-05 13:22:50 +00:00
Michael Kuperstein	d2b6fdbc31	Teach isDereferenceablePointer() to look through bitcast constant expressions. This fixes a LICM regression due to the new load+store pair canonicalization. Differential Revision: http://reviews.llvm.org/D7411 llvm-svn: 228284	2015-02-05 09:15:37 +00:00
Craig Topper	0857a1da0c	[X86] Add xrstors/xsavec/xsaves/clflushopt/clwb/pcommit instructions llvm-svn: 228283	2015-02-05 08:51:06 +00:00
Craig Topper	a898c2d737	[X86] Remove two feature flags that covered sets of instructions that have no patterns or intrinsics. Since we don't check feature flags in the assembler parser for any instruction sets, these flags don't provide any value. This frees up 2 of the fully utilized feature flags. llvm-svn: 228282	2015-02-05 08:51:02 +00:00
Matt Arsenault	abd271b4e8	R600/SI: Fix i64 truncate to i1 llvm-svn: 228273	2015-02-05 06:05:13 +00:00
Larisse Voufo	bc9f12e7bc	Disable enumeral mismatch warning when compiling llvm with gcc. Tested with gcc 4.9.2. Compiling with -Werror was producing: .../llvm/lib/Target/X86/X86ISelLowering.cpp: In function 'llvm::SDValue lowerVectorShuffleAsBitMask(llvm::SDLoc, llvm::MVT, llvm::SDValue, llvm::SDValue, llvm::ArrayRef<int>, llvm::SelectionDAG&)': .../llvm/lib/Target/X86/X86ISelLowering.cpp:7771:40: error: enumeral mismatch in conditional expression: 'llvm::X86ISD::NodeType' vs 'llvm::ISD::NodeType' [-Werror=enum-compare] V = DAG.getNode(VT.isFloatingPoint() ? X86ISD::FAND : ISD::AND, DL, VT, V, ^ llvm-svn: 228271	2015-02-05 04:54:51 +00:00
Matt Arsenault	f28cf0cbaf	Add addrspacecast node to tablegen The node is still defined oddly so that the address spaces are not operands and not accessible from tablegen, but as-is this can now be used to write a ComplexPattern with an addrspacecast root node. llvm-svn: 228270	2015-02-05 03:35:34 +00:00
Matt Arsenault	d931642cc7	Add support for double / float to EndianStream Also add new unit tests for endian::Writer llvm-svn: 228269	2015-02-05 03:30:08 +00:00
Michael Zolotukhin	a9aadd2903	Implement new heuristic for complete loop unrolling. Complete loop unrolling can make some loads constant, thus enabling a lot of other optimizations. To catch such cases, we look for loads that might become constants and estimate number of instructions that would be simplified or become dead after substitution. Example: Suppose we have: int a[] = {0, 1, 0}; v = 0; for (i = 0; i < 3; i ++) v += b[i]a[i]; If we completely unroll the loop, we would get: v = b[0]a[0] + b[1]a[1] + b[2]a[2] Which then will be simplified to: v = b[0]* 0 + b[1]* 1 + b[2]* 0 And finally: v = b[1] llvm-svn: 228265	2015-02-05 02:34:00 +00:00
Cameron Esfahani	17177d1e84	Value soft float calls as more expensive in the inliner. Summary: When evaluating floating point instructions in the inliner, ask the TTI whether it is an expensive operation. By default, it's not an expensive operation. This keeps the default behavior the same as before. The ARM TTI has been updated to return back TCC_Expensive for targets which don't have hardware floating point. Reviewers: chandlerc, echristo Reviewed By: echristo Subscribers: t.p.northover, aemerson, llvm-commits Differential Revision: http://reviews.llvm.org/D6936 llvm-svn: 228263	2015-02-05 02:09:33 +00:00
Ahmed Bougacha	c7f241cba9	[ARM] Use patterns instead of hardcoded regs in test. NFC. llvm-svn: 228259	2015-02-05 01:52:19 +00:00
Ahmed Bougacha	daaf3d9b58	[ARM] Make testcase more explicit. NFC. The q8/d16 thing is silly; I'd be happy to hear about a better way to write those tests where simple substitution isn't enough.. llvm-svn: 228258	2015-02-05 01:45:28 +00:00
Reid Kleckner	9ccec06623	Try to fix the build in MCValue.cpp llvm-svn: 228256	2015-02-05 01:23:14 +00:00
Sean Silva	32f24c49d2	Fixup. Didn't see these calls in my release build locally when testing. llvm-svn: 228254	2015-02-05 01:13:47 +00:00
Duncan P. N. Exon Smith	252b62d3c2	IR: Split out getOperandAs(), NFC llvm-svn: 228250	2015-02-05 01:07:47 +00:00
Sean Silva	0e1fe184c8	[MC] Remove various unused MCAsmInfo parameters. llvm-svn: 228244	2015-02-05 00:58:51 +00:00
Duncan P. N. Exon Smith	9c26d80c18	IR: Rename 'operator ==()' to 'isKeyOf()', NFC `isKeyOf()` is a clearer name than overloading `operator==()`. llvm-svn: 228242	2015-02-05 00:51:35 +00:00
Duncan P. N. Exon Smith	5a914a8c63	ADT: Add int64_t interoperability to APSInt Add some API to `APSInt` to make it easier to compare with `int64_t`. - `APSInt::compareValues(APSInt, APSInt)` returns 1, -1 or 0 for greater, lesser, or equal, doing the right thing for mismatched "has-sign" and bitwidths. This is just like `isSameValue()` (and is now the implementation of it). - `APSInt::get(int64_t)` gets a signed `APSInt`. - `operator<(int64_t)`, etc., are implemented trivially via `get()` and `compareValues()`. - Also added `APSInt::getUnsigned(uint64_t)` to make it easier to test `compareValues()`. llvm-svn: 228239	2015-02-05 00:17:43 +00:00
Colin LeMahieu	ceebe8659b	[Hexagon] Deleting unused instructions and adding isCodeGenOnly to some defs. llvm-svn: 228238	2015-02-05 00:10:16 +00:00
Colin LeMahieu	9cb9078ccf	[Hexagon] Updating load extend to i64 patterns. llvm-svn: 228237	2015-02-04 23:55:16 +00:00
Kostya Serebryany	92e0476c67	[fuzzer] add flag prefer_small_during_initial_shuffle, be a bit more verbose llvm-svn: 228235	2015-02-04 23:42:42 +00:00
Colin LeMahieu	712d5c393b	[Hexagon] Cleaning up i1 load and extension patterns. llvm-svn: 228232	2015-02-04 23:27:48 +00:00
Colin LeMahieu	90a91bbf43	[Hexagon] Simplifying more load and store patterns and using new addressing patterns. llvm-svn: 228231	2015-02-04 23:23:16 +00:00
Reid Kleckner	2c1990778d	Remove useless call to isOSCygMing() This used to do something when we modeled the Cygwin and MinGW environments as distinct OSs, but now it is not needed. llvm-svn: 228229	2015-02-04 23:17:19 +00:00
Tom Stellard	f6afc80cc0	R600/SI: Enable subreg liveness by default llvm-svn: 228228	2015-02-04 23:14:18 +00:00
Colin LeMahieu	ad13d4e8a6	[Hexagon] Simplifying some load and store patterns. llvm-svn: 228227	2015-02-04 23:10:21 +00:00
Duncan P. N. Exon Smith	9c93dc6453	AsmParser: Split out LineField, NFC Split out `LineField`, which restricts the legal line numbers. This will make it easier to be consistent between different node parsers. llvm-svn: 228226	2015-02-04 22:59:18 +00:00
Colin LeMahieu	68292c96da	[Hexagon] Converting absolute-address load patterns to use AddrGP. llvm-svn: 228225	2015-02-04 22:54:51 +00:00
Colin LeMahieu	8bf5de10c3	[Hexagon] Converting atomic store/load to use AddrGP addressing. llvm-svn: 228223	2015-02-04 22:40:36 +00:00
Reid Kleckner	ba77ad75d3	Don't warn or note if bash is missing We haven't needed bash on Windows to run the test suite for a long time now. Patch by Michael Edwards! llvm-svn: 228221	2015-02-04 22:36:52 +00:00
Colin LeMahieu	5149135369	[Hexagon] Simplifying some store patterns. Adding AddrGP addressing forms. llvm-svn: 228220	2015-02-04 22:36:28 +00:00
Filipe Cabecinhas	b36685ee85	Handle LLVM_USE_SANITIZER=Address;Undefined (and the other way around) Summary: Handle LLVM_USE_SANITIZER=Address;Undefined to enable ASan and UBSan If UBSan is compatible with more of the other sanitizers, maybe we should deal with this in a better way where we allow combining UBSan with any of the other sanitizers. Reviewers: samsonov Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7024 llvm-svn: 228219	2015-02-04 22:33:31 +00:00
Kostya Serebryany	33f866922a	[fuzzer] add -runs=N to limit the number of runs per session. Also, make sure we do some mutations w/o cross over. llvm-svn: 228214	2015-02-04 22:20:09 +00:00
Duncan P. N. Exon Smith	2f09c46f80	Fix GCC error caused by r228211 llvm-svn: 228213	2015-02-04 22:13:28 +00:00
Duncan P. N. Exon Smith	8af6cfcdb2	IR: Reduce boilerplate in DenseMapInfo overrides, NFC Minimize the boilerplate required for the `MDNode` subclass `DenseMapInfo<>` overrides in `LLVMContextImpl`. llvm-svn: 228212	2015-02-04 22:08:30 +00:00
Duncan P. N. Exon Smith	077c0318dc	AsmParser: Move MDField details to source file, NFC Move all the types of `MDField` to an anonymous namespace in the source file. This also eliminates the duplication of `ParseMDField()` declarations in the header for each new field type. llvm-svn: 228211	2015-02-04 22:05:21 +00:00
Duncan P. N. Exon Smith	fad631a90c	AsmParser: Simplify assertion, NFC llvm-svn: 228209	2015-02-04 22:02:18 +00:00
Duncan P. N. Exon Smith	7a968781a6	AsmParser: Remove dead code, NFC This condition is checked in the generic `ParseMDField()`. llvm-svn: 228208	2015-02-04 22:00:59 +00:00
Duncan P. N. Exon Smith	39b10c2cbb	AsmParser: Simplify MDUnsignedField We only need `uint64_t` for storage. llvm-svn: 228205	2015-02-04 21:57:52 +00:00
Duncan P. N. Exon Smith	6b7b680efd	IR: Initialize MDNode abbreviations en masse, NFC llvm-svn: 228203	2015-02-04 21:54:12 +00:00
Duncan P. N. Exon Smith	104e402610	IR: Define MDNode uniquing sets automatically, NFC llvm-svn: 228200	2015-02-04 21:46:12 +00:00
Kevin Enderby	10ba041188	Add code to llvm-objdump so the -section option with -macho will dump ‘C’ string sections with the Mach-O S_CSTRING_LITERALS section type. llvm-svn: 228198	2015-02-04 21:38:42 +00:00
Rafael Espindola	a092f17580	Don' try to make sections in comdats SHF_MERGE. Parts of llvm were not expecting it and we wouldn't print the entity size of the section. Given what comdats are used for, having SHF_MERGE sections would be just a small improvement, so just disable it for now. Fixes pr22463. llvm-svn: 228196	2015-02-04 21:27:24 +00:00
Sean Silva	b6472fe3da	[docs] Put an explicit link to InAlloca.rst llvm-svn: 228192	2015-02-04 20:51:19 +00:00
Tom Stellard	33e64c66ac	R600/SI: Expand misaligned 16-bit memory accesses llvm-svn: 228190	2015-02-04 20:49:52 +00:00
Tom Stellard	c7e448c92e	R600/SI: Make more store operations legal v2i32, i32, trunc i32 to i16, and truc i32 to i8 stores are legal for all address spaces. We had marked them as custom in order to lower them for the private address space, but this is no longer necessary. This enables lowering of misaligned stores of these types in the DAGLegalizer. llvm-svn: 228189	2015-02-04 20:49:51 +00:00
Tom Stellard	096b8c1e6d	R600: Don't promote i64 stores to v2i32 during DAG legalization We take care of this during instruction selection now. This fixes a potential infinite loop when lowering misaligned stores. llvm-svn: 228188	2015-02-04 20:49:49 +00:00
Tom Stellard	080209d573	StructurizeCFG: Remove obsolete fix for loop backedge detection This is no longer needed now that we are using a reverse post-order traversal. llvm-svn: 228187	2015-02-04 20:49:47 +00:00
Tom Stellard	071ec90b68	StructurizeCFG: Use a reverse post-order traversal We were previously doing a post-order traversal and operating on the list in reverse, however this would occasionaly cause backedges for loops to be visited before some of the other blocks in the loop. We know use a reverse post-order traversal, which avoids this issue. The reverse post-order traversal is not completely ideal, so we need to manually fixup the list to ensure that inner loop backedges are visited before outer loop backedges. llvm-svn: 228186	2015-02-04 20:49:44 +00:00
Colin LeMahieu	987b0943c8	[Hexagon] Adding selection for GlobalAddress and converting [z/i]ext load patterns to make use of them. llvm-svn: 228184	2015-02-04 20:38:01 +00:00
Bill Schmidt	caf7e8b147	Add missing test case from r228046 llvm-svn: 228182	2015-02-04 20:00:04 +00:00
Duncan P. N. Exon Smith	920df5c1bb	Utils: Resolve cycles under distinct MDNodes Track unresolved nodes under distinct `MDNode`s during `MapMetadata()`, and resolve them at the end. Previously, these cycles wouldn't get resolved. llvm-svn: 228180	2015-02-04 19:44:34 +00:00
Matthias Braun	26e7ea6267	MachineCSE: Clear dead-def flag on CSE. In case CSE reuses a previoulsy unused register the dead-def flag has to be cleared on the def operand, as exposed by the arm64-cse.ll test. This fixes PR22439 and the corresponding rdar://19694987 Differential Revision: http://reviews.llvm.org/D7395 llvm-svn: 228178	2015-02-04 19:35:16 +00:00
Reid Kleckner	c26a17a822	Add range adapters predecessors() and successors() for BBs Use them in two isolated transforms so we know they work and aren't dead code. llvm-svn: 228173	2015-02-04 19:14:57 +00:00
Kostya Serebryany	5b266a8a23	[fuzzer] make multi-process execution more verbose; fix mutation to actually respect mutation depth and to never produce empty units llvm-svn: 228170	2015-02-04 19:10:20 +00:00
Colin LeMahieu	86abe35ceb	[Hexagon] Replacing some load patterns with cleaner versions. llvm-svn: 228169	2015-02-04 19:05:32 +00:00
Michael Kuperstein	cd63c5fa73	Fixes a bug in vector load legalization that confused bits and bytes. Differential Revision: http://reviews.llvm.org/D7400 llvm-svn: 228168	2015-02-04 18:54:01 +00:00
Ismail Donmez	9559232f1d	Revert test commit llvm-svn: 228167	2015-02-04 18:46:00 +00:00
Ismail Donmez	3b25ef3831	Test commit llvm-svn: 228166	2015-02-04 18:45:43 +00:00
Juergen Ributzka	719615f6dd	Add missing include. llvm-svn: 228161	2015-02-04 18:16:53 +00:00
Colin LeMahieu	f856dcb75e	[Hexagon] Adding missing isCodeGenOnly = 0 llvm-svn: 228160	2015-02-04 18:11:32 +00:00
Colin LeMahieu	c0434466e4	[Hexagon] Adding encoding information for absolute-reg mode stores. Xfailing a test until constant extenders are correctly put in the same packet. llvm-svn: 228158	2015-02-04 17:52:06 +00:00
Alexey Samsonov	b9b8027cee	SpecialCaseList: Add support for parsing multiple input files. Summary: This change allows users to create SpecialCaseList objects from multiple local files. This is needed to implement a proper support for -fsanitize-blacklist flag (allow users to specify multiple blacklists, in addition to default blacklist, see PR22431). DFSan can also benefit from this change, as DFSan instrumentation pass now accepts ABI-lists both from -fsanitize-blacklist= and -mllvm -dfsan-abilist flags. Go bindings are fixed accordingly. Test Plan: regression test suite Reviewers: pcc Subscribers: llvm-commits, axw, kcc Differential Revision: http://reviews.llvm.org/D7367 llvm-svn: 228155	2015-02-04 17:39:48 +00:00
Colin LeMahieu	7d971056ed	[Hexagon] Adding encoding information for absolute-set stores. llvm-svn: 228154	2015-02-04 17:24:04 +00:00
Colin LeMahieu	0eb9727d42	[Hexagon] Adding encoding bits for indirect long load instructions. llvm-svn: 228152	2015-02-04 16:56:46 +00:00
Bradley Smith	9f4cd59e80	[ARM] Fix subtarget feature set truncation when using .cpu directive This is a bug that was caused due to storing the feature bitset in a 32-bit variable when it is a 64-bit mask, discarding the top half of the feature set. llvm-svn: 228151	2015-02-04 16:23:24 +00:00
Zoran Jovanovic	5a1a780c2a	[mips][microMIPS] Implement CodeGen support for SW16 and LW16 instructions Differential Revision: http://reviews.llvm.org/D6581 llvm-svn: 228149	2015-02-04 15:43:17 +00:00
Daniel Sanders	e67d27f5cc	[mips] Make MipsSubtarget::hasMips*() functions consistent. NFC. Reviewers: vmedic Reviewed By: vmedic Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7377 llvm-svn: 228147	2015-02-04 15:18:11 +00:00
Daniel Sanders	a9aab74304	[mips] Remove unused check prefix from tests. NFC. Reviewers: vmedic Reviewed By: vmedic Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7376 llvm-svn: 228145	2015-02-04 14:48:39 +00:00
Aaron Ballman	34c325e749	Fixing a -Wsign-compare warning; NFC llvm-svn: 228142	2015-02-04 14:01:08 +00:00
Renato Golin	6088504499	Adding support to LLVM for targeting Cortex-A72 Currently, Cortex-A72 is modelled as an Cortex-A57 except the fp load balancing pass isn't enabled for Cortex-A72 as it's not profitable to have it enabled for this core. Patch by Ranjeet Singh. llvm-svn: 228140	2015-02-04 13:31:29 +00:00
Rafael Espindola	a5eb775c4d	Fix warning: "function declaration isn’t a prototype" llvm-svn: 228139	2015-02-04 13:30:28 +00:00
Justin Bogner	1f92ce67c5	InstrProf: std::to_string needs to #include <string> llvm-svn: 228136	2015-02-04 11:19:16 +00:00
Chandler Carruth	4d31f58c88	[x86] Give movss and movsd execution domains in the x86 backend. This associates movss and movsd with the packed single and packed double execution domains (resp.). While this is largely cosmetic, as we now don't have weird ping-pong-ing between single and double precision, it is also useful because it avoids the domain fixing algorithm from seeing domain breaks that don't actually exist. It will also be much more important if we have an execution domain default other than packed single, as that would cause us to mix movss and movsd with integer vector code on a regular basis, a very bad mixture. llvm-svn: 228135	2015-02-04 10:58:53 +00:00
Chandler Carruth	78c8dcd9d3	[x86] Remove a low-value test that was just checking how we cleared a register. We have lots of tests covering this. llvm-svn: 228133	2015-02-04 10:47:34 +00:00
Chandler Carruth	bb525e336b	[x86] Mechanically update a bunch of tests' check lines using the latest version of the script. Changes include: - Using the VEX prefix - Skipping more detail when we have useful shuffle comments to match - Matching more shuffle comments that have been added to the printer (yay!) - Matching the destination registers of some AVX instructions - Stripping trailing whitespace that crept in - Fixing indentation issues Nothing interesting going on here. I'm just trying really hard to ensure these changes don't show up in the diffs with actual changes to the backend. llvm-svn: 228132	2015-02-04 10:46:53 +00:00
Chandler Carruth	e375095392	[x86] Teach the test update script to strip trailing whitespace. This is done in a bit of a strange way to use a multiline RE instead of looping over the lines. Suggestions welcome here for a more pythonic way of doing this as long as its reasonably fast. llvm-svn: 228131	2015-02-04 10:46:48 +00:00
Renato Golin	2a5c0a51ce	Reverting VLD1/VST1 base-updating/post-incrementing combining This reverts patches 223862, 224198, 224203, and 224754, which were all related to the vector load/store combining and were reverted/reaplied a few times due to the same alignment problems we're seeing now. Further tests, mainly self-hosting Clang, will be needed to reapply this patch in the future. llvm-svn: 228129	2015-02-04 10:11:59 +00:00
Chandler Carruth	22b1525ae8	[x86] Include the destination register in the check-lines for AVX instructions. No actual change here. llvm-svn: 228127	2015-02-04 09:18:27 +00:00
Chandler Carruth	18ba596609	[x86] Add some tests I missed in the prior commit to cover blends with zero for v8i16 as well. These exhibit the same domain badness, but also exhibit other weaknesses in our blend lowering. More fixes to come. llvm-svn: 228126	2015-02-04 09:15:46 +00:00
Chandler Carruth	024cf8efd7	[x86] Start to introduce bit-masking based blend lowering. This is the simplest form of bit-math based blending which only fires when we are blending with zero and is relatively profitable. I've only enabled this path on very specific lowering strategies. I'm planning to widen its applicability in subsequent patches, but so far you'll notice that even though we get fewer shufps instructions, we still do the bit math in the FP execution port. I'm looking into why this is still happening. llvm-svn: 228124	2015-02-04 09:06:05 +00:00
Chandler Carruth	f4a1c33c7c	[x86] Add missing patterns for andps, orps, xorps, and andnps. Specifically, the existing patterns were scalar-only. These cover the packed vector bitwise operations when specifically requested with pseudo instructions. This is particularly important in SSE1 where we can't actually emit a logical operation on a v2i64 as that isn't a legal type. This will be tested in subsequent patches which form the floating point and patterns in more places. llvm-svn: 228123	2015-02-04 09:06:01 +00:00
Chandler Carruth	872d80e7a4	[x86] Add tests for blends-with-zero on 4-element vectors. llvm-svn: 228122	2015-02-04 09:05:58 +00:00
Bill Schmidt	81638eabe6	Replace tabs with spaces from r228116. Oops. llvm-svn: 228117	2015-02-04 06:14:38 +00:00
Bill Schmidt	1354f7c5fa	[PowerPC] Handle 32-bit targets properly in PPCTLSDynamicCall.cpp llvm-svn: 228116	2015-02-04 05:51:56 +00:00
Philip Reames	72634d6af0	Fix a warning in non-asserts builds llvm-svn: 228114	2015-02-04 05:11:20 +00:00
Frederic Riss	b61f01f1c2	Fix some unnoticed/unwanted behavior change from r222319. The ARM assembler allows register alias redefinitions as long as it targets the same register. r222319 broke that. In the AArch64 case it would just produce a new warning, but in the ARM case it would error out on previously accepted assembler. llvm-svn: 228109	2015-02-04 03:10:03 +00:00
Kostya Serebryany	fe43aa8d19	[fuzzer]: fix exit code, add more diagnostics llvm-svn: 228103	2015-02-04 01:22:57 +00:00
Kostya Serebryany	77cc729ad7	[sanitizer] add another workaround for PR 17409: when over a threshold emit coverage instrumentation as calls. llvm-svn: 228102	2015-02-04 01:21:45 +00:00
Kevin Enderby	95df54c819	Add code to llvm-objdump so the -section option with -macho will disassemble sections that have attributes indicating they contain instructions. llvm-svn: 228101	2015-02-04 01:01:38 +00:00
Chandler Carruth	abd09a1f35	[x86] Refresh the checks of a number of tests using update_llc_test_checks.py. The exact format of the checks has changed over time. This includes different indenting rules, new shuffle comments that have been added, and more operand hiding behind regular expressions. No functional change to the tests are expected here, but this will make subsequent patches have a clean diff as they change shuffle lowering. llvm-svn: 228097	2015-02-04 00:58:42 +00:00
Chandler Carruth	abde67eb1c	[x86] Switch to using the long '--check-prefix' form which the update_llc_test_checks.py script uses, and refresh the checks in this test. No functionality changed here, just bringing this test up to work with automated updates using the python script. llvm-svn: 228096	2015-02-04 00:58:40 +00:00
Chandler Carruth	52332dc620	[x86] Port this test to use utils/update_llc_test_checks.py. This will make it easy to update as I change some parts of the X86 backend, makes it more clear what instruction differences are introduced, and I find it makes it a bit easier to read as well. llvm-svn: 228095	2015-02-04 00:58:37 +00:00
Peter Collingbourne	69ba0167b3	Misc documentation/comment fixes. llvm-svn: 228093	2015-02-04 00:42:45 +00:00
Philip Reames	5a9685dba6	Clang format of a file introduced in 228090 (NFC) llvm-svn: 228091	2015-02-04 00:39:57 +00:00
Philip Reames	47cc673e1f	Add a pass for inserting safepoints into (nearly) arbitrary IR This pass is responsible for figuring out where to place call safepoints and safepoint polls. It doesn't actually make the relocations explicit; that's the job of the RewriteStatepointsForGC pass (http://reviews.llvm.org/D6975). Note that this code is not yet finalized. Its moving in tree for incremental development, but further cleanup is needed and will happen over the next few days. It is not yet part of the standard pass order. Planned changes in the near future: - I plan on restructuring the statepoint rewrite to use the functions add to the IRBuilder a while back. - In the current pass, the function "gc.safepoint_poll" is treated specially but is not an intrinsic. I plan to make identifying the poll function a property of the GCStrategy at some point in the near future. - As follow on patches, I will be separating a collection of test cases we have out of tree and submitting them upstream. - It's not explicit in the code, but these two patches are introducing a new state for a statepoint which looks a lot like a patchpoint. There's no a transient form which doesn't yet have the relocations explicitly represented, but does prevent reordering of memory operations. Once this is in, I need to update actually make this explicit by reserving the 'unused' argument of the statepoint as a flag, updating the docs, and making the code explicitly check for such a thing. This wasn't really planned, but once I split the two passes - which was done for other reasons - the intermediate state fell out. Just reminds us once again that we need to merge statepoints and patchpoints at some point in the not that distant future. Future directions planned: - Identifying more cases where a backedge safepoint isn't required to ensure timely execution of a safepoint poll. - Tweaking the insertion process to generate easier to optimize IR. (For example, investigating making SplitBackedge) the default. - Adding opt-in flags for a GCStrategy to use this pass. Once done, add this pass to the actual pass ordering. Differential Revision: http://reviews.llvm.org/D6981 llvm-svn: 228090	2015-02-04 00:37:33 +00:00
Sanjay Patel	b82b8d6b84	improved CHECK llvm-svn: 228086	2015-02-04 00:24:06 +00:00
Galina Kistanova	d9b46a187f	Added missing header for the explicit dependency on MDNode. llvm-svn: 228085	2015-02-04 00:20:52 +00:00
Justin Bogner	0cca70a6e5	InstrProf: Add some unit tests for CoverageMapping The llvm-level tests for coverage mapping need a binary input file, which means they're hard to understand, hard to update, and it's difficult to add new ones. By adding some unit tests that build up the coverage data structures in C++, we can write more meaningful and targeted tests. llvm-svn: 228084	2015-02-04 00:15:12 +00:00
Justin Bogner	70e0c09e6c	InstrProf: Use a stable sort when reading coverage regions Keeping regions that start at the same location in insertion order makes this logic easier to test / more deterministic. llvm-svn: 228083	2015-02-04 00:12:18 +00:00
Colin LeMahieu	585316cb41	[Hexagon] Revert change to isCodeGenOnly = 1 in r228080 llvm-svn: 228082	2015-02-04 00:09:23 +00:00
Colin LeMahieu	510ba0c661	[Hexagon] Changing some isCodeGenOnly to isAsmParserOnly since we want them to asm parse but not cause decode conflicts. llvm-svn: 228080	2015-02-04 00:07:26 +00:00
Owen Anderson	21b1788ad0	Remove a gross usage of environment variables in MachineVerifier, replacing it with support for setting the -verify-machineinstrs flag via an environment variable in LIT. This preserves the handy functionality of force-enabling the MachineVerifier, without the need to embed usage of environment variables in LLVM client applications. llvm-svn: 228079	2015-02-04 00:02:59 +00:00
Justin Bogner	26b3142d34	InstrProf: Make CounterMappingRegions less confusing to construct Creating empty and expansion regions is awkward with the current API. Expose static methods to make this simpler. llvm-svn: 228075	2015-02-03 23:59:33 +00:00
Arnaud A. de Grandmaison	10797c5707	[PBQP] Provide more information in the debug prints Based on a patch by Jonas Paulsson llvm-svn: 228068	2015-02-03 23:40:24 +00:00
Philip Reames	0285c74261	Use ImmutableCallSite for statepoint verification. Patch by: Igor Laevsky "This change generalizes statepoint verification to use ImmutableCallSite instead of CallInst. This will allow to easily implement invoke statepoint verification (in a following change)." Differential Revision: http://reviews.llvm.org/D7308 llvm-svn: 228064	2015-02-03 23:18:47 +00:00
Adam Nemet	5add5d9d85	[LV] Split off memcheck block really at the first check I've noticed this while trying to move addRuntimeCheck to LoopAccessAnalysis. I think that the intention was to early exit from the overflow checking before the code for the memchecks. This is the entire reason why we compute FirstCheckInst but then we don't use that as the splitting instruction but the final check. Looks like an oversight. llvm-svn: 228056	2015-02-03 22:45:39 +00:00
Chandler Carruth	68fcb38328	[x86] Fix signed vs. unsigned comparison. llvm-svn: 228055	2015-02-03 22:43:30 +00:00
Simon Pilgrim	9eecb14bd9	Fixed unused variable warning. llvm-svn: 228054	2015-02-03 22:39:28 +00:00
Colin LeMahieu	e4101e2c9e	[Hexagon] Marking a bunch of non-encoded instructions with isCodeGenOnly = 1. llvm-svn: 228050	2015-02-03 22:09:51 +00:00

... 2 3 4 5 6 ...

113048 Commits