llvm-project

Commit Graph

Author	SHA1	Message	Date
Zachary Turner	3e76643a95	Change Path::filename_pos() to skip the drive letter. For Windows, filename_pos() tries to find the filename by searching for separators after the last :. Instead, it should really check for the only location that a : is valid, which is in the second character, and search for separators after that. llvm-svn: 228874	2015-02-11 21:16:35 +00:00
Rafael Espindola	d966522377	Remove unused argument. NFC. llvm-svn: 228873	2015-02-11 21:08:00 +00:00
Mehdi Amini	9730116bd6	Reassociate: cannot negate a INT_MIN value Summary: When trying to canonicalize negative constants out of multiplication expressions, we need to check that the constant is not INT_MIN which cannot be negated. Reviewers: mcrosier Reviewed By: mcrosier Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7286 From: Mehdi Amini <mehdi.amini@apple.com> llvm-svn: 228872	2015-02-11 19:54:44 +00:00
Tom Stellard	0648588e7d	R600/SI: Disable subreg liveness This is temporary while we try to fix a crash in the register coalescer. llvm-svn: 228861	2015-02-11 18:24:53 +00:00
Adrian Prantl	18a25b016e	Allow DIBuilder::replaceVTableHolder() to work with temporary nodes, tested via the clang test CodeGenCXX/vtable-holder-self-reference.cpp . llvm-svn: 228854	2015-02-11 17:45:10 +00:00
Adrian Prantl	9a8049238e	Add a trackIfUnresolved to DIBuilder::createInheritance(), tested via the clang test CodeGenCXX/vtable-holder-self-reference.cpp . llvm-svn: 228853	2015-02-11 17:45:08 +00:00
Adrian Prantl	534a81a9ec	Generalize DIBuilder's createReplaceableForwardDecl() to a more flexible createReplaceableCompositeType() that allows to create non-forward-declared temporary nodes. Paired commit with CFE. llvm-svn: 228852	2015-02-11 17:45:05 +00:00
Tom Stellard	de5b7b180a	R600: Split AMDGPUPassConfig into R600PassConfig and GCNPassConfig llvm-svn: 228850	2015-02-11 17:11:51 +00:00
Tom Stellard	c65b36061a	R600: Create an R600TargetMachine for pre-gcn GPUs No functinality change. R600TargetMachine inherits from AMDGPUTargetMachine. llvm-svn: 228849	2015-02-11 17:11:50 +00:00
Jonas Paulsson	bf8d0cc699	Fix SelectionDAG compile time issue with alias analysis. Add new token factor node and its users to worklist if alias analysis is turned on, in DAGCombiner::visitTokenFactor(). Alias analysis may cause a lot of new token factors to be inserted into the DAG, and they need to be optimized to avoid significant slow-downs. Reviewed by Hal Finkel. llvm-svn: 228841	2015-02-11 16:10:31 +00:00
Rafael Espindola	25d2c20c0c	Don't repeat name in comment and clang-format a function. llvm-svn: 228831	2015-02-11 14:44:17 +00:00
James Molloy	7c336576a5	[SimplifyCFG] Swap to using TargetTransformInfo for cost analysis. We're already using TTI in SimplifyCFG, so remove the hard-baked "cheapness" heuristic and use TTI directly. Generally NFC intended, but we're using a slightly different heuristic now so there is a slight test churn. Test changes: * combine-comparisons-by-cse.ll: Removed unneeded branch check. * 2014-08-04-muls-it.ll: Test now doesn't branch but emits muleq. * coalesce-subregs.ll: Superfluous block check. * 2008-01-02-hoist-fp-add.ll: fadd is safe to speculate. Change to udiv. * PhiBlockMerge.ll: Superfluous CFG checking code. Main checks still present. * select-gep.ll: A variable GEP is not expensive, just TCC_Basic, according to the TTI. llvm-svn: 228826	2015-02-11 12:15:41 +00:00
Daniel Sanders	a19216c8f4	[mips] Merge disassemblers into a single implementation. Summary: Currently we have Mips32 and Mips64 disassemblers and this causes the target triple to affect the disassembly despite all the relevant information being in the ELF header. These implementations do not need to be separate. This patch merges them together such that the appropriate tables are checked for the subtarget (e.g. Mips64 is checked when GP64 is enabled). Reviewers: vmedic Reviewed By: vmedic Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7498 llvm-svn: 228825	2015-02-11 11:28:56 +00:00
James Molloy	f147359376	[LoopReroll] Introduce the concept of DAGRootSets. A DAGRootSet models an induction variable being used in a rerollable loop. For example: x[i3+0] = y1 x[i3+1] = y2 x[i3+2] = y3 Base instruction -> i3 +---+----+ / \| \ ST[y1] +1 +2 <-- Roots \| \| ST[y2] ST[y3] There may be multiple DAGRootSets, for example: x[i2+0] = ... (1) x[i2+1] = ... (1) x[i2+4] = ... (2) x[i2+5] = ... (2) x[(i+1234)2+5678] = ... (3) x[(i+1234)2+5679] = ... (3) This concept is similar to the "Scale" member used previously, but allows multiple independent sets of roots based off the same induction variable. llvm-svn: 228821	2015-02-11 09:19:47 +00:00
David Majnemer	fad5a31160	AsmParser: Validate alloca's type An alloca's type should be weird things like metadata. llvm-svn: 228820	2015-02-11 09:13:11 +00:00
David Majnemer	04578fcfa5	DataLayout: Report when the preferred alignment is less than the ABI llvm-svn: 228819	2015-02-11 09:13:09 +00:00
David Majnemer	d7677e7a8d	Verifier: Check for null operands in !llvm.module.flags llvm-svn: 228818	2015-02-11 09:13:06 +00:00
Michael Kuperstein	1921d3d6f3	[X86] Split information collection from actual transformation in call frame optimization This splits collecting information from actually performing the transformation, so that we can add a heuristic in between the two. NFC. Differential Revision: http://reviews.llvm.org/D7497 llvm-svn: 228817	2015-02-11 08:53:55 +00:00
Arnaud A. de Grandmaison	de79026d5e	[PBQP] Cautiously update edge costs in the solver The NodeMetadata are maintained in an incremental way. When an edge between 2 nodes has its cost updated, in the course of graph reduction for example, the NodeMetadata need first to have the old edge cost removed, then the new edge cost added. Only once the NodeMetadata have been fully updated, it becomes safe to consider promoting the nodes to the ConservativelyAllocatable or OptimallyReducible sets. Previously, this promotion was occuring right after the removing the old cost, and this was breaking the assumption that a ConservativelyAllocatable should not be spilled. This patch also adds asserts to: - enforces the invariant that a node's reduction can not be downgraded, - only not provably allocatable or optimally reducible nodes can be spilled. llvm-svn: 228816	2015-02-11 08:25:36 +00:00
David Majnemer	9fd8cdc009	Verifier: Make sure !llvm.ident's operand isn't null llvm-svn: 228815	2015-02-11 08:23:20 +00:00
David Majnemer	300745351f	AsmParser: Don't crash when insertvalue has bad operands llvm-svn: 228813	2015-02-11 07:43:58 +00:00
David Majnemer	19b51054af	AsmParser: Switch some vectors to maps This speeds up parsing .ll files with metadata nodes with large IDs. llvm-svn: 228812	2015-02-11 07:43:56 +00:00
Peter Collingbourne	d20eff0ea6	Fix build for CMake < 2.8.12. llvm-svn: 228810	2015-02-11 05:58:57 +00:00
Zachary Turner	3bd47cee78	Use ADDITIONAL_HEADER_DIRS in all LLVM CMake projects. This allows IDEs to recognize the entire set of header files for each of the core LLVM projects. Differential Revision: http://reviews.llvm.org/D7526 Reviewed By: Chris Bieneman llvm-svn: 228798	2015-02-11 03:28:02 +00:00
Justin Bogner	d24e185784	InstrProf: Lower coverage mappings by setting their sections appropriately Add handling for __llvm_coverage_mapping to the InstrProfiling pass. We need to make sure the constant and any profile names it refers to are in the correct sections, which is easier and cleaner to do here where we have to know about profiling sections anyway. This is really tricky to test without a frontend, so I'm committing the test for the fix in clang. If anyone knows a good way to test this within LLVM, please let me know. Fixes PR22531. llvm-svn: 228793	2015-02-11 02:52:44 +00:00
Andrew Kaylor	7ad134a746	Temporary workaround to fix MSVC 2012 build problems llvm-svn: 228788	2015-02-11 02:16:34 +00:00
Reid Kleckner	96d011315a	Don't promote asynch EH invokes of nounwind functions to calls If the landingpad of the invoke is using a personality function that catches asynch exceptions, then it can catch a trap. Also add some landingpads to invalid LLVM IR test cases that lack them. Over-the-shoulder reviewed by David Majnemer. llvm-svn: 228782	2015-02-11 01:23:16 +00:00
Tom Stellard	94b7231740	R600/SI: Store immediate offsets > 12-bits in soffset This will save us from having to extend these offsets to 64-bits and storing them in a pair of vgprs. llvm-svn: 228776	2015-02-11 00:34:35 +00:00
Tom Stellard	c53861ab84	R600/SI: Add soffset operand to mubuf addr64 instruction We were previously hard-coding soffset to 0. llvm-svn: 228775	2015-02-11 00:34:32 +00:00
Zachary Turner	df3cc51f06	Fix some warnings due to -Wcovered-switch-default. llvm-svn: 228773	2015-02-11 00:13:39 +00:00
Zachary Turner	be6d1e49b0	Convert std::make_unique<> to llvm::make_unique<>. llvm-svn: 228768	2015-02-10 23:46:48 +00:00
Petar Jovanovic	d9f52043b1	Fix makeLibCall argument (signed) in SoftenFloatRes_XINT_TO_FP function The isSigned argument of makeLibCall function was hard-coded to false (unsigned). This caused zero extension on MIPS64 soft float. As the result SingleSource/Benchmarks/Stanford/FloatMM test and SingleSource/UnitTests/2005-07-17-INT-To-FP test failed. The solution was to use the proper argument. Patch by Strahinja Petrovic. Differential Revision: http://reviews.llvm.org/D7292 llvm-svn: 228765	2015-02-10 23:30:14 +00:00
Adrian Prantl	ca7e470221	Debug Info: Support variables that are described by more than one MMI table entry. This happens when SROA splits up an alloca and the resulting allocas cannot be lowered to SSA values because their address is passed to a function. Fixes PR22502. llvm-svn: 228764	2015-02-10 23:18:28 +00:00
Adrian Prantl	d49691f779	Fix indentation. llvm-svn: 228763	2015-02-10 23:18:15 +00:00
David Majnemer	7679300d93	EarlyCSE: It isn't safe to CSE across synchronization boundaries This fixes PR22514. llvm-svn: 228760	2015-02-10 23:09:43 +00:00
Zachary Turner	a5549178f1	Rewrite llvm-pdbdump in terms of LLVMDebugInfoPDB. This makes llvm-pdbdump available on all platforms, although it will currently fail to create a dumper if there is no PDB reader implementation for the current platform. It implements dumping of compilands and children, which is less information than was previously available, but it has to be rewritten from scratch using the new set of interfaces, so the rest of the functionality will be added back in subsequent commits. llvm-svn: 228755	2015-02-10 22:43:25 +00:00
David Majnemer	ca19485f08	X86: @llvm.frameaddress should defer to SelectionDAG for Win CFI llvm-svn: 228754	2015-02-10 22:00:34 +00:00
Simon Atanasyan	0ca59894aa	[Object] Reformat the code with clang-format No functional changes. llvm-svn: 228751	2015-02-10 21:38:25 +00:00
David Majnemer	13d0b11d7b	X86: Make @llvm.frameaddress work correctly with Windows unwind codes Simply loading or storing the frame pointer is not sufficient for Windows targets. Instead, create a synthetic frame object that we will lower later. References to this synthetic object will be replaced with the correct reference to the frame address. llvm-svn: 228748	2015-02-10 21:22:05 +00:00
Zachary Turner	cffff26b68	Provide DIA implementation of DebugInfoPDB. This implements DebugInfoPDB when the DIA SDK is present on the system. Specifically, this means that the following conditions are met: 1) You are building on Windows. 2) You are building with MSVC. 3) Visual Studio did not corrupt the installation of DIA due to a known issue with side-by-side installations of VS2012 and VS2013. If all of these conditions are true, you will be able to pass a value of PDB_Reader::DIA to PDB::createPdbReader(). There are no tests for this yet, as any test will be in the form of a lit test which tests the llvm-pdbdump.exe, which still needs to be rewritten in terms of this library. llvm-svn: 228747	2015-02-10 21:17:52 +00:00
Eric Christopher	f3e79e8714	Reformat (and remove some tabs) to make debugging this code a little easier to step through. llvm-svn: 228746	2015-02-10 21:15:06 +00:00
Andrew Kaylor	78b53dbcc1	Adding support for llvm.eh.begincatch and llvm.eh.endcatch intrinsics and beginning the documentation of native Windows exception handling. Differential Revision: http://reviews.llvm.org/D7398 llvm-svn: 228733	2015-02-10 19:52:43 +00:00
Tim Northover	43c0d2db50	DeadArgElim: arguments affect all returned sub-values by default. Unless we meet an insertvalue on a path from some value to a return, that value will be live if any of the return's components are live, so all of those components must be added to the MaybeLiveUses. Previously we were deleting arguments if sub-value 0 turned out to be dead. llvm-svn: 228731	2015-02-10 19:49:18 +00:00
Bill Schmidt	67f36bd0d8	Fix up r228725, missed change in PPCSubtarget definition llvm-svn: 228728	2015-02-10 19:31:55 +00:00
Duncan P. N. Exon Smith	4ee4a98eaa	IR: Add MDNode::replaceWithPermanent() Add new API for converting temporaries that may self-reference. Self-referencing nodes are not allowed to be uniqued, so sending them into `replaceWithUniqued()` is dangerous (and this commit adds assertions that prevent it). `replaceWithPermanent()` has similar semantics to `get()` followed by calls to `replaceOperandWith()`. In particular, if there's a self-reference, it returns a distinct node; otherwise, it returns a uniqued one. Like `replaceWithUniqued()` and `replaceWithDistinct()` (well, it calls out to them) it mutates the temporary node in place if possible, only calling `replaceAllUsesWith()` on a uniquing collision. llvm-svn: 228726	2015-02-10 19:13:46 +00:00
Bill Schmidt	82f1c775a0	[PowerPC] Fix reverted patch r227976 to avoid register assignment issues See full discussion in http://reviews.llvm.org/D7491. We now hide the add-immediate and call instructions together in a separate pseudo-op, which is tagged to define GPR3 and clobber the call-killed registers. The PPCTLSDynamicCall pass prior to RA now expands this op into the two separate addi and call ops, with explicit definitions of GPR3 on both instructions, and explicit clobbers on the call instruction. The pass is now marked as requiring and preserving the LiveIntervals and SlotIndexes analyses, and fixes these up after the replacement sequences are introduced. Self-hosting has been verified on LE P8 and BE P7 with various optimization levels, etc. It has also been verified with the --no-tls-optimize flag workaround removed. llvm-svn: 228725	2015-02-10 19:09:05 +00:00
David Majnemer	a7d908eb2b	X86: Emit Win64 SaveXMM opcodes at the right offset in the right order Walk the instructions marked FrameSetup and consider any stores of XMM registers to the stack as needing a SaveXMM opcode. This fixes PR22521. Differential Revision: http://reviews.llvm.org/D7527 llvm-svn: 228724	2015-02-10 19:01:47 +00:00
Hal Finkel	57c6ac5e41	[PowerPC] Support the (old) cntlz instruction alias Some old assembly code uses the cntlz alias for cntlzw, binutils supports this, and we should too. Fixes PR22519. llvm-svn: 228719	2015-02-10 18:45:02 +00:00
Colin LeMahieu	404d5b242d	[Hexagon] Adding vector load with post-increment instructions. Adding decoder function for 64bit control register class. llvm-svn: 228708	2015-02-10 16:59:36 +00:00
Zoran Jovanovic	416886793f	[mips][microMIPS] Implement movep instruction Differential Revision: http://reviews.llvm.org/D7465 llvm-svn: 228703	2015-02-10 16:36:20 +00:00
Jonas Paulsson	a25a3f4fea	Two comment typo fixes in lib/CodeGen/SelectionDAG/DAGCombiner.cpp. llvm-svn: 228700	2015-02-10 15:34:29 +00:00
Bradley Smith	e997b45076	[ARM] Add armv6s[-]m as an alias to armv6[-]m llvm-svn: 228696	2015-02-10 15:15:08 +00:00
Simon Pilgrim	d142ab7d08	[X86][AVX2] Missing AVX2 memory folding instructions Added most of the missing vector folding patterns for AVX2 (as well as fixing the vpermpd and verpmq patterns) Differential Revision: http://reviews.llvm.org/D7492 llvm-svn: 228688	2015-02-10 13:22:57 +00:00
Jonas Paulsson	afa6813816	Bugfix for missed dependency from store to load in buildSchedGraph(). Background: When handling underlying objects for a store, the vector of previous mem uses, mapped to the same Value, is afterwards cleared (regardless of ThisMayAlias). This means that during handling of the next store using the same Value, adjustChainDeps() must be called, otherwise a dependency might be missed. For example, three spill/reload (NonAliasing) memory accesses using the same Value 'a', with different offsets: SU(2): store @a SU(1): store @a, Offset:1 SU(0): load @a In this case we have: * SU(1) does not need a dep against SU(0). Therefore,SU(0) ends up in RejectMemNodes and is removed from the mem-uses list (AliasMemUses or NonAliasMemUses), as this list is cleared. * SU(2) needs a dep against SU(0). Therefore, SU(2) must check RejectMemNodes by calling adjustChainDeps(). Previously, for store SUs, adjustChainDeps() was only called if MayAlias was true, missing the S(2) to S(0) dependency in the case above. The fix is to always call adjustChainDeps(), regardless of MayAlias, since this applies both for AliasMemUses and NonAliasMemUses. No testcase found for any in-tree target. llvm-svn: 228686	2015-02-10 13:03:32 +00:00
Simon Pilgrim	cd32254a35	[X86][XOP] Added XOP memory folding patterns + tests This patch adds the complete AMD Bulldozer XOP instruction set to the memory folding pattern tables for stack folding, etc. Note: Many of the XOP instructions have multiple table entries as it can fold loads from different sources. Differential Revision: http://reviews.llvm.org/D7484 llvm-svn: 228685	2015-02-10 12:57:17 +00:00
Jozef Kolek	d68d424abf	[mips][microMIPS] Fix disassembling of 16-bit microMIPS instructions LWM16 and SWM16 Differential Revision: http://reviews.llvm.org/D7436 llvm-svn: 228683	2015-02-10 12:41:13 +00:00
Andrea Di Biagio	62622d2396	[X86][FastIsel] Avoid introducing legacy SSE instructions if the target has AVX. This patch teaches X86FastISel how to select AVX instructions for scalar float/double convert operations. Before this patch, X86FastISel always selected legacy SSE instructions for FPExt (from float to double) and FPTrunc (from double to float). For example: \code define double @foo(float %f) { %conv = fpext float %f to double ret double %conv } \end code Before (with -mattr=+avx -fast-isel) X86FastIsel selected a CVTSS2SDrr which is legacy SSE: cvtss2sd %xmm0, %xmm0 With this patch, X86FastIsel selects a VCVTSS2SDrr instead: vcvtss2sd %xmm0, %xmm0, %xmm0 Added test fast-isel-fptrunc-fpext.ll to check both the register-register and the register-memory float/double conversion variants. Differential Revision: http://reviews.llvm.org/D7438 llvm-svn: 228682	2015-02-10 12:04:41 +00:00
Chandler Carruth	2496910325	Revert r228556: InstCombine: propagate nonNull through assume This commit isn't using the correct context, and is transfoming calls that are operands to loads rather than calls that are operands to an icmp feeding into an assume. I've replied on the original review thread with a very reduced test case and some thoughts on how to rework this. llvm-svn: 228677	2015-02-10 08:07:32 +00:00
Craig Topper	9e71b82f40	[X86] Preserve mem refs on newly created 'Store' node instead of 'Load' node when handling store unfolding. Bug spotted by Steve King. I have no idea how to test this. llvm-svn: 228672	2015-02-10 06:29:28 +00:00
Craig Topper	f7e92f10b6	[X86] Remove unnecessary alignment checks from the load folding tables. llvm-svn: 228671	2015-02-10 05:10:50 +00:00
Zachary Turner	aeedd65c64	Teach llvm_add_library() to find include dirs. Since header files are not compilation units, CMake does not require you to specify them in the CMakeLists.txt file. As a result, unless a header file is explicitly added, CMake won't know about it, and when generating IDE-based projects, CMake won't put the header files into the IDE project. LLVM currently tries to deal with this in two ways: 1) It looks for all .h files that are in the project directory, and adds those. 2) llvm_add_library() understands the ADDITIONAL_HEADERS argument, which allows one to list an arbitrary list of headers. This patch takes things one step further. It adds the ability for llvm_add_library() to take an ADDITIONAL_HEADER_DIRS argument, which will specify a list of folders which CMake will glob for header files. Furthermore, it will glob not only for .h files, but also for .inc files. Included in this CL is an update to one of the existing users of ADDITIONAL_HEADERS to use this new argument instead, to serve as an illustration of how this cleans up the CMake. The big advantage of this new approach is that until now, there was no way for the IDE projects to locate the header files that are in the include tree. In other words, if you are in, for example, lib/DebugInfo/DWARF, the corresponding includes for this project will be located under include/llvm/DebugInfo/DWARF. Now, in the CMakeLists.txt for lib/DebugInfo/DWARF, you can simply write: ADDITIONAL_HEADER_DIRS ../../include/llvm/DebugInfo/DWARF as an argument to llvm_add_library(), and all header files will get added to the IDE project. Differential Revision: http://reviews.llvm.org/D7460 Reviewed By: Chris Bieneman llvm-svn: 228670	2015-02-10 05:04:37 +00:00
Chandler Carruth	b65d61a2e8	[x86] Fix PR22524: the DAG combiner was incorrectly handling illegal nodes when folding bitcasts of constants. We can't fold things and then check after-the-fact whether it was legal. Once we have formed the DAG node, arbitrary other nodes may have been collapsed to it. There is no easy way to go back. Instead, we need to test for the specific folding cases we're interested in and ensure those are legal first. This could in theory make this less powerful for bitcasting from an integer to some vector type, but AFAICT, that can't actually happen in the SDAG so its fine. Now, we only whitelist specific int->fp and fp->int bitcasts for post-legalization folding. I've added the test case from the PR. (Also as a note, this does not appear to be in 3.6, no backport needed) llvm-svn: 228656	2015-02-10 02:25:56 +00:00
Duncan P. N. Exon Smith	9e95f27eff	Verifier: reuse getInlinedAt() result, NFC llvm-svn: 228655	2015-02-10 02:25:18 +00:00
Duncan P. N. Exon Smith	bd33d375f0	IR: Remove unnecessary fields from MDTemplateParameter I noticed this fields were never used in r228607, but I neglected to propagate that into `MDTemplateParameter` until now. This really should have been done before commit in r228640; sorry for the churn. llvm-svn: 228652	2015-02-10 01:59:57 +00:00
Duncan P. N. Exon Smith	e4725beba7	Verifier: Check for valid tags in debug nodes Check that specialized `DebugNode`s have valid `DW_TAG`s. llvm-svn: 228649	2015-02-10 01:40:40 +00:00
Duncan P. N. Exon Smith	692bdb910d	Verifier: Add simple checks for MDLocation llvm-svn: 228647	2015-02-10 01:32:56 +00:00
Duncan P. N. Exon Smith	b0a19ad08a	Verifier: Create stubs for specialized metadata nodes llvm-svn: 228645	2015-02-10 01:09:50 +00:00
Duncan P. N. Exon Smith	ed458fa12c	AsmParser: Add stubs for specialized MDNodes, NFC Well, the exact error from the failed parse will change, but... llvm-svn: 228644	2015-02-10 01:08:16 +00:00
David Majnemer	93c22a45be	X86: Emit an ABI compliant prologue and epilogue for Win64 Win64 has specific contraints on what valid prologues and epilogues look like. This constraint is born from the flexibility and descriptiveness of Win64's unwind opcodes. Prologues previously emitted by LLVM could not be represented by the unwind opcodes, preventing operations powered by stack unwinding to successfully work. Differential Revision: http://reviews.llvm.org/D7520 llvm-svn: 228641	2015-02-10 00:57:42 +00:00
Duncan P. N. Exon Smith	01fc176977	IR: Add specialized debug info metadata nodes Add specialized debug info metadata nodes that match the `DIDescriptor` wrappers (used by `DIBuilder`) closely. Assembly and bitcode support to follow soon (it'll mostly just be obvious), but this sketches in today's schema. This is the first big commit (well, the only big one aside from the testcase changes that'll come when I move this into place) for PR22464. I've marked a bunch of obvious changes as `TODO`s in the source; I plan to make those changes promptly after this hierarchy is moved underneath `DIDescriptor`, but for now I'm aiming mostly to match the status quo. llvm-svn: 228640	2015-02-10 00:52:32 +00:00
Eric Christopher	d49868080e	Migrate PPCAsmPrinter's subtarget from reference to pointer in preparation for making it MachineFunction dependent. llvm-svn: 228638	2015-02-10 00:44:17 +00:00
David Blaikie	36a036909c	Fix the clang -Werror build (-Wunused-variable) llvm-svn: 228635	2015-02-10 00:16:36 +00:00
Philip Reames	7e7dc3e9df	Adjust how we avoid poll insertion inside the poll function (NFC) I realized that my early fix for this was overly complicated. Rather than scatter checks around in a bunch of places, just exit early when we visit the poll function itself. Thinking about it a bit, the whole inlining mechanism used with gc.safepoint_poll could probably be cleaned up a bit. Originally, poll insertion was fused with gc relocation rewriting. It might be worth going back to see if we can simplify the chain of events now that these two are seperated. As one thought, maybe it makes sense to rewrite calls inside the helper function before inlining it to the many callers. This would require us to visit the poll function before any other functions though.. llvm-svn: 228634	2015-02-10 00:04:53 +00:00
Adrian Prantl	34e7590e0d	Debug info: When updating debug info during SROA, do not emit debug info for any padding introduced by SROA. In particular, do not emit debug info for an alloca that represents only the padding introduced by a previous iteration. Fixes PR22495. llvm-svn: 228632	2015-02-09 23:57:22 +00:00
Adrian Prantl	27bd01f71c	Debug info: Use DW_OP_bit_piece instead of DW_OP_piece in the intermediate representation. This - increases consistency by using the same granularity everywhere - allows for pieces < 1 byte - DW_OP_piece didn't actually allow storing an offset. Part of PR22495. llvm-svn: 228631	2015-02-09 23:57:15 +00:00
Colin LeMahieu	328b1633d7	[Hexagon] Adding missing load instructions and removing an unused multiclass parameter. llvm-svn: 228630	2015-02-09 23:45:24 +00:00
Colin LeMahieu	4282e7cffd	[Hexagon] Factoring classes out of some load patterns and deleting some unused ones. llvm-svn: 228627	2015-02-09 23:05:44 +00:00
Ramkumar Ramachandra	3edf74fe29	[Statepoint] Improve two asserts, fix some style (NFC) Summary: It's important that our users immediately know what gc.safepoint_poll is. Also fix the style of the declaration of CreateGCStatepoint, in preparation for another change that will wrap it. Reviewers: reames Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7517 llvm-svn: 228626	2015-02-09 23:02:10 +00:00
Ramkumar Ramachandra	2e4b9e0a37	PlaceSafepoints: modernize gc.result.* -> gc.result Differential Revision: http://reviews.llvm.org/D7516 llvm-svn: 228625	2015-02-09 23:00:40 +00:00
Duncan P. N. Exon Smith	b407bb2789	DebugInfo: Remove DW_TAG_constant Remove handling for DW_TAG_constant. We started producing it in r110656, but reverted that in r110876 without dropping the support. Finish the job. llvm-svn: 228623	2015-02-09 22:48:04 +00:00
Philip Reames	d4a912fefd	Update file comment to clarify points highlighted in review (NFC) llvm-svn: 228621	2015-02-09 22:44:03 +00:00
Philip Reames	a29de87ea4	Use range for loops in PlaceSafepoints (NFC) llvm-svn: 228620	2015-02-09 22:26:11 +00:00
Duncan P. N. Exon Smith	bd75ad4d0c	IR: Take uint64_t in DIBuilder::createExpression() `DIExpression` deals with `uint64_t`, so it doesn't make sense that `createExpression()` is created from `int64_t`. Switch to `uint64_t` to unify them. I've temporarily left in the `int64_t` version, which forwards to the `uint64_t` version. I'll delete it once I've updated the callers. llvm-svn: 228619	2015-02-09 22:13:27 +00:00
Colin LeMahieu	4fd203d3e1	[Hexagon] Removing more V4 predicates since V4 is the required minimum. llvm-svn: 228614	2015-02-09 21:56:37 +00:00
Ramkumar Ramachandra	82ab65c7cd	MemDerefPrinter: Require DataLayoutPass for higher accuracy Without a valid data layout, deferenceable(N) doesn't get parsed or propagated. Since this is the key item we are testing, add a dependency on the pass. Differential Revision: http://reviews.llvm.org/D7508 llvm-svn: 228611	2015-02-09 21:50:03 +00:00
Philip Reames	b1ed02f728	Add basic tests for PlaceSafepoints This is just adding really simple tests which should have been part of the original submission. When doing so, I discovered that I'd mistakenly removed required pieces when preparing the patch for upstream submission. I fixed two such bugs in this submission. llvm-svn: 228610	2015-02-09 21:48:05 +00:00
Duncan P. N. Exon Smith	ac3ed7afc9	Verifier: Const-qualify Metadata, NFC llvm-svn: 228609	2015-02-09 21:30:05 +00:00
Ramkumar Ramachandra	a7343d65f4	isDereferenceablePointer: look through gc.relocate calls While a theoretical GC might change dereferenceability on collection, there is no such known collector and no need to account for the case with a flag yet. Differential Revision: http://reviews.llvm.org/D7454 llvm-svn: 228606	2015-02-09 21:08:03 +00:00
Colin LeMahieu	641c24b9bf	[Hexagon] Removing v2-4 flags. V4 is the minimum supported version. llvm-svn: 228605	2015-02-09 21:07:35 +00:00
Ben Langmuir	d2d52de229	Reduce the LockFileManager timeout, and provide unsafeRemoveLockFile 5 minutes is an eternity, so try to strike a better balance between waiting long enough for any reasonable module build and not so long that users kill the process because they think it's hanging. Also give the client a way to delete the lock file after a timeout. llvm-svn: 228603	2015-02-09 20:34:24 +00:00
Colin LeMahieu	955c4ff9c3	[Hexagon] Factoring classes out of store patterns. llvm-svn: 228602	2015-02-09 20:33:46 +00:00
Colin LeMahieu	ab5a8d6070	[Hexagon] Formatting v5 TD file. Removing commented defs. llvm-svn: 228598	2015-02-09 20:03:42 +00:00
Ramkumar Ramachandra	010b77c3a2	MemDepPrinter: cleanup a few loops (NFC) Make use of the newly introduced inst_range to clean up two loops. Clean up a third one while at it. Differential Revision: http://reviews.llvm.org/D7455 llvm-svn: 228596	2015-02-09 19:49:54 +00:00
Colin LeMahieu	38e6689276	[Hexagon] Cleaning up definition formatting. llvm-svn: 228593	2015-02-09 19:24:44 +00:00
Sanjoy Das	bf5d870dfa	Bugfix: SCEV incorrectly marks certain add recurrences as nsw When creating a scev for sext({X,+,Y}), scev checks if the expression is equivalent to {sext X,+,zext Y}. If it can prove that, it also tags the original {X,+,Y} as <nsw>, which is not correct. In the test case I run `-scalar-evolution` twice because the bug manifests only once SCEV has run through and seen the `sext` expressions (and then does a in-place mutation on {X,+,Y}). Differential Revision: http://reviews.llvm.org/D7495 llvm-svn: 228586	2015-02-09 18:34:55 +00:00
Kit Barton	0b0cdb1cd4	This change implements the following three logical vector operations: veqv (vector equivalence) vnand vorc I increased the AddedComplexity for these instructions to 500 to ensure they are generated instead of issuing other VSX instructions. Phabricator review: http://reviews.llvm.org/D7469 llvm-svn: 228580	2015-02-09 17:03:18 +00:00
Sanjay Patel	a7b893d5c0	rename variable to give it some meaning; remove obvious comments; NFC llvm-svn: 228579	2015-02-09 16:30:58 +00:00
Sanjay Patel	fc54c61c56	fix comment that didn't match the code; remove unnecessary braces; NFC llvm-svn: 228578	2015-02-09 16:04:52 +00:00
Johannes Doerfert	2683e5676c	Allow ScalarEvolution to catch more min/max cases For the attached test case different types are used in the ICmpInst and SelectInst that represent the min/max expressions. However, if the ICmpInst type is smaller a comparison with the sign/zero extended operands would have yielded the same result. This situation might arise after the instruction combination pass was applied. Differential Revision: http://reviews.llvm.org/D7338 llvm-svn: 228572	2015-02-09 12:34:23 +00:00
Akira Hatanaka	8d3cb829ce	Fix a bug in DemoteRegToStack where a reload instruction was inserted into the wrong basic block. This would happen when the result of an invoke was used by a phi instruction in the invoke's normal destination block. An instruction to reload the invoke's value would get inserted before the critical edge was split and a new basic block (which is the correct insertion point for the reload) was created. This commit fixes the bug by splitting the critical edge before all the reload instructions are inserted. Also, hoist up the code which computes the insertion point to the only place that need that computation. rdar://problem/15978721 llvm-svn: 228566	2015-02-09 06:38:23 +00:00
David Majnemer	1de3094d78	MC: Calculate intra-section symbol differences correctly for COFF This fixes PR22060. llvm-svn: 228565	2015-02-09 06:31:31 +00:00
Craig Topper	141e65e69c	[X86] Remove 256-bit and 512-bit memop pattern fragments. They are no longer used. llvm-svn: 228563	2015-02-09 04:04:53 +00:00
Craig Topper	820d49270d	[X86] Remove 'memop' uses from AVX512. Use 'load' instead. llvm-svn: 228562	2015-02-09 04:04:50 +00:00
Tim Northover	705d2af9e1	DeadArgElim: fix mismatch in accounting of array return types. Some parts of DeadArgElim were only considering the individual fields of StructTypes separately, but others (where insertvalue & extractvalue instructions occur) also looked into ArrayTypes. This one is an actual bug; the mismatch can lead to an argument being considered used by a return sub-value that isn't being tracked (and hence is dead by default). It then gets incorrectly eliminated. llvm-svn: 228559	2015-02-09 01:21:00 +00:00
Tim Northover	854c927de5	DeadArgElim: assess uses of entire return value aggregate. Previously, a non-extractvalue use of an aggregate return value meant the entire return was considered live (the algorithm gave up entirely). This was correct, but conservative. It's better to actually look at that Use, making the analysis results apply to all sub-values under consideration. E.g. %val = call { i32, i32 } @whatever() [...] ret { i32, i32 } %val The return is using the entire aggregate (sub-values 0 and 1). We can still simplify @whatever if we can prove that this return is itself unused. Also unifies the logic slightly between aggregate and non-aggregate cases.. llvm-svn: 228558	2015-02-09 01:20:53 +00:00
Lang Hames	114b4f324b	[Orc] Add a JITSymbol class to the Orc APIs, refactor APIs, update clients. This patch refactors a key piece of the Orc APIs: It removes the ::getSymbolAddress and ::lookupSymbolAddressIn methods, which returned target addresses (uint64_ts), and replaces them with ::findSymbol and ::findSymbolIn respectively, which return instances of the new JITSymbol type. Unlike the old methods, calling findSymbol or findSymbolIn does not cause the symbol to be immediately materialized when found. Instead, the symbol will be materialized if/when the getAddress method is called on the returned JITSymbol. This allows us to query for the existence of symbols without actually materializing them. In the future I expect more information to be attached to the JITSymbol class, for example whether the returned symbol is a weak or strong definition. This will allow us to properly handle weak symbols and multiple definitions. llvm-svn: 228557	2015-02-09 01:20:51 +00:00
Ramkumar Ramachandra	a021ee62ca	InstCombine: propagate nonNull through assume Make assume (load (call\|invoke) != null) set nonNull return attribute for the call and invoke. Also include tests. Differential Revision: http://reviews.llvm.org/D7107 llvm-svn: 228556	2015-02-09 01:13:13 +00:00
David Blaikie	e4698b9a6c	Fix -Wuninitialized build by referencing the relevant ctor parameter instead of the base class member variable. llvm-svn: 228554	2015-02-08 23:15:37 +00:00
Zachary Turner	98571ed996	Make PDBSymbol's IPDBSymbol reference const. llvm-svn: 228553	2015-02-08 22:53:53 +00:00
Sanjoy Das	f2e931cae9	Bugfix: ScalarEvolution incorrectly assumes that the start of certain add recurrences don't overflow. This change makes the optimization more restrictive. It still assumes that an overflowing `add nsw` is undefined behavior; and this change will need revisiting once we have a consistent semantics for poison values. Differential Revision: http://reviews.llvm.org/D7331 llvm-svn: 228552	2015-02-08 22:52:17 +00:00
Craig Topper	68ab0465a0	[X86] Remove the remaining uses of memop from AVX and AVX2 instruction patterns. AVX and AVX2 can handle unaligned loads being folded so we can just use 'load' llvm-svn: 228551	2015-02-08 22:38:25 +00:00
Benjamin Kramer	4c1f097bc8	Metadata: Use <algorithm> to simplify code. NFC. llvm-svn: 228550	2015-02-08 21:56:09 +00:00
Zachary Turner	bae16b3f53	DebugInfoPDB: Make the symbol base case hold an IPDBSession ref. Dumping a symbol often requires access to data that isn't inside the symbol hierarchy, but which is only accessible through the top-level session. This patch is a pure interface change to give symbols a reference to the session. llvm-svn: 228542	2015-02-08 20:58:09 +00:00
Sanjay Patel	3510bc7162	fix typos; NFC llvm-svn: 228529	2015-02-08 18:54:22 +00:00
Zachary Turner	afdff425d7	Make UTF8->UTF16 conversion null terminate output on empty input. llvm-svn: 228527	2015-02-08 18:08:51 +00:00
Simon Pilgrim	d11b013623	Moved AVX2 vbroadcast (reg) instruction foldings under the correct grouping. NFC. llvm-svn: 228526	2015-02-08 17:13:54 +00:00
Bjorn Steinbrink	5ec7522771	Correctly combine alias.scope metadata by a union instead of intersecting Summary: The alias.scope metadata represents sets of things an instruction might alias with. When generically combining the metadata from two instructions the result must be the union of the original sets, because the new instruction might alias with anything any of the original instructions aliased with. Reviewers: hfinkel Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7490 llvm-svn: 228525	2015-02-08 17:07:14 +00:00
Elena Demikhovsky	23a485a4ed	Masked Gather and Scatter Intrinsics. Gather and Scatter are new introduced intrinsics, comming after recently implemented masked load and store. This is the first patch for Gather and Scatter intrinsics. It includes only the syntax, parsing and verification. Gather and Scatter intrinsics allow to perform multiple memory accesses (read/write) in one vector instruction. The intrinsics are not target specific and will have the following syntax: Gather: declare <16 x i32> @llvm.masked.gather.v16i32(<16 x i32> <vector of ptrs>, i32 <alignment>, <16 x i1> <mask>, <16 x i32> <passthru>) declare <8 x float> @llvm.masked.gather.v8f32(<8 x float><vector of ptrs>, i32 <alignment>, <8 x i1> <mask>, <8 x float><passthru>) Scatter: declare void @llvm.masked.scatter.v8i32(<8 x i32><vector value to be stored> , <8 x i32><vector of ptrs> , i32 <alignment>, <8 x i1> <mask>) declare void @llvm.masked.scatter.v16i32(<16 x i32> <vector value to be stored> , <16 x i32> <vector of ptrs>, i32 <alignment>, <16 x i1><mask> ) Vector of ptrs - a set of source/destination addresses, to load/store the value. Mask - switches on/off vector lanes to prevent memory access for switched-off lanes vector of ptrs, value and mask should have the same vector width. These are code examples where gather / scatter should be used and will allow function vectorization ;void foo1(int * restrict A, int * restrict B, int * restrict C) { ; for (int i=0; i<SIZE; i++) { ; A[i] = B[C[i]]; ; } ;} ;void foo3(int * restrict A, int * restrict B) { ; for (int i=0; i<SIZE; i++) { ; A[B[i]] = i+5; ; } ;} Tests will come in the following patches, with CodeGen and Vectorizer. http://reviews.llvm.org/D7433 llvm-svn: 228521	2015-02-08 08:27:19 +00:00
Tim Northover	45aa89c925	ARM & AArch64: teach LowerVSETCC that output type size may differ from input. While various DAG combines try to guarantee that a vector SETCC operation will have the same output size as input, there's nothing intrinsic to either creation or LegalizeTypes that actually guarantees it, so the function needs to be ready to handle a mismatch. Fortunately this is easy enough, just extend or truncate the naturally compared result. I couldn't reproduce the failure in other backends that I know have SIMD, so it's probably only an issue for these two due to shared heritage. Should fix PR21645. llvm-svn: 228518	2015-02-08 00:50:47 +00:00
Zachary Turner	92545a03df	Removed unused function mistakenly left in, triggering -Werror. llvm-svn: 228517	2015-02-08 00:41:31 +00:00
Zachary Turner	21473f7bb6	Some cleanup for libpdb. This patch implements a few of the optional suggestions from the initial patch comitting libpdb. In particular, it implements a virtual function out of line for each of the concrete classes. A few other minor cleanups exist as well, such as using override instead of virtual, etc. llvm-svn: 228516	2015-02-08 00:29:29 +00:00
Craig Topper	e169c57188	[X86] Add register use/def for wrmsr and rdmsr. llvm-svn: 228515	2015-02-07 23:36:51 +00:00
Craig Topper	1d472db8cc	[X86] Add GETSEC instruction. llvm-svn: 228514	2015-02-07 23:36:36 +00:00
Simon Pilgrim	a2618679a8	[X86][AVX] Added missing stack folding support + test for vptest ymm instruction llvm-svn: 228509	2015-02-07 21:44:06 +00:00
Benjamin Kramer	f094d77de8	LoopIdiom: Use utility functions. The only difference between deleteIfDeadInstruction and RecursivelyDeleteTriviallyDeadInstructions is that the former also manually invalidates SCEV. That's unnecessary because SCEV automatically gets informed when an instruction is deleted via a ValueHandle. NFC. llvm-svn: 228508	2015-02-07 21:37:08 +00:00
Joerg Sonnenberger	a73284a2a4	Avoid integer overflows around realloc calls resulting in potential heap. Problem identified by Guido Vranken. Changes differ from original OpenBSD sources by not depending on non-portable reallocarray. llvm-svn: 228507	2015-02-07 21:24:06 +00:00
Benjamin Kramer	17d9015d27	ValueTracking: Make isBytewiseValue simpler and more powerful at the same time. Turns out there is a simpler way of checking that all bytes in a word are equal than binary decomposition. llvm-svn: 228503	2015-02-07 19:29:02 +00:00
Bjorn Steinbrink	71bf3b800a	Properly update AA metadata when performing call slot optimization Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7482 llvm-svn: 228500	2015-02-07 17:54:36 +00:00
Ahmed Bougacha	29efe3b287	[BasicAA] Try to disambiguate GEPs through arrays of structs into different fields. We can show that two GEPs off of the same (possibly multidimensional) array of structs, into different fields, can't alias. Quoting: For two GEPOperators GEP1 and GEP2, if we find that: - both GEPs begin indexing from the exact same pointer; - the last indices in both GEPs are constants, indexing into a struct; - said indices are different, hence,the pointed-to fields are different; - and both GEPs only index through arrays prior to that; this lets us determine that the struct that GEP1 indexes into and the struct that GEP2 indexes into must either precisely overlap or be completely disjoint. Because they cannot partially overlap, indexing into different non-overlapping fields of the struct will never alias. The other BasicAA::aliasGEP rules worked in some cases, but not all (for example, the i32x3 struct in the testcase). We can add this simple ad-hoc rule to complement them. rdar://19717375 Differential Revision: http://reviews.llvm.org/D7453 llvm-svn: 228498	2015-02-07 17:04:29 +00:00
Benjamin Kramer	d7e331e0f9	SCEV: Compress disposition pairs. Composing DenseMaps and SmallVectors is still somewhat suboptimal, but this at least halves the size of the vector elements. NFC. llvm-svn: 228497	2015-02-07 16:41:12 +00:00
Andrea Di Biagio	4f8bdcb738	Fix typos; NFC. llvm-svn: 228493	2015-02-07 13:56:20 +00:00
Benjamin Kramer	a9591b5880	Move DebugLocs around instead of copying. llvm-svn: 228491	2015-02-07 12:28:15 +00:00
David Majnemer	5614ea9aae	MC: Emit COFF section flags in the "proper" order COFF section flags are not idempotent: 'rd' will make a read-write section because 'd' implies write 'dr' will make a read-only section because 'r' disables write llvm-svn: 228490	2015-02-07 08:26:40 +00:00
Hal Finkel	291cc7bacd	[PowerPC] Handle loop predecessor invokes If a loop predecessor has an invoke as its terminator, and the return value from that invoke is used to determine the loop iteration space, then we can't insert a computation based on that value in the loop predecessor prior to the terminator (oops). If there's such an invoke, or just no predecessor for that matter, insert a new loop preheader. llvm-svn: 228488	2015-02-07 07:32:58 +00:00
Bruce Mitchener	7e575ed1ea	Add more DWARF 5 language constants. Differential Revision: http://reviews.llvm.org/D7430 llvm-svn: 228487	2015-02-07 06:35:30 +00:00
Zachary Turner	94269a4db3	Resubmit unittests for DebugInfoPDB. These were originally submitted as part of r228428, but this part caused a build breakage in LLVMConfig. The library portion was resubmitted independently since it was not causing breakage. There were two reasons this was causing the build to fail. The first is that there were no Makefiles added for the PDB tests. And the second is that the DebugInfoPDB library was only being built by CMake behind an "if (MSVC)" check. This is wrong since this the library hides platform specific details, and it was causing LLVM-Config to not find the library when trying to build unittests. llvm-svn: 228482	2015-02-07 01:47:14 +00:00
Duncan P. N. Exon Smith	e7e2abe6a2	Support: Add dwarf::getVirtuality() llvm-svn: 228474	2015-02-07 00:37:15 +00:00
Duncan P. N. Exon Smith	d6f3574210	Support: Use Dwarf.def for DW_VIRTUALITY, NFC Use definition file for `DW_VIRTUALITY_*`. Add a `DW_VIRTUALITY_max` both for ease of testing and for future use by the `LLParser`. llvm-svn: 228473	2015-02-07 00:36:23 +00:00
Duncan P. N. Exon Smith	e07f13ae35	Support: Add dwarf::getAttributeEncoding() llvm-svn: 228470	2015-02-06 23:46:49 +00:00
Duncan P. N. Exon Smith	dd563dd341	Support: Rewrite AttributeEncodingString(), NFC llvm-svn: 228469	2015-02-06 23:45:37 +00:00
Duncan P. N. Exon Smith	8d4eeb53f6	Support: Stop stringifying DW_ATE_{lo,hi}_user llvm-svn: 228468	2015-02-06 23:44:24 +00:00
Ahmed Bougacha	df956a2e78	[AArch64] Use the source location of the IR branch when creating Bcc from a conditional branch fed by an add/sub/mul-with-overflow node. We previously used the SDLoc of the overflow node, for no good reason. In some cases, this led to the Bcc and B terminators having different source orders, and DBG_VALUEs being inserted between them. The real issue is with the code that can't handle DBG_VALUEs between terminators: the few places affected by this will be fixed soon. In the meantime, fixing the SDLoc is a positive change no matter what. No tests, as I have no idea how to get .loc emitted for branches? rdar://19347133 llvm-svn: 228463	2015-02-06 23:15:39 +00:00
Hal Finkel	0d2a1515d5	Revert "r227976 - [PowerPC] Yet another approach to __tls_get_addr" and related fixups Unfortunately, even with the workaround of disabling the linker TLS optimizations in Clang restored (which has already been done), this still breaks self-hosting on my P7 machine (-O3 -DNDEBUG -mcpu=native). Bill is currently working on an alternate implementation to address the TLS issue in a way that also fully elides the linker bug (which, unfortunately, this approach did not fully), so I'm reverting this now. llvm-svn: 228460	2015-02-06 23:07:40 +00:00
Duncan P. N. Exon Smith	d40af00e66	Support: Add dwarf::getLanguage() llvm-svn: 228458	2015-02-06 22:55:13 +00:00
Duncan P. N. Exon Smith	0317944b1a	Support: Rewrite dwarf::LanguageString(), NFC llvm-svn: 228457	2015-02-06 22:53:19 +00:00
Duncan P. N. Exon Smith	af677ebb41	IR: Allow 32-bits for lines in debug location Remove unnecessary restriction of 24-bits for line numbers in `MDLocation`. The rest of the debug info schema (with the exception of local variables) uses 32-bits for line numbers. As I introduce the specialized nodes, it makes sense to canonicalize on one size or the other. llvm-svn: 228455	2015-02-06 22:50:13 +00:00
Sanjay Patel	3d982214b0	use local variables; NFC llvm-svn: 228452	2015-02-06 22:43:52 +00:00
Duncan P. N. Exon Smith	4031beb10a	Support: Stop stringifying DW_LANG_{lo,hi}_user llvm-svn: 228451	2015-02-06 22:34:48 +00:00
Duncan P. N. Exon Smith	a81f7e1a63	AsmParser: Use DW_TAG_hi_user instead of magic constant, NFC llvm-svn: 228448	2015-02-06 22:29:35 +00:00
Duncan P. N. Exon Smith	c6da7f1c74	AsmWriter: Extract writeTag(), NFC llvm-svn: 228447	2015-02-06 22:28:05 +00:00
Duncan P. N. Exon Smith	61a0933007	AsmWriter: Extract writeMetadataAsOperand(), NFC llvm-svn: 228446	2015-02-06 22:27:22 +00:00
Evgeniy Stepanov	4e12057760	[msan] Fix "missing origin" in atomic store. An atomic store always make the target location fully initialized (in the current implementation). It should not store origin. Initialized memory can't have meaningful origin, and, due to origin granularity (4 bytes) there is a chance that this extra store would overwrite meaningfull origin for an adjacent location. llvm-svn: 228444	2015-02-06 21:47:39 +00:00
Cameron Esfahani	f0bf1947b7	Test commit to see if it triggers an email to llvm-commits. No change. llvm-svn: 228442	2015-02-06 21:33:08 +00:00
Zachary Turner	8022f24fe8	Try to fix Makefile build for LLVMDebugInfoPDB. llvm-svn: 228437	2015-02-06 20:42:03 +00:00
Zachary Turner	0e9e66330a	Resubmit "Create lib/DebugInfo/PDB" (r228428) This change resubmits the patch that broke the build, this time without unittests. The unittests will be submitted separately after the problem has been addressed: --Original Commit Message-- Create lib/DebugInfo/PDB. This patch creates a platform-independent interface to a PDB reader. There is currently no implementation of this interface, which will be provided in future patches. This defines the basic object model which any implementation must conform to. Reviewed by: David Blaikie Differential Revision: http://reviews.llvm.org/D7356 llvm-svn: 228435	2015-02-06 20:30:52 +00:00
Michael Zolotukhin	7af83c1f39	Use estimated number of optimized insns in unroll-threshold computation. If complete-unroll could help us to optimize away N% of instructions, we might want to do this even if the final size would exceed loop-unroll threshold. However, we don't want to unroll huge loop, and we are add AbsoluteThreshold to avoid that - this threshold will never be crossed, even if we expect to optimize 99% instructions after that. llvm-svn: 228434	2015-02-06 20:20:40 +00:00
Michael Zolotukhin	4e8598eee3	[InstSimplify] Add SimplifyFPBinOp function. It is a variation of SimplifyBinOp, but it takes into account FastMathFlags. It is needed in inliner and loop-unroller to accurately predict the transformation's outcome (previously we dropped the flags and were too conservative in some cases). Example: float foo(float a, float b) { float r; if (a[1] b) r = /* a lot of expensive computations /; else r = 1; return r; } float boo(float a) { return foo(a, 0.0); } Without this patch, we don't inline 'foo' into 'boo'. llvm-svn: 228432	2015-02-06 20:02:51 +00:00
Zachary Turner	8c89a82c88	Revert "Create lib/DebugInfo/PDB." This reverts commit 21028, as it is causing failures in LLVMConfig. llvm-svn: 228431	2015-02-06 20:00:18 +00:00
Kostya Serebryany	db4d645714	[fuzzer] move default sanitizer options to a separate file llvm-svn: 228429	2015-02-06 19:52:07 +00:00
Zachary Turner	079ba9224f	Create lib/DebugInfo/PDB. This patch creates a platform-independent interface to a PDB reader. There is currently no implementation of this interface, which will be provided in future patches. This defines the basic object model which any implementation must conform to. Reviewed by: David Blaikie Differential Revision: http://reviews.llvm.org/D7356 llvm-svn: 228428	2015-02-06 19:44:09 +00:00
Lang Hames	8a6f35546b	[Orc] Move SectionMemoryManager's implementation from MCJIT to ExecutionEngine. This is a more sensible home for SectionMemoryManager, and allows the implementation to be shared between Orc and MCJIT. llvm-svn: 228427	2015-02-06 19:36:40 +00:00
Quentin Colombet	a8cb36ea43	[LiveIntervalAnalysis] Speed up creation of live ranges for physical registers by using a segment set. The patch addresses a compile-time performance regression in the LiveIntervals analysis pass (see http://llvm.org/bugs/show_bug.cgi?id=18580). This regression is especially critical when compiling long functions. Our analysis had shown that the most of time is taken for generation of live intervals for physical registers. Insertions in the middle of the array of live ranges cause quadratic algorithmic complexity, which is apparently the main reason for the slow-down. Overview of changes: - The patch introduces an additional std::set<Segment>* member in LiveRange for storing segments in the phase of initial creation. The set is used if this member is not NULL, otherwise everything works the old way. - The set of operations on LiveRange used during initial creation (i.e. used by createDeadDefs and extendToUses) have been reimplemented to use the segment set if it is available. - After a live range is created the contents of the set are flushed to the segment vector, because the set is not as efficient as the vector for the later uses of the live range. After the flushing, the set is deleted and cannot be used again. - The set is only for live ranges computed in LiveIntervalAnalysis::computeLiveInRegUnits() and getRegUnit() but not in computeVirtRegs(), because I did not bring any performance benefits to computeVirtRegs() and for some examples even brought a slow down. Patch by Vaidas Gasiunas <vaidas.gasiunas@sap.com> Differential Revision: http://reviews.llvm.org/D6013 llvm-svn: 228421	2015-02-06 18:42:41 +00:00
Adam Nemet	7206d7a5d2	[LV] Move addRuntimeCheck to LoopAccessAnalysis This will allow it to be shared with the new Loop Distribution pass. getFirstInst is currently duplicated across LoopVectorize.cpp and LoopAccessAnalysis.cpp. This is a short-term work-around until we figure out a better solution. NFC. (The code moved is adjusted a bit for the name of the Loop member and that PtrRtCheck is now a reference rather than a pointer.) llvm-svn: 228418	2015-02-06 18:31:04 +00:00
Reid Kleckner	526ec29370	Don't dllexport declarations Fixes PR22488 llvm-svn: 228411	2015-02-06 17:59:49 +00:00
Benjamin Kramer	970eac40bf	Make helper functions/classes/globals static. NFC. llvm-svn: 228410	2015-02-06 17:51:54 +00:00
Matthias Braun	2e404597f4	InstCombine: Combine select sequences into a single select Normalize select(C0, select(C1, a, b), b) -> select((C0 & C1), a, b) select(C0, a, select(C1, a, b)) -> select((C0 \| C1), a, b) This normal form may enable further combines on the And/Or and shortens paths for the values. Many targets prefer the other but can go back easily in CodeGen. Differential Revision: http://reviews.llvm.org/D7399 llvm-svn: 228409	2015-02-06 17:49:36 +00:00
Matthias Braun	7dc96dea0a	LiveInterval: Fix SubRange memory leak. llvm-svn: 228405	2015-02-06 17:28:47 +00:00
Benjamin Kramer	69b4ad29af	AArch64PromoteConstant: Modernize and resolve some Use<->User confusion. NFC. llvm-svn: 228399	2015-02-06 14:43:55 +00:00
Benjamin Kramer	39f76acb5c	IRCE: Demote template to ArrayRef and SmallVector to array. NFC. llvm-svn: 228398	2015-02-06 14:43:49 +00:00
Chad Rosier	92c1f363f4	Whitespace. llvm-svn: 228397	2015-02-06 14:14:41 +00:00
Arnaud A. de Grandmaison	77f62d45ef	[PBQP] Fix comment wording. NFC llvm-svn: 228390	2015-02-06 11:28:16 +00:00
Michel Danzer	d89a557f33	R600/SI: Don't enable WQM for V_INTERP_* instructions v2 Doesn't seem necessary anymore. I think this was mostly compensating for not enabling WQM for texture sampling instructions. v2: Add test coverage Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228373	2015-02-06 02:51:25 +00:00
Michel Danzer	494391ba47	R600/SI: Also enable WQM for image opcodes which calculate LOD v3 If whole quad mode isn't enabled for these, the level of detail is calculated incorrectly for pixels along diagonal triangle edges, causing artifacts. v2: Use a TSFlag instead of lots of switch cases v3: Add test coverage Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=88642 Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228372	2015-02-06 02:51:20 +00:00
Ramkumar Ramachandra	8378ac3684	Introduce print-memderefs to test isDereferenceablePointer Since testing the function indirectly is tricky, introduce a direct print-memderefs pass, in the same spirit as print-memdeps, which prints dereferenceability information matched by FileCheck. Differential Revision: http://reviews.llvm.org/D7075 llvm-svn: 228369	2015-02-06 01:46:42 +00:00
Daniel Jasper	4bb224decb	Small cleanup of MachineLICM.cpp Specifically: - Calculate the loop pre-header once at the stat of HoistOutOfLoop, so: - We don't-DFS walk the MachineDomTree if we aren't going to do anything - Don't call getCurPreheader for each Scope - Don't needlessly use a do-while loop - Use early exit for Scopes.size() == 0 No functional changes intended. llvm-svn: 228350	2015-02-05 22:39:46 +00:00
Colin LeMahieu	6e3e62fd13	[Hexagon] Renaming v4 compare-and-jump instructions. llvm-svn: 228349	2015-02-05 22:03:32 +00:00
Colin LeMahieu	6b3aac8ad9	[Hexagon] Deleting unused patterns. llvm-svn: 228348	2015-02-05 21:43:56 +00:00
Colin LeMahieu	de68b66b4d	[Hexagon] Simplifying and formatting several patterns. Changing a pattern multiply to be expanded. llvm-svn: 228347	2015-02-05 21:13:25 +00:00
Colin LeMahieu	99c5ce1ce4	[Hexagon] Factoring a class out of some store patterns, deleting unused definitions and reformatting some patterns. llvm-svn: 228345	2015-02-05 20:38:58 +00:00
Colin LeMahieu	d40bb5353d	[Hexagon] Factoring out a class for immediate transfers and cleaning up formatting. llvm-svn: 228343	2015-02-05 20:08:52 +00:00
Alexey Samsonov	19763c48df	[ASan] Enable -asan-stack-dynamic-alloca by default. By default, store all local variables in dynamic alloca instead of static one. It reduces the stack space usage in use-after-return mode (dynamic alloca will not be called if the local variables are stored in a fake stack), and improves the debug info quality for local variables (they will not be described relatively to %rbp/%rsp, which are assumed to be clobbered by function calls). llvm-svn: 228336	2015-02-05 19:39:20 +00:00
Eric Christopher	24f3f65196	Remove the use of getSubtarget in the creation of the X86 PassManager instance. In one case we can make the determination from the Triple, in the other (execution dependency pass) the pass will avoid running if we don't have any code that uses that register class so go ahead and add it to the pipeline. llvm-svn: 228334	2015-02-05 19:27:04 +00:00
Eric Christopher	d361ff8282	Use cached subtargets inside X86FixupLEAs. llvm-svn: 228333	2015-02-05 19:27:01 +00:00
Eric Christopher	d7dec666cc	Migrate the X86 AsmPrinter away from using the subtarget when dealing with module level emission. Currently this is using the Triple to determine, but eventually the logic should probably migrate to TLOF. llvm-svn: 228332	2015-02-05 19:06:45 +00:00
Sylvestre Ledru	9be0b77f3f	Fix an incorrect identifier Summary: EIEIO is not a correct declaration and breaks the build under Debian HURD. Instead, E_IEIO is used. // http://www.gnu.org/software/libc/manual/html_node/Reserved-Names.html Some additional classes of identifier names are reserved for future extensions to the C language or the POSIX.1 environment. While using these names for your own purposes right now might not cause a problem, they do raise the possibility of conflict with future versions of the C or POSIX standards, so you should avoid these names. ... Names beginning with a capital ‘E’ followed a digit or uppercase letter may be used for additional error code names. See Error Reporting.// Reported here: https://bugs.debian.org/cgi-bin/bugreport.cgi?bug=776965 And patch wrote by Svante Signell With this patch, LLVM, Clang & LLDB build under Debian HURD: https://buildd.debian.org/status/fetch.php?pkg=llvm-toolchain-3.6&arch=hurd-i386&ver=1%3A3.6~%2Brc2-2&stamp=1423040039 Reviewers: hfinkel Reviewed By: hfinkel Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D7437 llvm-svn: 228331	2015-02-05 18:57:02 +00:00
Colin LeMahieu	b882f2b5cf	[Hexagon] Renaming Y2_barrier. Fixing issues where doubleword variants of instructions can't be newvalue producers. llvm-svn: 228330	2015-02-05 18:56:28 +00:00
Hal Finkel	c9dd02066c	[PowerPC] Prepare loops for pre-increment loads/stores PowerPC supports pre-increment load/store instructions (except for Altivec/VSX vector load/stores). Using these on embedded cores can be very important, but most loops are not naturally set up to use them. We can often change that, however, by placing loops into a non-canonical form. Generically, this means transforming loops like this: for (int i = 0; i < n; ++i) array[i] = c; to look like this: T p = array[-1]; for (int i = 0; i < n; ++i) ++p = c; the key point is that addresses accessed are pulled into dedicated PHIs and "pre-decremented" in the loop preheader. This allows the use of pre-increment load/store instructions without loop peeling. A target-specific late IR-level pass (running post-LSR), PPCLoopPreIncPrep, is introduced to perform this transformation. I've used this code out-of-tree for generating code for the PPC A2 for over a year. Somewhat to my surprise, running the test suite + externals on a P7 with this transformation enabled showed no performance regressions, and one speedup: External/SPEC/CINT2006/483.xalancbmk/483.xalancbmk -2.32514% +/- 1.03736% So I'm going to enable it on everything for now. I was surprised by this because, on the POWER cores, these pre-increment load/store instructions are cracked (and, thus, harder to schedule effectively). But seeing no regressions, and feeling that it is generally easier to split instructions apart late than it is to combine them late, this might be the better approach regardless. In the future, we might want to integrate this functionality into LSR (but currently LSR does not create new PHI nodes, so (for that and other reasons) significant work would need to be done). llvm-svn: 228328	2015-02-05 18:43:00 +00:00
Hal Finkel	65d1cbf9df	[PowerPC] Generate pre-increment floating-point ld/st instructions PowerPC supports pre-increment floating-point load/store instructions, both r+r and r+i, and we had patterns for them, but they were not marked as legal. Mark them as legal (and add a test case). llvm-svn: 228327	2015-02-05 18:42:53 +00:00
Colin LeMahieu	27d50073b3	[Hexagon] Renaming A2_subri, A2_andir, A2_orir. Fixing formatting. llvm-svn: 228326	2015-02-05 18:38:08 +00:00
Ahmed Bougacha	e892d13d90	[CodeGen] Add hook/combine to form vector extloads, enabled on X86. The combine that forms extloads used to be disabled on vector types, because "None of the supported targets knows how to perform load and sign extend on vectors in one instruction." That's not entirely true, since at least SSE4.1 X86 knows how to do those sextloads/zextloads (with PMOVS/ZX). But there are several aspects to getting this right. First, vector extloads are controlled by a profitability callback. For instance, on ARM, several instructions have folded extload forms, so it's not always beneficial to create an extload node (and trying to match extloads is a whole 'nother can of worms). The interesting optimization enables folding of s/zextloads to illegal (splittable) vector types, expanding them into smaller legal extloads. It's not ideal (it introduces some legalization-like behavior in the combine) but it's better than the obvious alternative: form illegal extloads, and later try to split them up. If you do that, you might generate extloads that can't be split up, but have a valid ext+load expansion. At vector-op legalization time, it's too late to generate this kind of code, so you end up forced to scalarize. It's better to just avoid creating egregiously illegal nodes. This optimization is enabled unconditionally on X86. Note that the splitting combine is happy with "custom" extloads. As is, this bypasses the actual custom lowering, and just unrolls the extload. But from what I've seen, this is still much better than the current custom lowering, which does some kind of unrolling at the end anyway (see for instance load_sext_4i8_to_4i64 on SSE2, and the added FIXME). Also note that the existing combine that forms extloads is now also enabled on legal vectors. This doesn't have a big effect on X86 (because sext+load is usually combined to sext_inreg+aextload). On ARM it fires on some rare occasions; that's for a separate commit. Differential Revision: http://reviews.llvm.org/D6904 llvm-svn: 228325	2015-02-05 18:31:02 +00:00
Andrew Trick	7fc4583eda	X86 ABI fix for return values > 24 bytes. The return value's address must be returned in %rax. i.e. the callee needs to copy the sret argument (%rdi) into the return value (%rax). This probably won't manifest as a bug when the caller is LLVM-compiled code. But it is an ABI guarantee and tools expect it. llvm-svn: 228321	2015-02-05 18:09:05 +00:00
Colin LeMahieu	f297dbed48	[Hexagon] Renaming A2_addi and formatting. llvm-svn: 228318	2015-02-05 17:49:13 +00:00
Sanjay Patel	67667bcdc4	move fold comments to the corresponding fold; NFC llvm-svn: 228317	2015-02-05 17:33:59 +00:00
Colin LeMahieu	a66cf6f2df	[Hexagon] Since decoding conflicts have been resolved, isCodeGenOnly = 0 by default and remove explicitly setting it. llvm-svn: 228316	2015-02-05 17:32:17 +00:00
Hans Wennborg	8b4dbdf15d	LowerSwitch: Use ConstantInt for CaseRange::{Low,High} Case values are always ConstantInt. This allows us to remove a bunch of casts. NFC. llvm-svn: 228312	2015-02-05 16:58:10 +00:00
Hans Wennborg	8c82fbcb73	LowerSwitch: remove default args from CaseRange ctor; NFC llvm-svn: 228311	2015-02-05 16:50:27 +00:00
Tom Stellard	eea3f70432	R600/SI: Fix bug in TTI loop unrolling preferences We should be setting UnrollingPreferences::MaxCount to MAX_UINT instead of UnrollingPreferences::Count. Count is a 'forced unrolling factor', while MaxCount sets an upper limit to the unrolling factor. Setting Count to MAX_UINT was causing the loop in the testcase to be unrolled 15 times, when it only had a maximum of 4 iterations. llvm-svn: 228303	2015-02-05 15:32:18 +00:00
Tom Stellard	0f29de78e6	R600/SI: Fix bug from insertion of llvm.SI.end.cf into loop headers The llvm.SI.end.cf intrinsic is used to mark the end of if-then blocks, if-then-else blocks, and loops. It is responsible for updating the exec mask to re-enable threads that had been masked during the preceding control flow block. For example: s_mov_b64 exec, 0x3 ; Initial exec mask s_mov_b64 s[0:1], exec ; Saved exec mask v_cmpx_gt_u32 exec, s[2:3], v0, 0 ; llvm.SI.if do_stuff() s_or_b64 exec, exec, s[0:1] ; llvm.SI.end.cf The bug fixed by this patch was one where the llvm.SI.end.cf intrinsic was being inserted into the header of loops. This would happen when an if block terminated in a loop header and we would end up with code like this: s_mov_b64 exec, 0x3 ; Initial exec mask s_mov_b64 s[0:1], exec ; Saved exec mask v_cmpx_gt_u32 exec, s[2:3], v0, 0 ; llvm.SI.if do_stuff() LOOP: ; Start of loop header s_or_b64 exec, exec, s[0:1] ; llvm.SI.end.cf <-BUG: The exec mask has the same value at the beginning of each loop iteration. do_stuff(); s_cbranch_execnz LOOP The fix is to create a new basic block before the loop and insert the llvm.SI.end.cf there. This way the exec mask is restored before the start of the loop instead of at the beginning of each iteration. llvm-svn: 228302	2015-02-05 15:32:15 +00:00
Bill Schmidt	433b1c3aae	[PowerPC] Implement the vclz instructions for PWR8 Patch by Kit Barton. Add the vector count leading zeros instruction for byte, halfword, word, and doubleword sizes. This is a fairly straightforward addition after the changes made for vpopcnt: 1. Add the correct definitions for the various instructions in PPCInstrAltivec.td 2. Make the CTLZ operation legal on vector types when using P8Altivec in PPCISelLowering.cpp Test Plan Created new test case in test/CodeGen/PowerPC/vec_clz.ll to check the instructions are being generated when the CTLZ operation is used in LLVM. Check the encoding and decoding in test/MC/PowerPC/ppc_encoding_vmx.s and test/Disassembler/PowerPC/ppc_encoding_vmx.txt respectively. llvm-svn: 228301	2015-02-05 15:24:47 +00:00
Rafael Espindola	f8d662aa4b	Add a FIXME. Thanks to Eric for the suggestion. llvm-svn: 228300	2015-02-05 14:57:47 +00:00
Aaron Ballman	94d4d33a38	Removing an unused variable warning I accidentally introduced with my last warning fix; NFC. llvm-svn: 228295	2015-02-05 13:52:42 +00:00
Aaron Ballman	1b072b340b	Silencing an MSVC warning about a switch statement with no cases; NFC. llvm-svn: 228294	2015-02-05 13:40:04 +00:00
Bruno Cardoso Lopes	ab9ae87623	[X86][MMX] Handle i32->mmx conversion using movd Implement a BITCAST dag combine to transform i32->mmx conversion patterns into a X86 specific node (MMX_MOVW2D) and guarantee that moves between i32 and x86mmx are better handled, i.e., don't use store-load to do the conversion.. llvm-svn: 228293	2015-02-05 13:23:07 +00:00
Bruno Cardoso Lopes	e446aefcfe	[X86][MMX] Move MMX DAG node to proper file llvm-svn: 228291	2015-02-05 13:22:50 +00:00
Michael Kuperstein	d2b6fdbc31	Teach isDereferenceablePointer() to look through bitcast constant expressions. This fixes a LICM regression due to the new load+store pair canonicalization. Differential Revision: http://reviews.llvm.org/D7411 llvm-svn: 228284	2015-02-05 09:15:37 +00:00
Craig Topper	0857a1da0c	[X86] Add xrstors/xsavec/xsaves/clflushopt/clwb/pcommit instructions llvm-svn: 228283	2015-02-05 08:51:06 +00:00
Craig Topper	a898c2d737	[X86] Remove two feature flags that covered sets of instructions that have no patterns or intrinsics. Since we don't check feature flags in the assembler parser for any instruction sets, these flags don't provide any value. This frees up 2 of the fully utilized feature flags. llvm-svn: 228282	2015-02-05 08:51:02 +00:00
Matt Arsenault	abd271b4e8	R600/SI: Fix i64 truncate to i1 llvm-svn: 228273	2015-02-05 06:05:13 +00:00
Larisse Voufo	bc9f12e7bc	Disable enumeral mismatch warning when compiling llvm with gcc. Tested with gcc 4.9.2. Compiling with -Werror was producing: .../llvm/lib/Target/X86/X86ISelLowering.cpp: In function 'llvm::SDValue lowerVectorShuffleAsBitMask(llvm::SDLoc, llvm::MVT, llvm::SDValue, llvm::SDValue, llvm::ArrayRef<int>, llvm::SelectionDAG&)': .../llvm/lib/Target/X86/X86ISelLowering.cpp:7771:40: error: enumeral mismatch in conditional expression: 'llvm::X86ISD::NodeType' vs 'llvm::ISD::NodeType' [-Werror=enum-compare] V = DAG.getNode(VT.isFloatingPoint() ? X86ISD::FAND : ISD::AND, DL, VT, V, ^ llvm-svn: 228271	2015-02-05 04:54:51 +00:00
Michael Zolotukhin	a9aadd2903	Implement new heuristic for complete loop unrolling. Complete loop unrolling can make some loads constant, thus enabling a lot of other optimizations. To catch such cases, we look for loads that might become constants and estimate number of instructions that would be simplified or become dead after substitution. Example: Suppose we have: int a[] = {0, 1, 0}; v = 0; for (i = 0; i < 3; i ++) v += b[i]a[i]; If we completely unroll the loop, we would get: v = b[0]a[0] + b[1]a[1] + b[2]a[2] Which then will be simplified to: v = b[0]* 0 + b[1]* 1 + b[2]* 0 And finally: v = b[1] llvm-svn: 228265	2015-02-05 02:34:00 +00:00
Cameron Esfahani	17177d1e84	Value soft float calls as more expensive in the inliner. Summary: When evaluating floating point instructions in the inliner, ask the TTI whether it is an expensive operation. By default, it's not an expensive operation. This keeps the default behavior the same as before. The ARM TTI has been updated to return back TCC_Expensive for targets which don't have hardware floating point. Reviewers: chandlerc, echristo Reviewed By: echristo Subscribers: t.p.northover, aemerson, llvm-commits Differential Revision: http://reviews.llvm.org/D6936 llvm-svn: 228263	2015-02-05 02:09:33 +00:00
Reid Kleckner	9ccec06623	Try to fix the build in MCValue.cpp llvm-svn: 228256	2015-02-05 01:23:14 +00:00
Sean Silva	32f24c49d2	Fixup. Didn't see these calls in my release build locally when testing. llvm-svn: 228254	2015-02-05 01:13:47 +00:00
Sean Silva	0e1fe184c8	[MC] Remove various unused MCAsmInfo parameters. llvm-svn: 228244	2015-02-05 00:58:51 +00:00
Duncan P. N. Exon Smith	9c26d80c18	IR: Rename 'operator ==()' to 'isKeyOf()', NFC `isKeyOf()` is a clearer name than overloading `operator==()`. llvm-svn: 228242	2015-02-05 00:51:35 +00:00
Colin LeMahieu	ceebe8659b	[Hexagon] Deleting unused instructions and adding isCodeGenOnly to some defs. llvm-svn: 228238	2015-02-05 00:10:16 +00:00
Colin LeMahieu	9cb9078ccf	[Hexagon] Updating load extend to i64 patterns. llvm-svn: 228237	2015-02-04 23:55:16 +00:00
Kostya Serebryany	92e0476c67	[fuzzer] add flag prefer_small_during_initial_shuffle, be a bit more verbose llvm-svn: 228235	2015-02-04 23:42:42 +00:00
Colin LeMahieu	712d5c393b	[Hexagon] Cleaning up i1 load and extension patterns. llvm-svn: 228232	2015-02-04 23:27:48 +00:00
Colin LeMahieu	90a91bbf43	[Hexagon] Simplifying more load and store patterns and using new addressing patterns. llvm-svn: 228231	2015-02-04 23:23:16 +00:00
Tom Stellard	f6afc80cc0	R600/SI: Enable subreg liveness by default llvm-svn: 228228	2015-02-04 23:14:18 +00:00
Colin LeMahieu	ad13d4e8a6	[Hexagon] Simplifying some load and store patterns. llvm-svn: 228227	2015-02-04 23:10:21 +00:00
Duncan P. N. Exon Smith	9c93dc6453	AsmParser: Split out LineField, NFC Split out `LineField`, which restricts the legal line numbers. This will make it easier to be consistent between different node parsers. llvm-svn: 228226	2015-02-04 22:59:18 +00:00
Colin LeMahieu	68292c96da	[Hexagon] Converting absolute-address load patterns to use AddrGP. llvm-svn: 228225	2015-02-04 22:54:51 +00:00
Colin LeMahieu	8bf5de10c3	[Hexagon] Converting atomic store/load to use AddrGP addressing. llvm-svn: 228223	2015-02-04 22:40:36 +00:00
Colin LeMahieu	5149135369	[Hexagon] Simplifying some store patterns. Adding AddrGP addressing forms. llvm-svn: 228220	2015-02-04 22:36:28 +00:00
Kostya Serebryany	33f866922a	[fuzzer] add -runs=N to limit the number of runs per session. Also, make sure we do some mutations w/o cross over. llvm-svn: 228214	2015-02-04 22:20:09 +00:00
Duncan P. N. Exon Smith	2f09c46f80	Fix GCC error caused by r228211 llvm-svn: 228213	2015-02-04 22:13:28 +00:00
Duncan P. N. Exon Smith	8af6cfcdb2	IR: Reduce boilerplate in DenseMapInfo overrides, NFC Minimize the boilerplate required for the `MDNode` subclass `DenseMapInfo<>` overrides in `LLVMContextImpl`. llvm-svn: 228212	2015-02-04 22:08:30 +00:00
Duncan P. N. Exon Smith	077c0318dc	AsmParser: Move MDField details to source file, NFC Move all the types of `MDField` to an anonymous namespace in the source file. This also eliminates the duplication of `ParseMDField()` declarations in the header for each new field type. llvm-svn: 228211	2015-02-04 22:05:21 +00:00
Duncan P. N. Exon Smith	fad631a90c	AsmParser: Simplify assertion, NFC llvm-svn: 228209	2015-02-04 22:02:18 +00:00
Duncan P. N. Exon Smith	7a968781a6	AsmParser: Remove dead code, NFC This condition is checked in the generic `ParseMDField()`. llvm-svn: 228208	2015-02-04 22:00:59 +00:00
Duncan P. N. Exon Smith	39b10c2cbb	AsmParser: Simplify MDUnsignedField We only need `uint64_t` for storage. llvm-svn: 228205	2015-02-04 21:57:52 +00:00
Duncan P. N. Exon Smith	6b7b680efd	IR: Initialize MDNode abbreviations en masse, NFC llvm-svn: 228203	2015-02-04 21:54:12 +00:00
Duncan P. N. Exon Smith	104e402610	IR: Define MDNode uniquing sets automatically, NFC llvm-svn: 228200	2015-02-04 21:46:12 +00:00
Rafael Espindola	a092f17580	Don' try to make sections in comdats SHF_MERGE. Parts of llvm were not expecting it and we wouldn't print the entity size of the section. Given what comdats are used for, having SHF_MERGE sections would be just a small improvement, so just disable it for now. Fixes pr22463. llvm-svn: 228196	2015-02-04 21:27:24 +00:00
Tom Stellard	33e64c66ac	R600/SI: Expand misaligned 16-bit memory accesses llvm-svn: 228190	2015-02-04 20:49:52 +00:00
Tom Stellard	c7e448c92e	R600/SI: Make more store operations legal v2i32, i32, trunc i32 to i16, and truc i32 to i8 stores are legal for all address spaces. We had marked them as custom in order to lower them for the private address space, but this is no longer necessary. This enables lowering of misaligned stores of these types in the DAGLegalizer. llvm-svn: 228189	2015-02-04 20:49:51 +00:00
Tom Stellard	096b8c1e6d	R600: Don't promote i64 stores to v2i32 during DAG legalization We take care of this during instruction selection now. This fixes a potential infinite loop when lowering misaligned stores. llvm-svn: 228188	2015-02-04 20:49:49 +00:00
Tom Stellard	080209d573	StructurizeCFG: Remove obsolete fix for loop backedge detection This is no longer needed now that we are using a reverse post-order traversal. llvm-svn: 228187	2015-02-04 20:49:47 +00:00
Tom Stellard	071ec90b68	StructurizeCFG: Use a reverse post-order traversal We were previously doing a post-order traversal and operating on the list in reverse, however this would occasionaly cause backedges for loops to be visited before some of the other blocks in the loop. We know use a reverse post-order traversal, which avoids this issue. The reverse post-order traversal is not completely ideal, so we need to manually fixup the list to ensure that inner loop backedges are visited before outer loop backedges. llvm-svn: 228186	2015-02-04 20:49:44 +00:00
Colin LeMahieu	987b0943c8	[Hexagon] Adding selection for GlobalAddress and converting [z/i]ext load patterns to make use of them. llvm-svn: 228184	2015-02-04 20:38:01 +00:00
Duncan P. N. Exon Smith	920df5c1bb	Utils: Resolve cycles under distinct MDNodes Track unresolved nodes under distinct `MDNode`s during `MapMetadata()`, and resolve them at the end. Previously, these cycles wouldn't get resolved. llvm-svn: 228180	2015-02-04 19:44:34 +00:00
Matthias Braun	26e7ea6267	MachineCSE: Clear dead-def flag on CSE. In case CSE reuses a previoulsy unused register the dead-def flag has to be cleared on the def operand, as exposed by the arm64-cse.ll test. This fixes PR22439 and the corresponding rdar://19694987 Differential Revision: http://reviews.llvm.org/D7395 llvm-svn: 228178	2015-02-04 19:35:16 +00:00
Reid Kleckner	c26a17a822	Add range adapters predecessors() and successors() for BBs Use them in two isolated transforms so we know they work and aren't dead code. llvm-svn: 228173	2015-02-04 19:14:57 +00:00
Kostya Serebryany	5b266a8a23	[fuzzer] make multi-process execution more verbose; fix mutation to actually respect mutation depth and to never produce empty units llvm-svn: 228170	2015-02-04 19:10:20 +00:00
Colin LeMahieu	86abe35ceb	[Hexagon] Replacing some load patterns with cleaner versions. llvm-svn: 228169	2015-02-04 19:05:32 +00:00
Michael Kuperstein	cd63c5fa73	Fixes a bug in vector load legalization that confused bits and bytes. Differential Revision: http://reviews.llvm.org/D7400 llvm-svn: 228168	2015-02-04 18:54:01 +00:00
Colin LeMahieu	f856dcb75e	[Hexagon] Adding missing isCodeGenOnly = 0 llvm-svn: 228160	2015-02-04 18:11:32 +00:00
Colin LeMahieu	c0434466e4	[Hexagon] Adding encoding information for absolute-reg mode stores. Xfailing a test until constant extenders are correctly put in the same packet. llvm-svn: 228158	2015-02-04 17:52:06 +00:00

... 3 4 5 6 7 ...

76887 Commits