llvm-project

Commit Graph

Author	SHA1	Message	Date
Hal Finkel	57c6ac5e41	[PowerPC] Support the (old) cntlz instruction alias Some old assembly code uses the cntlz alias for cntlzw, binutils supports this, and we should too. Fixes PR22519. llvm-svn: 228719	2015-02-10 18:45:02 +00:00
Colin LeMahieu	404d5b242d	[Hexagon] Adding vector load with post-increment instructions. Adding decoder function for 64bit control register class. llvm-svn: 228708	2015-02-10 16:59:36 +00:00
Zoran Jovanovic	416886793f	[mips][microMIPS] Implement movep instruction Differential Revision: http://reviews.llvm.org/D7465 llvm-svn: 228703	2015-02-10 16:36:20 +00:00
Simon Pilgrim	d142ab7d08	[X86][AVX2] Missing AVX2 memory folding instructions Added most of the missing vector folding patterns for AVX2 (as well as fixing the vpermpd and verpmq patterns) Differential Revision: http://reviews.llvm.org/D7492 llvm-svn: 228688	2015-02-10 13:22:57 +00:00
Simon Pilgrim	cd32254a35	[X86][XOP] Added XOP memory folding patterns + tests This patch adds the complete AMD Bulldozer XOP instruction set to the memory folding pattern tables for stack folding, etc. Note: Many of the XOP instructions have multiple table entries as it can fold loads from different sources. Differential Revision: http://reviews.llvm.org/D7484 llvm-svn: 228685	2015-02-10 12:57:17 +00:00
Jozef Kolek	d68d424abf	[mips][microMIPS] Fix disassembling of 16-bit microMIPS instructions LWM16 and SWM16 Differential Revision: http://reviews.llvm.org/D7436 llvm-svn: 228683	2015-02-10 12:41:13 +00:00
Andrea Di Biagio	62622d2396	[X86][FastIsel] Avoid introducing legacy SSE instructions if the target has AVX. This patch teaches X86FastISel how to select AVX instructions for scalar float/double convert operations. Before this patch, X86FastISel always selected legacy SSE instructions for FPExt (from float to double) and FPTrunc (from double to float). For example: \code define double @foo(float %f) { %conv = fpext float %f to double ret double %conv } \end code Before (with -mattr=+avx -fast-isel) X86FastIsel selected a CVTSS2SDrr which is legacy SSE: cvtss2sd %xmm0, %xmm0 With this patch, X86FastIsel selects a VCVTSS2SDrr instead: vcvtss2sd %xmm0, %xmm0, %xmm0 Added test fast-isel-fptrunc-fpext.ll to check both the register-register and the register-memory float/double conversion variants. Differential Revision: http://reviews.llvm.org/D7438 llvm-svn: 228682	2015-02-10 12:04:41 +00:00
Craig Topper	9e71b82f40	[X86] Preserve mem refs on newly created 'Store' node instead of 'Load' node when handling store unfolding. Bug spotted by Steve King. I have no idea how to test this. llvm-svn: 228672	2015-02-10 06:29:28 +00:00
Craig Topper	f7e92f10b6	[X86] Remove unnecessary alignment checks from the load folding tables. llvm-svn: 228671	2015-02-10 05:10:50 +00:00
David Majnemer	93c22a45be	X86: Emit an ABI compliant prologue and epilogue for Win64 Win64 has specific contraints on what valid prologues and epilogues look like. This constraint is born from the flexibility and descriptiveness of Win64's unwind opcodes. Prologues previously emitted by LLVM could not be represented by the unwind opcodes, preventing operations powered by stack unwinding to successfully work. Differential Revision: http://reviews.llvm.org/D7520 llvm-svn: 228641	2015-02-10 00:57:42 +00:00
Eric Christopher	d49868080e	Migrate PPCAsmPrinter's subtarget from reference to pointer in preparation for making it MachineFunction dependent. llvm-svn: 228638	2015-02-10 00:44:17 +00:00
David Blaikie	36a036909c	Fix the clang -Werror build (-Wunused-variable) llvm-svn: 228635	2015-02-10 00:16:36 +00:00
Colin LeMahieu	328b1633d7	[Hexagon] Adding missing load instructions and removing an unused multiclass parameter. llvm-svn: 228630	2015-02-09 23:45:24 +00:00
Colin LeMahieu	4282e7cffd	[Hexagon] Factoring classes out of some load patterns and deleting some unused ones. llvm-svn: 228627	2015-02-09 23:05:44 +00:00
Colin LeMahieu	4fd203d3e1	[Hexagon] Removing more V4 predicates since V4 is the required minimum. llvm-svn: 228614	2015-02-09 21:56:37 +00:00
Colin LeMahieu	641c24b9bf	[Hexagon] Removing v2-4 flags. V4 is the minimum supported version. llvm-svn: 228605	2015-02-09 21:07:35 +00:00
Colin LeMahieu	955c4ff9c3	[Hexagon] Factoring classes out of store patterns. llvm-svn: 228602	2015-02-09 20:33:46 +00:00
Colin LeMahieu	ab5a8d6070	[Hexagon] Formatting v5 TD file. Removing commented defs. llvm-svn: 228598	2015-02-09 20:03:42 +00:00
Colin LeMahieu	38e6689276	[Hexagon] Cleaning up definition formatting. llvm-svn: 228593	2015-02-09 19:24:44 +00:00
Kit Barton	0b0cdb1cd4	This change implements the following three logical vector operations: veqv (vector equivalence) vnand vorc I increased the AddedComplexity for these instructions to 500 to ensure they are generated instead of issuing other VSX instructions. Phabricator review: http://reviews.llvm.org/D7469 llvm-svn: 228580	2015-02-09 17:03:18 +00:00
Sanjay Patel	a7b893d5c0	rename variable to give it some meaning; remove obvious comments; NFC llvm-svn: 228579	2015-02-09 16:30:58 +00:00
Sanjay Patel	fc54c61c56	fix comment that didn't match the code; remove unnecessary braces; NFC llvm-svn: 228578	2015-02-09 16:04:52 +00:00
Craig Topper	141e65e69c	[X86] Remove 256-bit and 512-bit memop pattern fragments. They are no longer used. llvm-svn: 228563	2015-02-09 04:04:53 +00:00
Craig Topper	820d49270d	[X86] Remove 'memop' uses from AVX512. Use 'load' instead. llvm-svn: 228562	2015-02-09 04:04:50 +00:00
Craig Topper	68ab0465a0	[X86] Remove the remaining uses of memop from AVX and AVX2 instruction patterns. AVX and AVX2 can handle unaligned loads being folded so we can just use 'load' llvm-svn: 228551	2015-02-08 22:38:25 +00:00
Sanjay Patel	3510bc7162	fix typos; NFC llvm-svn: 228529	2015-02-08 18:54:22 +00:00
Simon Pilgrim	d11b013623	Moved AVX2 vbroadcast (reg) instruction foldings under the correct grouping. NFC. llvm-svn: 228526	2015-02-08 17:13:54 +00:00
Tim Northover	45aa89c925	ARM & AArch64: teach LowerVSETCC that output type size may differ from input. While various DAG combines try to guarantee that a vector SETCC operation will have the same output size as input, there's nothing intrinsic to either creation or LegalizeTypes that actually guarantees it, so the function needs to be ready to handle a mismatch. Fortunately this is easy enough, just extend or truncate the naturally compared result. I couldn't reproduce the failure in other backends that I know have SIMD, so it's probably only an issue for these two due to shared heritage. Should fix PR21645. llvm-svn: 228518	2015-02-08 00:50:47 +00:00
Craig Topper	e169c57188	[X86] Add register use/def for wrmsr and rdmsr. llvm-svn: 228515	2015-02-07 23:36:51 +00:00
Craig Topper	1d472db8cc	[X86] Add GETSEC instruction. llvm-svn: 228514	2015-02-07 23:36:36 +00:00
Simon Pilgrim	a2618679a8	[X86][AVX] Added missing stack folding support + test for vptest ymm instruction llvm-svn: 228509	2015-02-07 21:44:06 +00:00
Andrea Di Biagio	4f8bdcb738	Fix typos; NFC. llvm-svn: 228493	2015-02-07 13:56:20 +00:00
Hal Finkel	291cc7bacd	[PowerPC] Handle loop predecessor invokes If a loop predecessor has an invoke as its terminator, and the return value from that invoke is used to determine the loop iteration space, then we can't insert a computation based on that value in the loop predecessor prior to the terminator (oops). If there's such an invoke, or just no predecessor for that matter, insert a new loop preheader. llvm-svn: 228488	2015-02-07 07:32:58 +00:00
Ahmed Bougacha	df956a2e78	[AArch64] Use the source location of the IR branch when creating Bcc from a conditional branch fed by an add/sub/mul-with-overflow node. We previously used the SDLoc of the overflow node, for no good reason. In some cases, this led to the Bcc and B terminators having different source orders, and DBG_VALUEs being inserted between them. The real issue is with the code that can't handle DBG_VALUEs between terminators: the few places affected by this will be fixed soon. In the meantime, fixing the SDLoc is a positive change no matter what. No tests, as I have no idea how to get .loc emitted for branches? rdar://19347133 llvm-svn: 228463	2015-02-06 23:15:39 +00:00
Hal Finkel	0d2a1515d5	Revert "r227976 - [PowerPC] Yet another approach to __tls_get_addr" and related fixups Unfortunately, even with the workaround of disabling the linker TLS optimizations in Clang restored (which has already been done), this still breaks self-hosting on my P7 machine (-O3 -DNDEBUG -mcpu=native). Bill is currently working on an alternate implementation to address the TLS issue in a way that also fully elides the linker bug (which, unfortunately, this approach did not fully), so I'm reverting this now. llvm-svn: 228460	2015-02-06 23:07:40 +00:00
Sanjay Patel	3d982214b0	use local variables; NFC llvm-svn: 228452	2015-02-06 22:43:52 +00:00
Cameron Esfahani	f0bf1947b7	Test commit to see if it triggers an email to llvm-commits. No change. llvm-svn: 228442	2015-02-06 21:33:08 +00:00
Reid Kleckner	526ec29370	Don't dllexport declarations Fixes PR22488 llvm-svn: 228411	2015-02-06 17:59:49 +00:00
Benjamin Kramer	970eac40bf	Make helper functions/classes/globals static. NFC. llvm-svn: 228410	2015-02-06 17:51:54 +00:00
Benjamin Kramer	69b4ad29af	AArch64PromoteConstant: Modernize and resolve some Use<->User confusion. NFC. llvm-svn: 228399	2015-02-06 14:43:55 +00:00
Michel Danzer	d89a557f33	R600/SI: Don't enable WQM for V_INTERP_* instructions v2 Doesn't seem necessary anymore. I think this was mostly compensating for not enabling WQM for texture sampling instructions. v2: Add test coverage Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228373	2015-02-06 02:51:25 +00:00
Michel Danzer	494391ba47	R600/SI: Also enable WQM for image opcodes which calculate LOD v3 If whole quad mode isn't enabled for these, the level of detail is calculated incorrectly for pixels along diagonal triangle edges, causing artifacts. v2: Use a TSFlag instead of lots of switch cases v3: Add test coverage Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=88642 Reviewed-by: Tom Stellard <tom@stellard.net> llvm-svn: 228372	2015-02-06 02:51:20 +00:00
Colin LeMahieu	6e3e62fd13	[Hexagon] Renaming v4 compare-and-jump instructions. llvm-svn: 228349	2015-02-05 22:03:32 +00:00
Colin LeMahieu	6b3aac8ad9	[Hexagon] Deleting unused patterns. llvm-svn: 228348	2015-02-05 21:43:56 +00:00
Colin LeMahieu	de68b66b4d	[Hexagon] Simplifying and formatting several patterns. Changing a pattern multiply to be expanded. llvm-svn: 228347	2015-02-05 21:13:25 +00:00
Colin LeMahieu	99c5ce1ce4	[Hexagon] Factoring a class out of some store patterns, deleting unused definitions and reformatting some patterns. llvm-svn: 228345	2015-02-05 20:38:58 +00:00
Colin LeMahieu	d40bb5353d	[Hexagon] Factoring out a class for immediate transfers and cleaning up formatting. llvm-svn: 228343	2015-02-05 20:08:52 +00:00
Eric Christopher	24f3f65196	Remove the use of getSubtarget in the creation of the X86 PassManager instance. In one case we can make the determination from the Triple, in the other (execution dependency pass) the pass will avoid running if we don't have any code that uses that register class so go ahead and add it to the pipeline. llvm-svn: 228334	2015-02-05 19:27:04 +00:00
Eric Christopher	d361ff8282	Use cached subtargets inside X86FixupLEAs. llvm-svn: 228333	2015-02-05 19:27:01 +00:00
Eric Christopher	d7dec666cc	Migrate the X86 AsmPrinter away from using the subtarget when dealing with module level emission. Currently this is using the Triple to determine, but eventually the logic should probably migrate to TLOF. llvm-svn: 228332	2015-02-05 19:06:45 +00:00

1 2 3 4 5 ...

31855 Commits