llvm-project

Commit Graph

Author	SHA1	Message	Date
Matheus Almeida	be8681b461	[mips][msa] Direct Object Emission support for the LSA instruction. llvm-svn: 193240	2013-10-23 13:20:07 +00:00
Daniel Sanders	a952160078	[mips][msa] Added support for matching fexp2 from normal IR (i.e. not intrinsics) llvm-svn: 193239	2013-10-23 10:36:52 +00:00
Artyom Skrobov	fc12e7016c	Make ARM hint ranges consistent, and add tests for these ranges llvm-svn: 193238	2013-10-23 10:14:40 +00:00
Tom Stellard	03a5c08de6	R600/SI: Replace ffs(x) - 1 with countTrailingZeros(x) ffs(x) broke the mingw buildbot. llvm-svn: 193225	2013-10-23 03:50:25 +00:00
Tom Stellard	54774e5681	R600/SI: fix MIMG writemask adjustement This fixes piglit: - shaders/glsl-fs-texture2d-masked - shaders/glsl-fs-texture2d-masked-4 Patch by: Marek Olšák Signed-off-by: Marek Olšák <marek.olsak@amd.com> Reviewed-by: Tom Stellard <thomas.stellard@amd.com> llvm-svn: 193222	2013-10-23 02:53:47 +00:00
Tom Stellard	af77543244	R600: Fix handling of vector kernel arguments The SelectionDAGBuilder was promoting vector kernel arguments to legal types, but this won't work for R600 and SI since kernel arguments are stored in memory and can't be promoted. In order to handle vector arguments correctly we need to look at the original types from the LLVM IR function. llvm-svn: 193215	2013-10-23 00:44:32 +00:00
Tom Stellard	fb9616905a	R600/SI: Add support for i64 bitwise or llvm-svn: 193213	2013-10-23 00:44:19 +00:00
Tom Stellard	a66cafa096	R600/SI: Use S_LOAD_DWORD instructions for v8i32 and v16i32 llvm-svn: 193212	2013-10-23 00:44:12 +00:00
Quentin Colombet	f34568b0af	[X86][FastISel] Add a comment to help understanding changes made in r192636. <rdar://problem/15192473> llvm-svn: 193199	2013-10-22 21:29:08 +00:00
Matt Arsenault	65864e3182	R600/SI: Don't assert on SCC usage llvm-svn: 193198	2013-10-22 21:11:31 +00:00
Tim Northover	08a8660260	ARM: provide diagnostics on more writeback LDM/STM instructions The set of circumstances where the writeback register is allowed to be in the list of registers is rather baroque, but I think this implements them all on the assembly parsing side. For disassembly, we still warn about an ARM-mode LDM even if the architecture revision is < v7 (the required architecture information isn't available). It's a silly instruction anyway, so hopefully no-one will mind. rdar://problem/15223374 llvm-svn: 193185	2013-10-22 19:00:39 +00:00
Tom Stellard	debb4cf5ea	R600/SI: Use llvm_unreachable() for an always false assert llvm-svn: 193183	2013-10-22 18:42:03 +00:00
Tom Stellard	8be4dd234a	R600/SI: Fix warning on non-asserts build llvm-svn: 193180	2013-10-22 18:31:45 +00:00
Tom Stellard	26a3b67b3b	R600: Simplify handling of private address space The AMDGPUIndirectAddressing pass was previously responsible for lowering private loads and stores to indirect addressing instructions. However, this pass was buggy and way too complicated. The only advantage it had over the new simplified code was that it saved one instruction per direct write to private memory. This optimization likely has a minimal impact on performance, and we may be able to duplicate it using some other transformation. For the private address space, we now: 1. Lower private loads/store to Register(Load\|Store) instructions 2. Reserve part of the register file as 'private memory' 3. After regalloc lower the Register(Load\|Store) instructions to MOV instructions that use indirect addressing. llvm-svn: 193179	2013-10-22 18:19:10 +00:00
Tom Stellard	c460b0dcf1	R600: Remove unused InstrInfo::getMovImmInstr() function llvm-svn: 193178	2013-10-22 18:19:01 +00:00
Matheus Almeida	eb68d9d985	[mips][msa] Direct Object Emission support for conditional branches. These branches have a 16-bit offset (R_MIPS_PC16). List of conditional branch instructions: bnz.{b,h,w,d} bnz.v bz.{b,h,w,d} bz.v llvm-svn: 193157	2013-10-22 09:43:32 +00:00
Elena Demikhovsky	1f3ed4169c	AVX-512: aligned / unaligned load and store for 512-bit integer vectors. llvm-svn: 193156	2013-10-22 09:19:28 +00:00
Craig Topper	f7290f7194	Replace (V)MOVZDI2PDIrr/rm instructions with patterns that select (V)MOVDI2PDIrr/rm. llvm-svn: 193146	2013-10-22 04:35:20 +00:00
Jim Grosbach	dba14ddd4f	ARM: Thumb2 copy for GPRPair needs to use thumb instructions. Use tMOVr instead of plain MOVr. rdar://15193017 llvm-svn: 193139	2013-10-22 02:29:37 +00:00
Jim Grosbach	8815bef000	ARM: Clean up copyPhysReg() a bit. No functional change, just cleaning things up for readability. llvm-svn: 193138	2013-10-22 02:29:35 +00:00
Chad Rosier	e012cb3783	[AArch64] Add the constraint to NEON scalar mla/mls instructions. llvm-svn: 193117	2013-10-21 20:11:47 +00:00
Lang Hames	2783993fca	X86 vector element shift-by-immediate instructions take i8 immediates. Make the instruction defenitions and ISEL reflect this. Prior to this patch these instructions took an i32i8imm, and the high bits were dropped during encoding. This led to incorrect behavior for shifts by immediates higher than 255. This patch fixes that issue by detecting large immediate shifts and returning constant zero (for logical shifts) or capping the shift amount at an encodable value (for arithmetic shifts). Fixes <rdar://problem/14968098> llvm-svn: 193096	2013-10-21 17:51:24 +00:00
Elena Demikhovsky	665c90e184	AVX-512: MUL operation lowering for v8i64 llvm-svn: 193083	2013-10-21 13:27:34 +00:00
Matheus Almeida	fe0bf9f618	[mips][msa] Direct Object Emission support for LD/ST instructions. llvm-svn: 193082	2013-10-21 13:07:13 +00:00
Matheus Almeida	8ddad15177	[mips][msa] Direct Object Emission support for LDI instructions. llvm-svn: 193081	2013-10-21 12:56:20 +00:00
Matheus Almeida	83d797de4a	[mips][msa] Direct Object Emission support for MOVE.v. llvm-svn: 193080	2013-10-21 12:43:54 +00:00
Matheus Almeida	a591fdc63c	[mips][msa] Direct Object Emission support for CTCMSA and CFCMSA. These instructions are logically related as they allow read/write of MSA control registers. Currently MSA control registers are emitted by number but hopefully that will change as soon as GAS starts accepting them by name as that would make the assembly easier to read. llvm-svn: 193078	2013-10-21 12:26:50 +00:00
Matheus Almeida	5798c6f3bb	[mips][msa] Direct Object Emission of SPLAT instruction. llvm-svn: 193077	2013-10-21 12:07:26 +00:00
Matheus Almeida	70fbf77546	[mips][msa] Fix definition of SLD instruction. The second parameter of the SLD intrinsic is the number of columns (GPR) to slide left the source array. llvm-svn: 193076	2013-10-21 11:47:56 +00:00
Nadav Rotem	7f27e0b0ce	Mark some command line flags as hidden llvm-svn: 193013	2013-10-18 23:38:13 +00:00
Hans Wennborg	ce69d77cec	MC asm parser: allow ?'s in symbol names, and handle @'s in names in MS asm This is another (final?) stab at making us able to parse our own asm output on Windows. Symbols on Windows often contain @'s and ?'s in their names. Our asm parser didn't like this. ?'s were not allowed, and @'s were intepreted as trying to reference PLT/GOT/etc. We can't just add quotes around the bad names, since e.g. for MinGW, we use gas to assemble, and it doesn't like quotes in some places (notably in .def directives). This commit makes us allow ?'s in symbol names, and @'s in symbol names for MS assembly. Differential Revision: http://llvm-reviews.chandlerc.com/D1978 llvm-svn: 193000	2013-10-18 20:46:28 +00:00
Richard Barton	a661b44a5d	Pure refactoring change. Patch by Artyom Skrobov. llvm-svn: 192977	2013-10-18 14:41:50 +00:00
Benjamin Kramer	a9fe95b6c2	R600: Remove \ at EOL from ascii art comments. Completely harmless, but GCC likes to warn about it even when the next line is a comment. llvm-svn: 192974	2013-10-18 14:12:50 +00:00
Richard Barton	87dacc38b8	Add hint disassembly syntax for 16-bit Thumb hint instructions. Patch by Artyom Skrobov llvm-svn: 192972	2013-10-18 14:09:49 +00:00
Chad Rosier	fe2f58c8a1	[AArch64] Add support for NEON scalar extract narrow instructions. llvm-svn: 192970	2013-10-18 14:03:24 +00:00
Silviu Baranga	314e58fdcc	Add hardware division as a default feature on Cortex-A15. Also add test cases to check this, and change diagnostics for the hwdiv-arm feature to something useful. llvm-svn: 192963	2013-10-18 10:18:40 +00:00
Hans Wennborg	7ddcdc82a5	Revert "Re-commit r192758 - MC: quote tricky symbol names in asm output" This caused the clang-native-mingw32-win7 buildbot to break. The assembler was complaining about the following lines that were showing up in the asm for CrashRecoveryContext.cpp: movl $"__ZL16ExceptionHandlerP19_EXCEPTION_POINTERS@4", 4(%eax) calll "_AddVectoredExceptionHandler@8" .def "__ZL16ExceptionHandlerP19_EXCEPTION_POINTERS@4"; "__ZL16ExceptionHandlerP19_EXCEPTION_POINTERS@4": calll "_RemoveVectoredExceptionHandler@4" Reverting for now. llvm-svn: 192940	2013-10-18 02:14:40 +00:00
David Peixotto	8e5abc52cb	17309 ARM backend incorrectly lowers COPY_STRUCT_BYVAL_I32 for thumb1 targets This commit implements the correct lowering of the COPY_STRUCT_BYVAL_I32 pseudo-instruction for thumb1 targets. Previously, the lowering of COPY_STRUCT_BYVAL_I32 generated the post-increment forms of ldr/ldrh/ldrb instructions. Thumb1 does not have the post-increment form of these instructions so the generated assembly contained invalid instructions. Passing the generated assembly to gcc caused it to complain with an error like this: Error: cannot honor width suffix -- `ldrb r3,[r0],#1' and the integrated assembler would generate an object file with an invalid instruction encoding. This commit contains a small test case that demonstrates the problem with thumb1 targets as well as an expanded test case that more throughly tests the lowering of byval struct passing for arm, thumb1, and thumb2 targets. llvm-svn: 192916	2013-10-17 19:52:05 +00:00
David Peixotto	c32e24a1b7	Refactor lowering for COPY_STRUCT_BYVAL_I32 This commit refactors the lowering of the COPY_STRUCT_BYVAL_I32 pseudo-instruction in the ARM backend. We introduce a new helper class that encapsulates all of the operations needed during the lowering. The operations are implemented for each subtarget in different subclasses. Currently only arm and thumb2 subtargets are supported. This refactoring was done to easily implement support for thumb1 subtargets. This initial patch does not add support for thumb1, but is only a refactoring. A follow on patch will implement the support for thumb1 subtargets. No intended functionality change. llvm-svn: 192915	2013-10-17 19:49:22 +00:00
Anders Waldenborg	959f04077c	llvm-c: Add LLVMIntPtrType{,ForAS}InContext All of the Core API functions have versions which accept explicit context, in addition to ones which work on global context. This commit adds functions which accept explicit context to the Target API for consistency. Patch by Peter Zotov Differential Revision: http://llvm-reviews.chandlerc.com/D1912 llvm-svn: 192913	2013-10-17 18:51:01 +00:00
Chad Rosier	37d29173aa	[AArch64] Add support for NEON scalar three register different instruction class. The instruction class includes the signed saturating doubling multiply-add long, signed saturating doubling multiply-subtract long, and the signed saturating doubling multiply long instructions. llvm-svn: 192908	2013-10-17 18:12:29 +00:00
Daniel Sanders	a4eaf59f9e	[mips][msa] Added lsa instruction llvm-svn: 192895	2013-10-17 13:38:20 +00:00
Daniel Sanders	0390568a09	[mips][msa] Removed ldx.[bhwd] and stx.[bhwd]. These were present in a previous version of the MSA spec but are not present in the published version. There is no hardware that uses these instructions. llvm-svn: 192888	2013-10-17 12:16:03 +00:00
Anders Waldenborg	39f5d7d5f0	llvm-c: Don't assert in LLVMTargetMachineEmitToFile on nonexistent file Error handling code for raw_fd_ostream constructor is present, but never used, because formatted_raw_ostream will always assert on closed fd's before. Patch by Peter Zotov Differential Revision: http://llvm-reviews.chandlerc.com/D1909 llvm-svn: 192881	2013-10-17 10:39:35 +00:00
Daniel Sanders	199b731b42	[mips][msa] Correct definition order of ftrunc_[su], ftint_[su], and ftq. Define these three instructions in alphabetical order (like the rest of the file). No functional change. llvm-svn: 192880	2013-10-17 10:30:12 +00:00
Anders Waldenborg	a89c1e3145	llvm-c: Return NULL from LLVMGetFirstTarget instead of asserting If no targets are registered, LLVMGetFirstTarget currently fails with an assertion. This patch makes it return NULL instead, similarly to how LLVMGetNextTarget would. Patch by Peter Zotov Differential Revision: http://llvm-reviews.chandlerc.com/D1908 llvm-svn: 192878	2013-10-17 10:25:24 +00:00
Jim Grosbach	c044c65470	x86: Move bitcasts outside concat_vector. Consider the following: typedef unsigned short ushort4U __attribute__((ext_vector_type(4), aligned(2))); typedef unsigned short ushort4 __attribute__((ext_vector_type(4))); typedef unsigned short ushort8 __attribute__((ext_vector_type(8))); typedef int int4 __attribute__((ext_vector_type(4))); int4 __bbase_cvt_int(ushort4 v) { ushort8 a; a.lo = v; return _mm_cvtepu16_epi32(a); } This generates the, not unreasonable, IR: define <4 x i32> @foo0(double %v.coerce) nounwind ssp { %tmp = bitcast double %v.coerce to <4 x i16> %tmp1 = shufflevector <4 x i16> %tmp, <4 x i16> undef, <8 x i32> <i32 %0, i32 1, i32 2, i32 3, i32 undef, i32 undef, i32 undef, i32 undef> %tmp2 = tail call <4 x i32> @llvm.x86.sse41.pmovzxwd(<8 x i16> %tmp1) ret <4 x i32> %tmp2 } The problem is when type legalization gets hold of the v4i16. It legalizes that by spilling to the stack, then doing a zero-extending load. Things go even more silly from there, ending up with something like: _foo0: movsd %xmm0, -8(%rsp) <== Spill to the stack. movq -8(%rsp), %xmm0 <== Reload it right back out. pmovzxwd %xmm0, %xmm1 <== Here's what we actually asked for. pblendw $1, %xmm1, %xmm0 <== We don't need this at all pmovzxwd %xmm0, %xmm0 <== We already did this ret The v8i8 to v8i16 zext intrinsic gives even worse results, with two table lookups via pshufb instructions(!!). To avoid all that, we can move the bitcasting until after we've formed the wider (legal) vector type. Then our normal codegen flows along nicely and we get the expected: _foo0: pmovzxwd %xmm0, %xmm0 ret rdar://15245794 llvm-svn: 192866	2013-10-17 02:58:06 +00:00
Hans Wennborg	69918bccab	Re-commit r192758 - MC: quote tricky symbol names in asm output The reason this got reverted was that the @feat.00 symbol which was emitted for every TU became quoted, and on cygwin/mingw we use the gas assembler which couldn't handle the quotes. This commit fixes the problem by only emitting @feat.00 for win32, where we use clang -cc1as to assemble. gas would just drop this symbol anyway, so there is no loss there. With @feat.00 gone, there shouldn't be quoted symbols showing up on cygwin since it uses the Itanium ABI, which doesn't put these funny characters in symbols. > Because of win32 mangling, we produce symbol and section names with > funny characters in them, most notably @ characters. > > MC would choke on trying to parse its own assembly output. This patch addresses > that by: > > - Making @ trigger quoting of symbol names > - Also quote section names in the same way > - Just parse section names like other identifiers (to allow for quotes) > - Don't assume @ signifies a symbol variant if it is in a string. llvm-svn: 192859	2013-10-17 01:13:02 +00:00
Chad Rosier	846a72539c	[AArch64] Add support for NEON scalar negate instruction. llvm-svn: 192843	2013-10-16 21:04:39 +00:00
Chad Rosier	175601d997	[AArch64] Add support for NEON scalar absolute value instruction. llvm-svn: 192842	2013-10-16 21:04:34 +00:00

1 2 3 4 5 ...

25950 Commits