llvm-project

Commit Graph

Author	SHA1	Message	Date
Eric Christopher	fc2beaa136	Recommit r179497 after fixing uninitialized variable. llvm-svn: 179512	2013-04-15 07:07:21 +00:00
Nadav Rotem	a440a356e0	Document our desire to enable the loop vectorizer on -Os in future releases. llvm-svn: 179511	2013-04-15 05:56:55 +00:00
Nadav Rotem	57da1fdd55	Docs: merge the description of the BB and SLP vectorizers and document the -fslp-vectorize-aggressive flag. llvm-svn: 179510	2013-04-15 05:53:23 +00:00
Nadav Rotem	d4dcc003df	Add an option -vectorize-slp-aggressive for running the BB vectorizer. Make -fslp-vectorize run the slp-vectorizer. llvm-svn: 179508	2013-04-15 05:39:58 +00:00
Nadav Rotem	a1e5e44eb3	Rename the slp-vectorizer clang/llvm flags. No functionality change. llvm-svn: 179505	2013-04-15 04:54:42 +00:00
Nadav Rotem	5d393c416f	SLPVectorizer: Add support for vectorizing trees that start at compare instructions. llvm-svn: 179504	2013-04-15 04:25:27 +00:00
Jia Liu	f3076492c2	fix include path in doc Extending LLVM llvm-svn: 179503	2013-04-15 03:26:13 +00:00
Hal Finkel	95e6ea69be	Mark all PPC comparison instructions as not having side effects Now that the CR spilling issues have been resolved, we can remove the unmodeled-side-effect attributes from the comparison instructions (and also mark them as isCompare). By allowing these, by default, to have unmodeled side effects, we were hiding problems with CR spilling; but everything seems much happier now. llvm-svn: 179502	2013-04-15 02:37:46 +00:00
Hal Finkel	6736988ae2	Fix PPC64 CR spill location for callee-saved registers This fixes an ABI bug for non-Darwin PPC64. For the callee-saved condition registers, the spill location is specified relative to the stack pointer (SP + 8). However, this is not relative to the SP after the new stack frame is established, but instead relative to the caller's stack pointer (it is stored into the linkage area of the parent's stack frame). So, like with the link register, we don't directly spill the CRs with other callee-saved registers, but just mark them to be spilled during prologue generation. In practice, this reverts r179457 for PPC64 (but leaves it in place for PPC32). llvm-svn: 179500	2013-04-15 02:07:05 +00:00
Eric Christopher	1f140317e3	Revert "Remove some unused triple and data layout." This reverts commit r179497 and the accompanying commit as it broke random platforms that aren't osx. llvm-svn: 179499	2013-04-14 23:35:36 +00:00
Eric Christopher	4eebd14ad0	Remove some unused triple and data layout. llvm-svn: 179498	2013-04-14 23:32:44 +00:00
Eric Christopher	e1876a2b79	If we've specified a triple on the command line then go ahead and use that as the default triple for the module and target data layout. llvm-svn: 179497	2013-04-14 23:32:40 +00:00
Nico Rieck	334c7bc7eb	Use object file specific section type for initial text section llvm-svn: 179494	2013-04-14 21:18:36 +00:00
David Majnemer	1fae195557	Reorders two transforms that collide with each other One performs: (X == 13 \| X == 14) -> X-13 <u 2 The other: (A == C1 \|\| A == C2) -> (A & ~(C1 ^ C2)) == C1 The problem is that there are certain values of C1 and C2 that trigger both transforms but the first one blocks out the second, this generates suboptimal code. Reordering the transforms should be better in every case and allows us to do interesting stuff like turn: %shr = lshr i32 %X, 4 %and = and i32 %shr, 15 %add = add i32 %and, -14 %tobool = icmp ne i32 %add, 0 into: %and = and i32 %X, 240 %tobool = icmp ne i32 %and, 224 llvm-svn: 179493	2013-04-14 21:15:43 +00:00
Nadav Rotem	6ebddae118	Make the command line triple match the module triple. llvm-svn: 179492	2013-04-14 20:13:05 +00:00
Benjamin Kramer	7d62ea86e5	Miscellaneous cleanups for VecUtils.h llvm-svn: 179483	2013-04-14 09:33:08 +00:00
Nadav Rotem	efa56e18be	Document the SLP infrastructure. llvm-svn: 179480	2013-04-14 07:42:25 +00:00
Nadav Rotem	3403c11529	SLP: Document the scalarization cost method. llvm-svn: 179479	2013-04-14 07:22:22 +00:00
Nadav Rotem	0db0690a70	Document the decision to assume that the cost of floats is twice as much as integers. llvm-svn: 179478	2013-04-14 05:55:18 +00:00
Jakob Stoklund Olesen	eed1072ff8	Use i32 for all SPARC shift amounts, even in 64-bit mode. Test case by llvm-stress. llvm-svn: 179477	2013-04-14 05:48:50 +00:00
Nadav Rotem	029208ceeb	Remove unused function attributes. llvm-svn: 179476	2013-04-14 05:47:04 +00:00
Nadav Rotem	54b413d157	SLPVectorizer: Add support for trees that don't start at binary operators, and add the cost of extracting values from the roots of the tree. llvm-svn: 179475	2013-04-14 05:15:53 +00:00
Jakob Stoklund Olesen	c3c28f8599	Add support for the abs64 SPARC v9 code model. For when 16 TB just isn't enough. llvm-svn: 179474	2013-04-14 05:10:36 +00:00
Jakob Stoklund Olesen	c8fc76b078	Add support for the SPARC v9 abs44 code model. This is the default model for non-PIC 64-bit code. It supports text+data+bss linked anywhere in the low 16 TB of the address space. llvm-svn: 179473	2013-04-14 04:57:51 +00:00
Jakob Stoklund Olesen	2e64d7ab1d	Use target flags for printing SPARC asm operands. 64-bit code models need multiple relocations that can't be inferred from the opcode like they can in 32-bit code. llvm-svn: 179472	2013-04-14 04:35:19 +00:00
Jakob Stoklund Olesen	e0fc832b77	Also put target flags on SPARC constant pool references. Constant pool entries are accessed exactly the same way as global variables. llvm-svn: 179471	2013-04-14 04:35:16 +00:00
Nadav Rotem	0b9cf8567b	SLPVectorizer: add initial support for reduction variable vectorization. llvm-svn: 179470	2013-04-14 03:22:20 +00:00
Jakob Stoklund Olesen	dc1ed57858	Fix patterns for 64-bit pointers. This fixes the pic32 code model for SPARC v9. llvm-svn: 179469	2013-04-14 01:53:23 +00:00
Jakob Stoklund Olesen	1fb08a8b08	Add target flags to SPARC address operands. SDNodes and MachineOperands get target flags representing the %hi() and %lo() assembly annotations that eventually become relocations. Also define flags to be used by the 64-bit code models. llvm-svn: 179468	2013-04-14 01:33:32 +00:00
Hal Finkel	2f29391504	Mark all PPC CR registers to be spilled as live-in and tag MFCR appropriately Leaving MFCR has having unmodeled side effects is not enough to prevent unwanted instruction reordering post-RA. We could probably apply a stronger barrier attribute, but there is a better way: Add all (not just the first) CR to be spilled as live-in to the entry block, and add all CRs to the MFCR instruction as implicitly killed. Unfortunately, I don't have a small test case. llvm-svn: 179465	2013-04-13 23:06:15 +00:00
Jakob Stoklund Olesen	15b3e90081	Define SPARC code models. Currently, only abs32 and pic32 are implemented. Add a test case for abs32 with 64-bit code. 64-bit PIC code is currently broken. llvm-svn: 179463	2013-04-13 19:02:23 +00:00
Jakob Stoklund Olesen	6a0a3eb53e	Use the correct types when matching ADDRri patterns from frame indexes. It doesn't seem like anybody is checking types this late in isel, so no test case. llvm-svn: 179462	2013-04-13 19:02:16 +00:00
Benjamin Kramer	adc1727c39	GlobalDCE: Fix an oversight in my last commit that could lead to crashes. There is a Constant with non-constant operands: blockaddress. llvm-svn: 179460	2013-04-13 16:11:14 +00:00
Benjamin Kramer	89ca4bc6d4	Fix a scalability issue with complex ConstantExprs. This is basically the same fix in three different places. We use a set to avoid walking the whole tree of a big ConstantExprs multiple times. For example: (select cmp, (add big_expr 1), (add big_expr 2)) We don't want to visit big_expr twice here, it may consist of thousands of nodes. The testcase exercises this by creating an insanely large ConstantExprs out of a loop. It's questionable if the optimizer should ever create those, but this can be triggered with real C code. Fixes PR15714. llvm-svn: 179458	2013-04-13 12:53:18 +00:00
Hal Finkel	d85a04b3df	Spill and restore PPC CR registers using the FP when we have one For functions that need to spill CRs, and have dynamic stack allocations, the value of the SP during the restore is not what it was during the save, and so we need to use the FP in these cases (as for all of the other spills and restores, but the CR restore has a special code path because its reserved slot, like the link register, is specified directly relative to the adjusted SP). llvm-svn: 179457	2013-04-13 08:09:20 +00:00
Andrew Trick	3d957c0ead	Further generalize this scheduler test. The order of copies depends on queue order, which is not very stable. llvm-svn: 179456	2013-04-13 07:37:27 +00:00
Andrew Trick	e6f9fc0cdb	Fix a dislexic regex. llvm-svn: 179455	2013-04-13 07:29:21 +00:00
Andrew Trick	88a1285b4f	Add a missing REQUIRES: asserts llvm-svn: 179453	2013-04-13 06:12:46 +00:00
Andrew Trick	1f0bb69b6c	MI-Sched: DEBUG formatting. llvm-svn: 179452	2013-04-13 06:07:49 +00:00
Andrew Trick	be2bccbce9	MI-Sched cleanup. If an instruction has no valid sched class, do not attempt to check for a variant. llvm-svn: 179451	2013-04-13 06:07:45 +00:00
Andrew Trick	f7fd6b9e3a	X86 machine model: reduce SandyBridge and Haswell ILPWindow. The initial values were arbitrary. I want them to be more conservative. This represents the number of latency cycles hidden by OOO execution. In practice, I think it should be within a small factor of the complex floating point operation latency so the scheduler can make some attempt to hide latency even for smallish blocks. These are by no means the best values, just a starting point for tuning heuristics. Some benchmarks such as TSVC run faster with this lower value for SandyBridge. I haven't run anything on Haswell, but it's shouldn't be 2x SB. llvm-svn: 179450	2013-04-13 06:07:43 +00:00
Andrew Trick	e833e1cd6e	MI-Sched: schedule physreg copies. The register allocator expects minimal physreg live ranges. Schedule physreg copies accordingly. This is slightly tricky when they occur in the middle of the scheduling region. For now, this is handled by rescheduling the copy when its associated instruction is scheduled. Eventually we may instead bundle them, but only if we can preserve the bundles as parallel copies during regalloc. llvm-svn: 179449	2013-04-13 06:07:40 +00:00
Andrew Trick	52b8387fd1	Catch another case where SD fails to propagate node order. I need to handle this for the test case in my following scheduler commit. Work is already under way to redesign the mechanism for node order propagation because this case by case approach is unmaintainable. llvm-svn: 179448	2013-04-13 06:07:36 +00:00
Rafael Espindola	98c0eaecf5	Add typenames to see if bot goes green. I hope this brings http://lab.llvm.org:8011/builders/clang-x86_64-darwin11-self-mingw32 back. llvm-svn: 179446	2013-04-13 02:31:34 +00:00
Akira Hatanaka	a6bbde5839	[mips] Move MipsTargetLowering::lowerINTRINSIC_W_CHAIN and lowerINTRINSIC_WO_CHAIN into MipsSETargetLowering. No functionality changes. llvm-svn: 179444	2013-04-13 02:13:30 +00:00
Rafael Espindola	6e8bb9eed5	Some versions of gcc don't like typenames in these places. Should fix the bots. llvm-svn: 179441	2013-04-13 01:55:34 +00:00
Rafael Espindola	9b709259e1	Finish templating MachObjectFile over endianness. We are now able to handle big endian macho files in llvm-readobject. Thanks to David Fang for providing the object files. llvm-svn: 179440	2013-04-13 01:45:40 +00:00
Akira Hatanaka	2f08822f9d	[mips] Reapply r179420 and r179421. llvm-svn: 179434	2013-04-13 00:55:41 +00:00
Akira Hatanaka	48996b0608	[mips] Override TargetLoweringBase::isShuffleMaskLegal. llvm-svn: 179433	2013-04-13 00:45:02 +00:00
Chad Rosier	43554eed5e	[ms-inline asm] Simplify the logic by using parsePrimaryExpr. No functional change intended. Test case previously added in r178568. Part of rdar://13611297 llvm-svn: 179425	2013-04-12 23:03:20 +00:00

1 2 3 4 5 ...

91122 Commits