llvm-project

Commit Graph

Author	SHA1	Message	Date
Bjorn Steinbrink	236446cd4c	Remove conflicting attributes before adding deduced readonly/readnone Summary: In case of functions that have a pointer argument and only pass it to each other, the function attributes pass deduces that the pointer should get the readnone attribute, but fails to remove a readonly attribute that may already have been present. Reviewers: nlewycky Subscribers: llvm-commits Differential Revision: http://reviews.llvm.org/D9995 llvm-svn: 238152	2015-05-25 19:46:38 +00:00
Rafael Espindola	cd62518369	Move HasInstructions to MCSection. llvm-svn: 238150	2015-05-25 18:34:26 +00:00
Simon Pilgrim	0be4fa761f	[X86][AVX2] Vectorized i16 shift operators Part of D9474, this patch extends AVX2 v16i16 types to 2 x 8i32 vectors and uses i32 shift variable shifts before packing back to i16. Adds AVX2 tests for v8i16 and v16i16 llvm-svn: 238149	2015-05-25 17:49:13 +00:00
Tom Stellard	50828163a1	R600/SI: Remove some unnecessary patterns from VINTRP multiclass DisableEncoding and Constraints can be set using let statements around the multiclass defs. llvm-svn: 238148	2015-05-25 16:15:56 +00:00
Tom Stellard	ec87f841c6	R600/SI: Fix bug with v_interp_p1_f32 instructions on 16 bank lds chips The src and dst register cannot be the same on chips with 16 lds banks. llvm-svn: 238147	2015-05-25 16:15:54 +00:00
Tom Stellard	c70cf90d09	R600/SI: Use NAME rather than opName as the key to the MCOpcode tables This lets us drop a parameter the opName parameter to the VINTRP multiclass and makes it possible to create multiple VINTRP defs with the same asm mnemonic. llvm-svn: 238146	2015-05-25 16:15:50 +00:00
Kit Barton	6646033e6e	This patch adds support for the vector quadword add/sub instructions introduced in POWER8: vadduqm vaddeuqm vaddcuq vaddecuq vsubuqm vsubeuqm vsubcuq vsubecuq In addition to adding the instructions themselves, it also adds support for the v1i128 type for intrinsics (Intrinsics.td, Function.cpp, and IntrinsicEmitter.cpp). http://reviews.llvm.org/D9081 llvm-svn: 238144	2015-05-25 15:49:26 +00:00
Rafael Espindola	b028cc8098	Move bundle info from MCSectionData to MCSection. llvm-svn: 238143	2015-05-25 15:04:26 +00:00
Rafael Espindola	c1cd944854	Add a isBundleLocked helper to MCELFStreamer. llvm-svn: 238142	2015-05-25 14:57:35 +00:00
Rafael Espindola	b2ac19ed6e	Move LayoutOrder to MCSection. llvm-svn: 238141	2015-05-25 14:25:28 +00:00
Rafael Espindola	6e6820a7e6	Stop forwarding getOrdinal and setOrdinal. llvm-svn: 238139	2015-05-25 14:12:48 +00:00
Rafael Espindola	1f02022027	Move Ordinal from MCSectionData to MCSection. NFC. Part of the work to merge MCSectionData and MCSection. llvm-svn: 238137	2015-05-25 14:00:56 +00:00
Rafael Espindola	e7134b2d78	Simplify boolean conditional return statements. Patch by Richard <legalize@xmission.com> llvm-svn: 238134	2015-05-25 13:50:21 +00:00
Benjamin Kramer	68a29562f9	Refactor: Simplify boolean conditional return statements in llvm/lib/DebugInfo/DWARF Use clang-tidy to simplify boolean conditional return statements. Patch by Richard Thomson <legalize@xmission.com>! Differential Revision: http://reviews.llvm.org/D9972 llvm-svn: 238132	2015-05-25 13:28:03 +00:00
Michael Kuperstein	f145228676	[X86] When pattern-matching scalar FMA3 intrinsics, don't re-arrange the first and second operands. The semantics of the scalar FMA intrinsics are that the high vector elements are copied from the first source. The existing pattern switches src1 and src2 around, to match the "213" order, which ends up tying the original src2 to the dest. Since the actual scalar fma3 instructions copy the high elements from the dest register, the wrong values are copied. This modifies the pattern to leave src1 and src2 in their original order. Differential Revision: http://reviews.llvm.org/D9908 llvm-svn: 238131	2015-05-25 12:35:25 +00:00
Elena Demikhovsky	1c1391ba24	Added promotion to EXTRACT_SUBVECTOR operand. I encountered with this case in one of KNL tests for i1 vectors. v16i1 = EXTRACT_SUBVECTOR v32i1, x llvm-svn: 238130	2015-05-25 11:33:13 +00:00
NAKAMURA Takumi	5582a6a4a5	Reformat. llvm-svn: 238126	2015-05-25 01:43:34 +00:00
NAKAMURA Takumi	fb3bd7127a	Prune CRLFs. llvm-svn: 238125	2015-05-25 01:43:23 +00:00
Chandler Carruth	04cc665cef	[Unroll] Switch from an eagerly populated SCEV cache to one that is lazily built. Also, make it a much more generic SCEV cache, which today exposes only a reduced GEP model description but could be extended in the future to do other profitable caching of SCEV information. llvm-svn: 238124	2015-05-25 01:00:46 +00:00
Duncan P. N. Exon Smith	882a2b5a7d	AsmPrinter: Avoid creating symbols in DwarfStringPool Stop creating symbols we don't need in `DwarfStringPool`. The consumers only call `DwarfStringPoolEntryRef::getSymbol()` when DWARF is relocatable, so this just stops creating the unused symbols when it's not. This drops memory usage from 851 MB to 845 MB, around 0.7%. (I'm looking at `llc` memory usage on `verify-uselistorder.lto.opt.bc`; see r236629 for details.) llvm-svn: 238122	2015-05-24 16:58:59 +00:00
Duncan P. N. Exon Smith	9d50e82fb2	AsmPrinter: Prune an include, NFC llvm-svn: 238121	2015-05-24 16:54:59 +00:00
Duncan P. N. Exon Smith	e344705ade	AsmPrinter: Remove dead code, NFC llvm-svn: 238120	2015-05-24 16:51:29 +00:00
Duncan P. N. Exon Smith	1e0d94e7bb	AsmPrinter: Avoid EmitLabelDifference() in DwarfAccelTable Mint a new function, `AsmPrinter::emitDwarfStringOffset()`, which takes a `DwarfStringPoolEntryRef`. When DWARF is relocatable across sections, this defers to `emitSectionOffset()` and emits the `MCSymbol`; otherwise, just emit the offset directly, without using any intermediate symbols. `EmitLabelDifference()` is already optimized to emit absolute label differences cheaply when possible, so there aren't any major memory savings here (853 MB down to 851 MB, or 0.2%). However, it prepares for making the `MCSymbol`s in the `DwarfStringPool` optional. (I'm looking at `llc` memory usage on `verify-uselistorder.lto.opt.bc`; see r236629 for details.) llvm-svn: 238119	2015-05-24 16:48:54 +00:00
Duncan P. N. Exon Smith	f4599942fb	AsmPrinter: Use DwarfStringPoolEntry in DwarfAccelTable, NFC This is just an API change, but it prepares to stop using `EmitLabelDifference()` when possible. llvm-svn: 238118	2015-05-24 16:44:32 +00:00
Duncan P. N. Exon Smith	f73bcf4020	AsmPrinter: Make DIEString small Expose the `DwarfStringPool` entry in a header, and store a pointer to it directly in `DIEString`. Instead of choosing at creation time how to emit it, use the `dwarf::Form` to determine that at emission time. Besides avoiding the other `DIEValue`, this shaves two pointers off of `DIEString`; the data is now a single pointer. This is a nice cleanup on its own -- and drops memory usage from 861 MB down to 853 MB, around 0.9% -- but it's also preparation for passing `DIEValue`s by value. (I'm looking at `llc` memory usage on `verify-uselistorder.lto.opt.bc`; see r236629 for details.) llvm-svn: 238117	2015-05-24 16:40:47 +00:00
Duncan P. N. Exon Smith	03b7a1cf93	AsmPrinter: Extract DwarfStringPoolEntry from DwarfStringPool, NFC Extract out `DwarfStringPoolEntry` and `DwarfStringPoolRef` from `DwarfStringPool` so that downstream users can start using `DwarfStringPool::getEntry()` directly. This will allow users to delay the decision between emitting a symbol or an offset until later. llvm-svn: 238116	2015-05-24 16:33:33 +00:00
Duncan P. N. Exon Smith	1a65e4ade4	AsmPrinter: Emit the DwarfStringPool offset directly when possible Change `DwarfStringPool` to calculate byte offsets on-the-fly, and update `DwarfUnit::getLocalString()` to use a `DIEInteger` instead of a `DIEDelta` when Dwarf doesn't use relocations (i.e., Mach-O). This eliminates another call to `EmitLabelDifference()`, and drops memory usage from 865 MB down to 861 MB, around 0.5%. (I'm looking at `llc` memory usage on `verify-uselistorder.lto.opt.bc`; see r236629 for details.) llvm-svn: 238114	2015-05-24 16:14:59 +00:00
Duncan P. N. Exon Smith	8c6499fa6d	AsmPrinter: Refactor DwarfStringPool::getEntry(), NFC Move `DwarfStringPool`'s `getEntry()` to the header (and make it a member function) in preparation for calculating symbol offsets on-the-fly. llvm-svn: 238112	2015-05-24 16:06:08 +00:00
Renato Golin	fe54d34bc6	Move parseSubArch to ARMTargetParser. NFC Using getCanonicalArchName() is the right way to parse ARM arch names. Mapping ARMTargetParser IDs to Triple Arch IDs is temporary, until they are merged into a TargetDescription class. This was the last LLVM FIXME to move things to ARMTargetParser. Now on to Clang and beyond. llvm-svn: 238110	2015-05-24 11:18:44 +00:00
Matt Arsenault	65ad1602b0	Add target hook to allow merging stores of nonzero constants On GPU targets, materializing constants is cheap and stores are expensive, so only doing this for zero vectors was silly. Most of the new testcases aren't optimally merged, and are for later improvements. llvm-svn: 238108	2015-05-24 00:51:27 +00:00
Benjamin Kramer	1577f1f484	Bump SmallString to the minimum required amount for raw_ostream to avoid allocation. NFC. llvm-svn: 238104	2015-05-23 17:20:53 +00:00
Benjamin Kramer	33b4691fd0	[Mips] Prefer Twine::utohexstr over utohexstr, saves a string copy. NFC. llvm-svn: 238103	2015-05-23 16:53:07 +00:00
Benjamin Kramer	be48c40475	[AArch64] Clean up the ELF streamer a bit. llvm-svn: 238102	2015-05-23 16:39:10 +00:00
Benjamin Kramer	1d1b9243d5	[AArch64] Move AArch64TargetStreamer out of MCStreamer.h It doesn't belong in the shared MC layer. NFC. llvm-svn: 238101	2015-05-23 16:15:10 +00:00
Aaron Ballman	c681c3d890	Silencing a spurious -Wreturn-type warning; NFC. llvm-svn: 238099	2015-05-23 14:46:49 +00:00
Hal Finkel	5f2a1379ef	[PowerPC] Fix fast-isel when compare is split from branch When the compare feeding a branch was in a different BB from the branch, we'd try to "regenerate" the compare in the block with the branch, possibly trying to make use of values not available there. Copy a page from AArch64's play book here to fix the problem (at least in terms of correctness). Fixes PR23640. llvm-svn: 238097	2015-05-23 12:18:10 +00:00
NAKAMURA Takumi	b558e5fb65	Update ExecutionEngine/LLVMBuild.txt, to add LLVMCodeGen. llvm-svn: 238096	2015-05-23 10:44:30 +00:00
Craig Topper	10949ae742	Give more meaningful names than I and J to some for loop variables after converting to range-based loops. llvm-svn: 238095	2015-05-23 08:45:10 +00:00
Craig Topper	37d0d866f4	Fix an unused variable warning in release builds. llvm-svn: 238094	2015-05-23 08:20:33 +00:00
Craig Topper	77b9941ab9	Use range-based for loops. NFC. llvm-svn: 238093	2015-05-23 08:01:41 +00:00
Kostya Serebryany	e0d60ba876	[lib/Fuzzer] doxygen-ify the comments for the user interface llvm-svn: 238086	2015-05-23 02:12:05 +00:00
Duncan P. N. Exon Smith	68b3f30778	AsmPrinter: Remove the vtable-entry from DIEValue Remove all virtual functions from `DIEValue`, dropping the vtable pointer from its layout. Instead, create "impl" functions on the subclasses, and use the `DIEValue::Type` to implement the dynamic dispatch. This is necessary -- obviously not sufficient -- for passing `DIEValue`s around by value. However, this change stands on its own: we make tons of these. I measured a drop in memory usage from 888 MB down to 860 MB, or around 3.2%. (I'm looking at `llc` memory usage on `verify-uselistorder.lto.opt.bc`; see r236629 for details.) llvm-svn: 238084	2015-05-23 01:45:07 +00:00
Duncan P. N. Exon Smith	d5aa33525c	CodeGen: Remove redundant DIETypeSignature::dump(), NFC We already have this in `DIEValue`; no reason to shadow it. llvm-svn: 238082	2015-05-23 01:26:26 +00:00
Kostya Serebryany	7c180eafc1	[lib/Fuzzer] fully get rid of std::cerr in libFuzzer llvm-svn: 238081	2015-05-23 01:22:35 +00:00
Akira Hatanaka	ddf76aa36f	Stop resetting NoFramePointerElim in TargetMachine::resetTargetOptions. This is part of the work to remove TargetMachine::resetTargetOptions. In this patch, instead of updating global variable NoFramePointerElim in resetTargetOptions, its use in DisableFramePointerElim is replaced with a call to TargetFrameLowering::noFramePointerElim. This function determines on a per-function basis if frame pointer elimination should be disabled. There is no change in functionality except that cl:opt option "disable-fp-elim" can now override function attribute "no-frame-pointer-elim". llvm-svn: 238080	2015-05-23 01:14:08 +00:00
Akira Hatanaka	bd881834c5	Simplify and rename function overrideFunctionAttributes. NFC. This is in preparation to making changes needed to stop resetting NoFramePointerElim in resetTargetOptions. llvm-svn: 238079	2015-05-23 01:12:26 +00:00
Kostya Serebryany	20e9bcbfc8	[lib/Fuzzer] start getting rid of std::cerr. Sadly, these parts of C++ library used in libFuzzer badly interract with the same code used in the target function and also with dfsan. It's easier to just not use std::cerr than to defeat these issues. llvm-svn: 238078	2015-05-23 01:07:46 +00:00
Rafael Espindola	445712264d	Revert "make reciprocal estimate code generation more flexible by adding command-line options" This reverts commit r238051. It broke some bots: http://lab.llvm.org:8011/builders/llvm-ppc64-linux1/builds/18190 llvm-svn: 238075	2015-05-23 00:22:44 +00:00
Rafael Espindola	5960cee1f5	Produce a single string table in a ELF .o Normally an ELF .o has two string tables, one for symbols, one for section names. With the scheme of naming sections like ".text.foo" where foo is a symbol, there is a big potential saving in using a single one. Building llvm+clang+lld with master and with this patch the results were: master: 193,267,008 bytes patch: 186,107,952 bytes master non unique section names: 183,260,192 bytes patch non unique section names: 183,118,632 bytes So using non usique saves 10,006,816 bytes, and the patch saves 7,159,056 while still using distinct names for the sections. llvm-svn: 238073	2015-05-22 23:58:30 +00:00
Philip Reames	7c78ef7dd9	Extend EarlyCSE to handle basic cases from JumpThreading and CVP This patch extends EarlyCSE to take advantage of the information that a controlling branch gives us about the value of a Value within this and dominated basic blocks. If the current block has a single predecessor with a controlling branch, we can infer what the branch condition must have been to execute this block. The actual change to support this is downright simple because EarlyCSE's existing scoped hash table logic deals with most of the complexity around merging. The patch actually implements two optimizations. 1) The first is analogous to JumpThreading in that it enables EarlyCSE's CSE handling to fold branches which are exactly redundant due to a previous branch to branches on constants. (It doesn't actually replace the branch or change the CFG.) This is pretty clearly a win since it enables substantial CFG simplification before we start trying to inline. 2) The second is analogous to CVP in that it exploits the knowledge gained to replace dominated uses of the original value. EarlyCSE does not otherwise reason about specific uses, so this is the more arguable one. It does enable further simplication and constant folding within the rest of the visit by EarlyCSE. In both cases, the added code only handles the easy dominance based case of each optimization. The general case is deferred to the existing passes. Differential Revision: http://reviews.llvm.org/D9763 llvm-svn: 238071	2015-05-22 23:53:24 +00:00

1 2 3 4 5 ...

79881 Commits