llvm-project

Commit Graph

Author	SHA1	Message	Date
Dan Gohman	20af5a0fe7	Check to see if a two-entry PHI block can be simplified before trying to merge the block into its predecessors. This allows two-entry-phi-return.ll to be simplified into a single basic block. llvm-svn: 48252	2008-03-11 21:53:06 +00:00
Devang Patel	64d0f07085	Restore optimization that merges blocks when inline function has single return value. llvm-svn: 48162	2008-03-10 18:34:00 +00:00
Devang Patel	72ea2dc9a9	Simplify llvm-svn: 48161	2008-03-10 18:22:16 +00:00
Devang Patel	c0325b2040	simplify llvm-svn: 48160	2008-03-10 18:11:41 +00:00
Nick Lewycky	fb2c1a999a	Turn unwind_to into "unwinds to". llvm-svn: 48123	2008-03-10 02:20:00 +00:00
Nick Lewycky	42445be0df	Firstly, having a BranchInst isn't exclusive with having an unwind_to. Secondly, we have to check whether the branch is actually pointing to the block with the unwind in it. We could have gotten here because of the unwind_to alone. llvm-svn: 48099	2008-03-09 07:50:37 +00:00
Nick Lewycky	f3d637fa14	A BB that unwind_to an "unwind" inst is that same as one that doesn't unwind_to at all. llvm-svn: 48096	2008-03-09 07:36:38 +00:00
Nick Lewycky	11fc6f8765	Update the block cloner which fixes bugpoint on code using unwind_to (phew!) and also update the cloning interface's major user, the loop optimizations. llvm-svn: 48088	2008-03-09 05:24:34 +00:00
Nick Lewycky	5ce9b521d7	Update the inliner and simplifycfg to handle unwind_to. llvm-svn: 48086	2008-03-09 05:10:13 +00:00
Nick Lewycky	cc24104703	Two things. Preserve the unwind_to when splitting a BB. Add the ability to remove just one instance of a BB from a phi node. This fixes the compile error in the tree now. llvm-svn: 48085	2008-03-09 05:04:48 +00:00
Devang Patel	780b3ca64b	Update inliner to handle functions that return multiple values. llvm-svn: 48020	2008-03-07 20:06:16 +00:00
Devang Patel	3b1c95f885	Handle 'ret' with multiple values. llvm-svn: 47965	2008-03-05 21:50:24 +00:00
Devang Patel	e516aa1127	Skip functions that return multiple values. llvm-svn: 47924	2008-03-05 00:36:59 +00:00
Devang Patel	4566d885dd	Use while loop. llvm-svn: 47909	2008-03-04 21:59:49 +00:00
Devang Patel	941ab37ea8	Use cast instead of dyn_cast. Update test to use multiple return value directly, instead of relying on -sretpromotion. llvm-svn: 47907	2008-03-04 21:45:28 +00:00
Devang Patel	841322b32a	Handle multiple return values. llvm-svn: 47904	2008-03-04 21:15:15 +00:00
Anton Korobeynikov	18991d78fa	Fix newly-introduced 4.3 warnings llvm-svn: 47375	2008-02-20 12:07:57 +00:00
Anton Korobeynikov	1bfd121321	Make Transforms to be 4.3 warnings-clean llvm-svn: 47371	2008-02-20 11:26:25 +00:00
Chris Lattner	c3591a0d48	remove the LowerSelect pass. The last client was the old Sparc backend, which is long dead by now. llvm-svn: 47323	2008-02-19 07:49:17 +00:00
Chris Lattner	6b39cb907b	switch simplifycfg from using vectors for most things to smallvectors, this speeds it up 2.3% on eon. llvm-svn: 47261	2008-02-18 07:42:56 +00:00
Chris Lattner	70e294660a	Fix PR2029 llvm-svn: 47129	2008-02-14 19:18:13 +00:00
Chris Lattner	a838141957	Make RenamePass faster by making the 'is this a new phi node' check more intelligent. This speeds up mem2reg from 5.29s to 0.79s on a synthetic testcase with tons of predecessors and phi nodes. llvm-svn: 46767	2008-02-05 21:26:23 +00:00
Duncan Sands	053c9871cd	Revert r46393: readonly/readnone functions are no longer allowed to write through byval arguments. llvm-svn: 46416	2008-01-27 18:12:58 +00:00
Duncan Sands	c4dc3dc3a2	Create an explicit copy for byval parameters even when inlining a readonly function. llvm-svn: 46393	2008-01-26 06:41:49 +00:00
Duncan Sands	f52faf9a64	Do this more neatly. llvm-svn: 46369	2008-01-25 22:06:51 +00:00
Chris Lattner	4f6c81ac68	we don't have to make an explicit copy of a byval argument when inlining a function if we know that the function does not write to any memory. This implements test/Transforms/Inline/byval2.ll llvm-svn: 45912	2008-01-12 18:54:29 +00:00
Chris Lattner	908117bf69	When inlining a functino with a byval argument, make an explicit copy of it in case the callee modifies the struct. llvm-svn: 45853	2008-01-11 06:09:30 +00:00
Chris Lattner	f391883670	don't hoist FP additions into unconditional adds + selects. This could theoretically introduce a trap, but is also a performance issue. This speeds up ptrdist/ks by 8%. llvm-svn: 45533	2008-01-03 07:25:26 +00:00
Chris Lattner	f3ebc3f3d2	Remove attribution from file headers, per discussion on llvmdev. llvm-svn: 45418	2007-12-29 20:36:04 +00:00
Chris Lattner	a087a8d2ce	remove attribution from lib Makefiles. llvm-svn: 45415	2007-12-29 20:09:26 +00:00
Chris Lattner	e96658392d	dead calls to llvm.stacksave can be deleted, even though they have potential side-effects. llvm-svn: 45392	2007-12-29 00:59:12 +00:00
Gordon Henriksen	b969c5981b	GC poses hazards to the inliner. Consider: define void @f() { ... call i32 @g() ... } define void @g() { ... } The hazards are: - @f and @g have GC, but they differ GC. Inlining is invalid. This may never occur. - @f has no GC, but @g does. g's GC must be propagated to @f. The other scenarios are safe: - @f and @g have the same GC. - @f and @g have no GC. - @g has no GC. This patch adds inliner checks for the former two scenarios. llvm-svn: 45351	2007-12-25 03:10:07 +00:00
Devang Patel	7a2c66b11e	If succ has succ itself as one of the predecessors then do not merge current bb and succ even if bb's terminator is unconditional branch to succ. llvm-svn: 45305	2007-12-22 01:32:53 +00:00
Duncan Sands	aa31b92508	When inlining through an 'nounwind' call, mark inlined calls 'nounwind'. It is important for correct C++ exception handling that nounwind markings do not get lost, so this transformation is actually needed for correctness. llvm-svn: 45218	2007-12-19 21:13:37 +00:00
Duncan Sands	3353ed09ac	Rename isNoReturn to doesNotReturn, and isNoUnwind to doesNotThrow. llvm-svn: 45160	2007-12-18 09:59:50 +00:00
Duncan Sands	b5a79d0eaa	Make invokes of inline asm legal. Teach codegen how to lower them (with no attempt made to be efficient, since they should only occur for unoptimized code). llvm-svn: 45108	2007-12-17 18:08:19 +00:00
David Greene	71eae8a5ee	GLIBCXX_DEBUG fix. std::vector<>::end() is invalidated by erase. llvm-svn: 45101	2007-12-17 17:42:03 +00:00
Christopher Lamb	edf0788758	Change the PointerType api for creating pointer types. The old functionality of PointerType::get() has become PointerType::getUnqual(), which returns a pointer in the generic address space. The new prototype of PointerType::get() requires both a type and an address space. llvm-svn: 45082	2007-12-17 01:12:55 +00:00
Duncan Sands	56ed48036b	Revert this part of r45073 until the verifier is changed not to reject invoke of inline asm. llvm-svn: 45077	2007-12-16 21:01:21 +00:00
Duncan Sands	8e4847ee95	Make instcombine promote inline asm calls to 'nounwind' calls. Remove special casing of inline asm from the inliner. There is a potential problem: the verifier rejects invokes of inline asm (not sure why). If an asm call is not marked "nounwind" in some .ll, and instcombine is not run, but the inliner is run, then an illegal module will be created. This is bad but I'm not sure what the best approach is. I'm tempted to remove the check in the verifier... llvm-svn: 45073	2007-12-16 15:51:49 +00:00
Chris Lattner	d2265b45ae	Fix PR1850 by removing an unsafe transformation from VMCore/ConstantFold.cpp. Reimplement the xform in Analysis/ConstantFolding.cpp where we can use targetdata to validate that it is safe. While I'm in there, fix some const correctness issues and generalize the interface to the "operand folder". llvm-svn: 44817	2007-12-10 22:53:04 +00:00
Gordon Henriksen	71183b6739	Adding a collector name attribute to Function in the IR. These methods are new to Function: bool hasCollector() const; const std::string &getCollector() const; void setCollector(const std::string &); void clearCollector(); The assembly representation is as such: define void @f() gc "shadow-stack" { ... The implementation uses an on-the-side table to map Functions to collector names, such that there is no overhead. A StringPool is further used to unique collector names, which are extremely likely to be unique per process. llvm-svn: 44769	2007-12-10 03:18:06 +00:00
Duncan Sands	38ef3a8ec7	Rather than having special rules like "intrinsics cannot throw exceptions", just mark intrinsics with the nounwind attribute. Likewise, mark intrinsics as readnone/readonly and get rid of special aliasing logic (which didn't use anything more than this anyway). llvm-svn: 44544	2007-12-03 20:06:50 +00:00
Duncan Sands	ad0ea2d430	Fix PR1146: parameter attributes are longer part of the function type, instead they belong to functions and function calls. This is an updated and slightly corrected version of Reid Spencer's original patch. The only known problem is that auto-upgrading of bitcode files doesn't seem to work properly (see test/Bitcode/AutoUpgradeIntrinsics.ll). Hopefully a bitcode guru (who might that be? :) ) will fix it. llvm-svn: 44359	2007-11-27 13:23:08 +00:00
Owen Anderson	b0dd27ee91	Make LoopInfoBase more generic, in preparation for having MachineLoopInfo. This involves a small interface change. llvm-svn: 44348	2007-11-27 03:43:35 +00:00
Anton Korobeynikov	550b98e147	Fix indent llvm-svn: 43941	2007-11-09 12:34:20 +00:00
Anton Korobeynikov	98638aede6	Forget to commit users part of value mapper interface llvm-svn: 43940	2007-11-09 12:27:04 +00:00
Anton Korobeynikov	8eeca1c252	And delete this one llvm-svn: 43939	2007-11-09 12:22:04 +00:00
Gordon Henriksen	d568767ecb	Finishing initial docs for all transformations in Passes.html. Also cleaned up some comments in source files. llvm-svn: 43674	2007-11-04 16:15:04 +00:00
Dan Gohman	d7917b6248	Add std:: to sort calls. llvm-svn: 43652	2007-11-02 22:24:01 +00:00
Dan Gohman	c981d72d1a	Change illegal uses of ++ to uses of STLExtra.h's next function. llvm-svn: 43651	2007-11-02 22:22:02 +00:00
Duncan Sands	44b8721de8	Executive summary: getTypeSize -> getTypeStoreSize / getABITypeSize. The meaning of getTypeSize was not clear - clarifying it is important now that we have x86 long double and arbitrary precision integers. The issue with long double is that it requires 80 bits, and this is not a multiple of its alignment. This gives a primitive type for which getTypeSize differed from getABITypeSize. For arbitrary precision integers it is even worse: there is the minimum number of bits needed to hold the type (eg: 36 for an i36), the maximum number of bits that will be overwriten when storing the type (40 bits for i36) and the ABI size (i.e. the storage size rounded up to a multiple of the alignment; 64 bits for i36). This patch removes getTypeSize (not really - it is still there but deprecated to allow for a gradual transition). Instead there is: (1) getTypeSizeInBits - a number of bits that suffices to hold all values of the type. For a primitive type, this is the minimum number of bits. For an i36 this is 36 bits. For x86 long double it is 80. This corresponds to gcc's TYPE_PRECISION. (2) getTypeStoreSizeInBits - the maximum number of bits that is written when storing the type (or read when reading it). For an i36 this is 40 bits, for an x86 long double it is 80 bits. This is the size alias analysis is interested in (getTypeStoreSize returns the number of bytes). There doesn't seem to be anything corresponding to this in gcc. (3) getABITypeSizeInBits - this is getTypeStoreSizeInBits rounded up to a multiple of the alignment. For an i36 this is 64, for an x86 long double this is 96 or 128 depending on the OS. This is the spacing between consecutive elements when you form an array out of this type (getABITypeSize returns the number of bytes). This is TYPE_SIZE in gcc. Since successive elements in a SequentialType (arrays, pointers and vectors) need to be aligned, the spacing between them will be given by getABITypeSize. This means that the size of an array is the length times the getABITypeSize. It also means that GEP computations need to use getABITypeSize when computing offsets. Furthermore, if an alloca allocates several elements at once then these too need to be aligned, so the size of the alloca has to be the number of elements multiplied by getABITypeSize. Logically speaking this doesn't have to be the case when allocating just one element, but it is simpler to also use getABITypeSize in this case. So alloca's and mallocs should use getABITypeSize. Finally, since gcc's only notion of size is that given by getABITypeSize, if you want to output assembler etc the same as gcc then getABITypeSize is the size you want. Since a store will overwrite no more than getTypeStoreSize bytes, and a read will read no more than that many bytes, this is the notion of size appropriate for alias analysis calculations. In this patch I have corrected all type size uses except some of those in ScalarReplAggregates, lib/Codegen, lib/Target (the hard cases). I will get around to auditing these too at some point, but I could do with some help. Finally, I made one change which I think wise but others might consider pointless and suboptimal: in an unpacked struct the amount of space allocated for a field is now given by the ABI size rather than getTypeStoreSize. I did this because every other place that reserves memory for a type (eg: alloca) now uses getABITypeSize, and I didn't want to make an exception for unpacked structs, i.e. I did it to make things more uniform. This only effects structs containing long doubles and arbitrary precision integers. If someone wants to pack these types more tightly they can always use a packed struct. llvm-svn: 43620	2007-11-01 20:53:16 +00:00
Chris Lattner	4a15e04aee	Fix PR1752 and LoopSimplify/2007-10-28-InvokeCrash.ll: terminators can have uses too. Wouldn't it be nice if invoke didn't exist? :) llvm-svn: 43426	2007-10-29 02:30:37 +00:00
Anton Korobeynikov	7499a3b092	Reg2Mem cleanup and optimizations: - enable phi instructions demotion to stack - create alloca instructions in the entry block llvm-svn: 43208	2007-10-21 23:05:16 +00:00
Owen Anderson	ca831a829d	Move Split<...>() into DomTreeBase. This should make the #include's of DominatorInternals.h in CodeExtractor and LoopSimplify unnecessary. Hartmut, could you confirm that this fixes the issues you were seeing? llvm-svn: 43115	2007-10-18 05:13:52 +00:00
Hartmut Kaiser	2f842e613f	Fixed linker errors (unresolved externals: split<>(...)) when compiling with VC++. Please review. llvm-svn: 43081	2007-10-17 18:37:09 +00:00
Devang Patel	9d1af9b63d	Fix comment. llvm-svn: 42048	2007-09-17 20:07:40 +00:00
Chris Lattner	0625bd6472	Merge DenseMapKeyInfo & DenseMapValueInfo into DenseMapInfo Add a new DenseMapInfo::isEqual method to allow clients to redefine the equality predicate used when probing the hash table. llvm-svn: 42042	2007-09-17 18:34:04 +00:00
Devang Patel	f6ef552f3d	Insert cloned loop basic blocks before original loop header. llvm-svn: 41713	2007-09-04 20:46:35 +00:00
David Greene	c656cbb8c2	Update GEP constructors to use an iterator interface to fix GLIBCXX_DEBUG issues. llvm-svn: 41697	2007-09-04 15:46:09 +00:00
Anton Korobeynikov	35322d745c	Silence warning while compiling with gcc 4.2 llvm-svn: 41676	2007-09-02 22:11:14 +00:00
David Greene	703623d571	Update InvokeInst to work like CallInst llvm-svn: 41506	2007-08-27 19:04:21 +00:00
Anton Korobeynikov	24fb6b2f8c	Don't promote volatile loads/stores. This is needed (for example) to handle setjmp/longjmp properly. This fixes PR1520. llvm-svn: 41461	2007-08-26 21:43:30 +00:00
Devang Patel	b5933bbbd5	Use SmallVector instead of std::vector. llvm-svn: 41207	2007-08-21 00:31:24 +00:00
Devang Patel	d1fcfcc76c	When one branch of condition is eliminated then head of the other branch is not necessary immediate dominators of merge blcok in all cases. llvm-svn: 41144	2007-08-17 21:59:16 +00:00
Devang Patel	22c7993ecf	Break infinite loop. llvm-svn: 41091	2007-08-14 23:59:17 +00:00
Devang Patel	da48cf40db	If NewBB dominates DestBB then DestBB is not part of NewBB's dominance frontier. llvm-svn: 41051	2007-08-13 21:59:17 +00:00
Devang Patel	aa36a43908	Add utility to clone loops. llvm-svn: 40997	2007-08-10 17:59:47 +00:00
Chris Lattner	c7ba225705	remove some dead lines llvm-svn: 40859	2007-08-06 06:21:06 +00:00
Chris Lattner	edce70d2fe	rewrite the code used to construct pruned SSA form with the IDF method. In the old way, we computed and inserted phi nodes for the whole IDF of the definitions of the alloca, then computed which ones were dead and removed them. In the new method, we first compute the region where the value is live, and use that information to only insert phi nodes that are live. This eliminates the need to compute liveness later, and stops the algorithm from inserting a bunch of phis which it then later removes. This speeds up the testcase in PR1432 from 2.00s to 0.15s (14x) in a release build and 6.84s->0.50s (14x) in a debug build. llvm-svn: 40825	2007-08-04 22:50:14 +00:00
Chris Lattner	d91576b01e	Factor out a whole bunch of code into it's own method. llvm-svn: 40824	2007-08-04 21:14:29 +00:00
Chris Lattner	4e1b4140eb	Use getNumPreds(BB) instead of computing them manually. This is a very small but measurable speedup. llvm-svn: 40823	2007-08-04 21:06:15 +00:00
Chris Lattner	b6a4ba808b	Change the rename pass to be "tail recursive", only adding N-1 successors to the worklist, and handling the last one with a 'tail call'. This speeds up PR1432 from 2.0578s to 2.0012s (2.8%) llvm-svn: 40822	2007-08-04 20:40:27 +00:00
Chris Lattner	840259c8d3	cache computation of #preds for a BB. This speeds up mem2reg from 2.0742->2.0522s on PR1432. llvm-svn: 40821	2007-08-04 20:24:50 +00:00
Chris Lattner	050bac4bed	reserve operand space for phi nodes when we insert them. llvm-svn: 40820	2007-08-04 20:14:34 +00:00
Chris Lattner	9318785df5	use continue to avoid nesting, no functionality change. llvm-svn: 40819	2007-08-04 20:07:06 +00:00
Chris Lattner	6b04ecbaf9	Promoting allocas with the 'single store' fastpath is faster than with the 'local to a block' fastpath. This speeds up PR1432 from 2.1232 to 2.0686s (2.6%) llvm-svn: 40818	2007-08-04 20:03:23 +00:00
Chris Lattner	4a930f9444	When PromoteLocallyUsedAllocas promoted allocas, it didn't remember to increment NumLocalPromoted, and didn't actually delete the dead alloca, leading to an extra iteration of mem2reg. llvm-svn: 40817	2007-08-04 20:01:43 +00:00
Chris Lattner	63c039780c	std::map -> DenseMap llvm-svn: 40816	2007-08-04 19:52:20 +00:00
Chris Lattner	7d382f7680	fix a logic bug where we wouldn't promote single store allocas if the stored value was a non-instruction value. Doh. This increase the # single store allocas from 8982 to 9026, and speeds up mem2reg on the testcase in PR1432 from 2.17 to 2.13s. llvm-svn: 40813	2007-08-04 02:45:02 +00:00
Chris Lattner	1b215f0661	When we do the single-store optimization, delete both the store and the alloca so they don't get reprocessed. This speeds up PR1432 from 2.20s to 2.17s. llvm-svn: 40812	2007-08-04 02:38:38 +00:00
Chris Lattner	862f125457	Three improvements: 1. Check for revisiting a block before checking domination, which is faster. 2. If the stored value isn't an instruction, we don't have to check for domination. 3. If we have a value used in the same block more than once, make sure to remove the block from the UsingBlocks vector. Not doing so forces us to go through the slow path for the alloca. The combination of these improvements increases the number of allocas on the fastpath from 8935 to 8982 on PR1432. This speeds it up from 2.90s to 2.20s (31%) llvm-svn: 40811	2007-08-04 02:32:22 +00:00
Chris Lattner	ae1e00eb36	switch from using a std::set to using a SmallPtrSet. This speeds up the testcase in PR1432 from 6.33s to 2.90s (2.22x) llvm-svn: 40810	2007-08-04 02:21:22 +00:00
Chris Lattner	9181801bb7	In mem2reg, when handling the single-store case, make sure to remove a using block from the list if we handle it. Not doing this caused us to not be able to promote (with the fast path) allocas which have uses (whoops). This increases the # allocas hitting this fastpath from 4042 to 8935 on the testcase in PR1432, speeding up mem2reg by 2.6x llvm-svn: 40809	2007-08-04 02:15:24 +00:00
Chris Lattner	886a41a007	split rewriting of single-store allocas into its own method. llvm-svn: 40806	2007-08-04 01:47:41 +00:00
Chris Lattner	3cede09c67	refactor some code to shrink PromoteMem2Reg::run a bit llvm-svn: 40805	2007-08-04 01:41:18 +00:00
Chris Lattner	d524537fe9	add a typedef, no other change. llvm-svn: 40804	2007-08-04 01:19:38 +00:00
Chris Lattner	df138be527	avoid an unneeded vector copy. This speeds up mem2reg on the testcase in PR1432 by 6% llvm-svn: 40803	2007-08-04 01:07:49 +00:00
Chris Lattner	fd838f0770	make RenamePassWorkList a local var instead of an ivar. llvm-svn: 40802	2007-08-04 01:04:40 +00:00
Dan Gohman	34d442f274	More explicit keywords. llvm-svn: 40673	2007-08-01 15:32:29 +00:00
David Greene	17a5dfe6f7	New CallInst interface to address GLIBCXX_DEBUG errors caused by indexing an empty std::vector. Updates to all clients. llvm-svn: 40660	2007-08-01 03:43:44 +00:00
Devang Patel	c5e340eded	LCSSA preserves dom info. llvm-svn: 40604	2007-07-30 20:23:45 +00:00
Devang Patel	e3206cb425	Use SmallPtrSet. llvm-svn: 40560	2007-07-27 18:34:27 +00:00
Dan Gohman	6e853bc73f	Move the GET_SIDE_EFFECT_INFO logic from isInstructionTriviallyDead to Instruction::mayWriteToMemory, fixing a FIXME, and helping various places that call mayWriteToMemory directly. llvm-svn: 40533	2007-07-26 16:06:08 +00:00
Devang Patel	33227115b9	Add BasicInliner interface. This interface allows clients to inline bunch of functions with module level call graph information.:wq llvm-svn: 40486	2007-07-25 18:00:25 +00:00
Devang Patel	a273d1cd3a	Verify loop info. llvm-svn: 40062	2007-07-19 18:02:32 +00:00
Devang Patel	186e0d8b0a	After a basic block is split into two parts, second part dominates all the blocks dominated by original basic block. And first part dominates second part. llvm-svn: 40035	2007-07-19 02:29:24 +00:00
Devang Patel	de5901523c	Now this temp. fix is not required. llvm-svn: 40034	2007-07-19 02:22:21 +00:00
Reid Spencer	3363f4ad96	Return Undef if the block has no dominator. This was required to allow llvm-gcc build to succeed. Without this change it fails in libstdc++ compilation. This causes no regressions in dejagnu tests. However, someone who knows this code better might want to review it. llvm-svn: 39924	2007-07-16 21:03:44 +00:00
Dan Gohman	06c60b6032	Fix comments about vectors to use the current wording. llvm-svn: 39921	2007-07-16 14:29:03 +00:00

1 2 3 4 5 ...

816 Commits