llvm-project

Commit Graph

Author	SHA1	Message	Date
Nate Begeman	e14fdfaecd	SSE 4.1 Intrinsics and detection llvm-svn: 46681	2008-02-03 07:18:54 +00:00
Evan Cheng	32e5347eb8	Get rid of the annoying blank lines before labels. llvm-svn: 46667	2008-02-02 08:39:46 +00:00
Nick Lewycky	f5b9938ef6	Don't use uninitialized values. Fixes vec_align.ll on X86 Linux. llvm-svn: 46666	2008-02-02 08:29:58 +00:00
Evan Cheng	efd142a920	SDIsel processes llvm.dbg.declare by recording the variable debug information descriptor and its corresponding stack frame index in MachineModuleInfo. This only works if the local variable is "homed" in the stack frame. It does not work for byval parameter, etc. Added ISD::DECLARE node type to represent llvm.dbg.declare intrinsic. Now the intrinsic calls are lowered into a SDNode and lives on through out the codegen passes. For now, since all the debugging information recording is done at isel time, when a ISD::DECLARE node is selected, it has the side effect of also recording the variable. This is a short term solution that should be fixed in time. llvm-svn: 46659	2008-02-02 04:07:54 +00:00
Evan Cheng	4e7ff941f1	Frame index can be negative. llvm-svn: 46655	2008-02-02 00:17:00 +00:00
Evan Cheng	d6e44ab5ec	Remove the nasty LABEL hack with a much less evil one. Now llvm.dbg.func.start implies a stoppoint is set. SelectionDAGISel records a new source line but does not create a ISD::LABEL node for this special stoppoint. Asm printer will magically print this label. This ensures nothing is emitted before. llvm-svn: 46635	2008-02-01 09:10:45 +00:00
Evan Cheng	27b32b87ed	Revert 46556 and 46585. Dan please fix the PseudoSourceValue problem and re-commit. llvm-svn: 46623	2008-01-31 21:00:00 +00:00
Evan Cheng	1c6c16ea11	Add an extra operand to LABEL nodes which distinguishes between debug, EH, or misc labels. This fixes the EH breakage. However I am not convinced this is the solution. llvm-svn: 46609	2008-01-31 09:59:15 +00:00
Evan Cheng	6332dbec69	Add x86 specific getFrameIndexOffset(). This fixes local variable debugging info. llvm-svn: 46598	2008-01-31 04:06:00 +00:00
Dan Gohman	ed346f2ed5	Avoid unnecessarily casting away const. llvm-svn: 46590	2008-01-31 01:01:48 +00:00
Dan Gohman	9ba4d76816	Rename ISD::FLT_ROUNDS to ISD::FLT_ROUNDS_ to avoid conflicting with the real FLT_ROUNDS (defined in <float.h>). llvm-svn: 46587	2008-01-31 00:41:03 +00:00
Dan Gohman	3646fdda67	Create a new class, MemOperand, for describing memory references in the backend. Introduce a new SDNode type, MemOperandSDNode, for holding a MemOperand in the SelectionDAG IR, and add a MemOperand list to MachineInstr, and code to manage them. Remove the offset field from SrcValueSDNode; uses of SrcValueSDNode that were using it are all all using MemOperandSDNode now. Also, begin updating some getLoad and getStore calls to use the PseudoSourceValue objects. Most of this was written by Florian Brander, some reorganization and updating to TOT by me. llvm-svn: 46585	2008-01-31 00:25:39 +00:00
Evan Cheng	a3395a61cc	Treat the label for the first @llvm.dbg.stoppoint the same way as the dbg_func_start label. Make sure nothing else is inserted before them. Note this solution might be somewhat fragile since ISD::LABEL may be used for other purposes. If that ends up to be an issue, we may need to introduce a different node for debug labels. llvm-svn: 46571	2008-01-30 20:08:35 +00:00
Evan Cheng	29cfb67e28	Even though InsertAtEndOfBasicBlock is an ugly hack it still deserves a proper name. Rename it to EmitInstrWithCustomInserter since it does not necessarily insert instruction at the end. llvm-svn: 46562	2008-01-30 18:18:23 +00:00
Evan Cheng	ed17ef7e18	Skip over the label which marks the beginning of the function before inserting prologue code. llvm-svn: 46546	2008-01-30 03:57:33 +00:00
Evan Cheng	084a1cdcdd	Work in progress. This patch fixes x86-64 calls which are modelled as StructRet but really should be return in registers, e.g. _Complex long double, some 128-bit aggregates. This is a short term solution that is necessary only because llvm, for now, cannot model i128 nor call's with multiple results. Status: This only works for direct calls, and only the caller side is done. Disabled for now. llvm-svn: 46527	2008-01-29 19:34:22 +00:00
Dale Johannesen	2b3bc30420	Handle 'X' constraint in asm's better. llvm-svn: 46485	2008-01-29 02:21:21 +00:00
Chris Lattner	2e4719ec55	add a note llvm-svn: 46413	2008-01-27 07:31:41 +00:00
Chris Lattner	d05d2011d0	Use fldz and fld1 for long double constants instead of a constant pool load. llvm-svn: 46411	2008-01-27 06:19:31 +00:00
Chris Lattner	2dd23b9f32	Add some notes. llvm-svn: 46405	2008-01-26 20:12:07 +00:00
Chris Lattner	250789f1bd	Remove some code for inferring alignment info from the x86 backend now that the dag combiner does it. llvm-svn: 46404	2008-01-26 20:07:42 +00:00
Bill Wendling	1a17ef02c8	If there's no instructions being emitted on X86 for a function, emit a nop. Emit the nop directly for PPC. llvm-svn: 46398	2008-01-26 09:03:52 +00:00
Chris Lattner	f4523c35cb	optimize fxor like for llvm-svn: 46345	2008-01-25 06:14:17 +00:00
Chris Lattner	84ab724e06	Add target-specific dag combines for FAND(x,0) and FOR(x,0). This allows us to compile: double test(double X) { return copysign(0.0, X); } into: _test: andpd LCPI1_0(%rip), %xmm0 ret instead of: _test: pxor %xmm1, %xmm1 andpd LCPI1_0(%rip), %xmm1 movapd %xmm0, %xmm2 andpd LCPI1_1(%rip), %xmm2 movapd %xmm1, %xmm0 orpd %xmm2, %xmm0 ret llvm-svn: 46344	2008-01-25 05:46:26 +00:00
Anton Korobeynikov	fcde616864	Provide correct DWARF register numbering for debug information emission on x86-32/Darwin. This should fix bunch of issues. llvm-svn: 46337	2008-01-25 00:34:13 +00:00
Chris Lattner	a91f77eaac	Significantly simplify and improve handling of FP function results on x86-32. This case returns the value in ST(0) and then has to convert it to an SSE register. This causes significant codegen ugliness in some cases. For example in the trivial fp-stack-direct-ret.ll testcase we used to generate: _bar: subl $28, %esp call L_foo$stub fstpl 16(%esp) movsd 16(%esp), %xmm0 movsd %xmm0, 8(%esp) fldl 8(%esp) addl $28, %esp ret because we move the result of foo() into an XMM register, then have to move it back for the return of bar. Instead of hacking ever-more special cases into the call result lowering code we take a much simpler approach: on x86-32, fp return is modeled as always returning into an f80 register which is then truncated to f32 or f64 as needed. Similarly for a result, we model it as an extension to f80 + return. This exposes the truncate and extensions to the dag combiner, allowing target independent code to hack on them, eliminating them in this case. This gives us this code for the example above: _bar: subl $12, %esp call L_foo$stub addl $12, %esp ret The nasty aspect of this is that these conversions are not legal, but we want the second pass of dag combiner (post-legalize) to be able to hack on them. To handle this, we lie to legalize and say they are legal, then custom expand them on entry to the isel pass (PreprocessForFPConvert). This is gross, but less gross than the code it is replacing :) This also allows us to generate better code in several other cases. For example on fp-stack-ret-conv.ll, we now generate: _test: subl $12, %esp call L_foo$stub fstps 8(%esp) movl 16(%esp), %eax cvtss2sd 8(%esp), %xmm0 movsd %xmm0, (%eax) addl $12, %esp ret where before we produced (incidentally, the old bad code is identical to what gcc produces): _test: subl $12, %esp call L_foo$stub fstpl (%esp) cvtsd2ss (%esp), %xmm0 cvtss2sd %xmm0, %xmm0 movl 16(%esp), %eax movsd %xmm0, (%eax) addl $12, %esp ret Note that we generate slightly worse code on pr1505b.ll due to a scheduling deficiency that is unrelated to this patch. llvm-svn: 46307	2008-01-24 08:07:48 +00:00
Evan Cheng	35abd840a6	Let each target decide byval alignment. For X86, it's 4-byte unless the aggregare contains SSE vector(s). For x86-64, it's max of 8 or alignment of the type. llvm-svn: 46286	2008-01-23 23:17:41 +00:00
Duncan Sands	95d46ef887	The last pieces needed for loading arbitrary precision integers. This won't actually work (and most of the code is dead) unless the new legalization machinery is turned on. While there, I rationalized the handling of i1, and removed some bogus (and unused) sextload patterns. For i1, this could result in microscopically better code for some architectures (not X86). It might also result in worse code if annotating with AssertZExt nodes turns out to be more harmful than helpful. llvm-svn: 46280	2008-01-23 20:39:46 +00:00
Dale Johannesen	7f1ff5fedd	Honor explicit section information on Darwin. llvm-svn: 46267	2008-01-23 00:58:14 +00:00
Evan Cheng	1e0d4d2aa8	SSE varargs arguments are passed in memory. llvm-svn: 46262	2008-01-22 23:26:53 +00:00
Anton Korobeynikov	da19b1c875	Honour ByVal parameter attribute for name decoration llvm-svn: 46200	2008-01-20 14:00:07 +00:00
Anton Korobeynikov	c7ffe0f4db	Remove Darwin'ism llvm-svn: 46199	2008-01-20 13:59:37 +00:00
Anton Korobeynikov	28d4302807	Enable PIC codegen on x86-64/linux llvm-svn: 46198	2008-01-20 13:58:16 +00:00
Duncan Sands	3e95d963e9	Need to handle any 'nest' parameter before integer parameters, since otherwise it won't be passed in the right register. With this change trampolines work on x86-64 (thanks to Luke Guest for providing access to an x86-64 box). llvm-svn: 46192	2008-01-19 16:42:10 +00:00
Chris Lattner	7dc00e8021	make a method public llvm-svn: 46159	2008-01-18 06:52:41 +00:00
Dale Johannesen	60a9855799	Revert the part of 45848 that treated weak globals as weak globals rather than commons. While not wrong, this change tickled a latent bug in Darwin's strip, so revert it for now as a workaround. llvm-svn: 46144	2008-01-17 23:04:07 +00:00
Chris Lattner	1ea55cf816	This commit changes: 1. Legalize now always promotes truncstore of i1 to i8. 2. Remove patterns and gunk related to truncstore i1 from targets. 3. Rename the StoreXAction stuff to TruncStoreAction in TLI. 4. Make the TLI TruncStoreAction table a 2d table to handle from/to conversions. 5. Mark a wide variety of invalid truncstores as such in various targets, e.g. X86 currently doesn't support truncstore of any of its integer types. 6. Add legalize support for truncstores with invalid value input types. 7. Add a dag combine transform to turn store(truncate) into truncstore when safe. The later allows us to compile CodeGen/X86/storetrunc-fp.ll to: _foo: fldt 20(%esp) fldt 4(%esp) faddp %st(1) movl 36(%esp), %eax fstps (%eax) ret instead of: _foo: subl $4, %esp fldt 24(%esp) fldt 8(%esp) faddp %st(1) fstps (%esp) movl 40(%esp), %eax movss (%esp), %xmm0 movss %xmm0, (%eax) addl $4, %esp ret llvm-svn: 46140	2008-01-17 19:59:44 +00:00
Chris Lattner	72733e573b	* Introduce a new SelectionDAG::getIntPtrConstant method and switch various codegen pieces and the X86 backend over to using it. * Add some comments to SelectionDAGNodes.h * Introduce a second argument to FP_ROUND, which indicates whether the FP_ROUND changes the value of its input. If not it is safe to xform things like fp_extend(fp_round(x)) -> x. llvm-svn: 46125	2008-01-17 07:00:52 +00:00
Duncan Sands	32b0ff6814	Trampoline support for x86-64. This looks like it should work, but I have no machine to test it on. Committed because it will at least cause no harm, and maybe someone can test it for me! llvm-svn: 46098	2008-01-16 22:55:25 +00:00
Chris Lattner	e8bb9f2190	make it more clear that this predicate only applies to scalar FP types. llvm-svn: 46058	2008-01-16 06:24:21 +00:00
Chris Lattner	14e616ef0b	introduce a isTypeInSSEReg predicate, which allows us to simplify some code. No functionality change. llvm-svn: 46055	2008-01-16 06:19:45 +00:00
Chris Lattner	8f7cec859e	My previous commit had an incomplete message, it should have been: make the 'fp return in ST(0)' optimization smart enough to look through token factor nodes. THis allows us to compile testcases like CodeGen/X86/fp-stack-retcopy.ll into: _carg: subl $12, %esp call L_foo$stub fstpl (%esp) fldl (%esp) addl $12, %esp ret instead of: _carg: subl $28, %esp call L_foo$stub fstpl 16(%esp) movsd 16(%esp), %xmm0 movsd %xmm0, 8(%esp) fldl 8(%esp) addl $28, %esp ret Still not optimal, but much better and this is a trivial patch. Fixing the rest requires invasive surgery that is is not llvm 2.2 material. llvm-svn: 46054	2008-01-16 05:56:59 +00:00
Chris Lattner	ea001f1db7	make the 'fp return in ST(0)' optimization smart enough to look through token factor llvm-svn: 46053	2008-01-16 05:53:06 +00:00
Chris Lattner	de5c74f18e	various whitespace cleanups, no functionality change. llvm-svn: 46052	2008-01-16 05:52:18 +00:00
Dale Johannesen	59a2250b0d	Fix and enable EH for x86-64 Darwin. Adds ShortenEHDataFor64Bits as a not-very-accurate abstraction to cover all the changes in DwarfWriter. Some cosmetic changes to Darwin assembly code for gcc testsuite compatibility. llvm-svn: 46029	2008-01-15 23:24:56 +00:00
Chris Lattner	9a249b0ce5	rename SDTRet -> SDTNone. Move definition of 'trap' sdnode up from x86 instrinfo to targetselectiondag.td. llvm-svn: 46017	2008-01-15 22:02:54 +00:00
Chris Lattner	3c3fefde06	no need to expand ISD::TRAP to X86ISD::TRAP, just match ISD::TRAP. llvm-svn: 46015	2008-01-15 21:58:22 +00:00
Anton Korobeynikov	59e6d533bd	Fix JIT encoding of trap/ud2 instruction llvm-svn: 46012	2008-01-15 21:40:02 +00:00
Anton Korobeynikov	6bbbc4cbfa	For PR1839: add initial support for __builtin_trap. llvm-gcc part is missed as well as PPC codegen llvm-svn: 46001	2008-01-15 07:02:33 +00:00
Evan Cheng	4d70ba3134	Rename CCIfStruct to CCIfByVal and CCStructAssign to CCPassByVal. Remove unused parameters of CCStructAssign and add size and alignment requirement info. llvm-svn: 45997	2008-01-15 03:34:58 +00:00
Evan Cheng	48bdfe63e2	Both x86-32 and x86-64 handle byval parameter attributes. llvm-svn: 45996	2008-01-15 03:15:41 +00:00
Chris Lattner	3c43efc9d1	Improve the FP stackifier to decide all on its own whether an instruction kills a register or not. This is cheap and easy to do now that instructions record this on their flags, and this eliminates the second pass of LiveVariables from the x86 backend. This speeds up a release llc by ~2.5%. llvm-svn: 45955	2008-01-14 06:41:29 +00:00
Duncan Sands	51fe7bbcf5	Whitespace tweak. llvm-svn: 45940	2008-01-13 21:20:29 +00:00
Evan Cheng	7411b510b2	Code clean up. llvm-svn: 45898	2008-01-12 01:08:07 +00:00
Chris Lattner	18df33d0c8	fix a wordo that gordon noticed :) llvm-svn: 45896	2008-01-12 00:53:16 +00:00
Chris Lattner	6da61c2515	Any x86 instruction that reads from an invariant location is invariant. This allows us to sink things like: cvtsi2sd 32(%esp), %xmm1 when reading from the argument area, for example. llvm-svn: 45895	2008-01-12 00:35:08 +00:00
Chris Lattner	596875118c	rename MachineInstr::setInstrDescriptor -> setDesc llvm-svn: 45871	2008-01-11 18:10:50 +00:00
Chris Lattner	806dd0e2ac	remove xchg and shift-reg-by-1 instructions, which are dead. llvm-svn: 45870	2008-01-11 18:00:50 +00:00
Chris Lattner	ff5998e66b	add a note, remove a done deed. llvm-svn: 45869	2008-01-11 18:00:13 +00:00
Arnold Schwaighofer	06da9e2d43	hrm - correct spelling. Actually were not riding any arguments. Sadly there is no semantic spell checker that is going to safe you from such a mistake. llvm-svn: 45868	2008-01-11 17:10:15 +00:00
Arnold Schwaighofer	6cf72fbbaf	Improve tail call optimized call's argument lowering. Before this commit all arguments where moved to the stack slot where they would reside on a normal function call before the lowering to the tail call stack slot. This was done to prevent arguments overwriting each other. Now only arguments sourcing from a FORMAL_ARGUMENTS node or a CopyFromReg node with virtual register (could also be a caller's argument) are lowered indirectly. --This line, and those below, will be ignored-- M X86/X86ISelLowering.cpp M X86/README.txt llvm-svn: 45867	2008-01-11 16:49:42 +00:00
Arnold Schwaighofer	bf1816ea7b	Correct a copy and paste error. llvm-svn: 45865	2008-01-11 14:34:56 +00:00
Evan Cheng	8c51394e01	Rename Int_CVTSI642SSr* to Int_CVTSI2SS64r* for naming consistency and remove unused instructions. llvm-svn: 45861	2008-01-11 07:37:44 +00:00
Chris Lattner	9283173061	more flags set right llvm-svn: 45860	2008-01-11 07:18:17 +00:00
Chris Lattner	f4b0c99d63	add some missing flags. llvm-svn: 45859	2008-01-11 06:59:07 +00:00
Dale Johannesen	2ff66f08f2	Weak things initialized to 0 don't go in bss on Darwin. Cosmetic changes to spacing to match gcc (some dejagnu tests actually care). llvm-svn: 45848	2008-01-11 00:54:37 +00:00
Chris Lattner	c8226f32e9	Simplify the side effect stuff a bit more and make licm/sinking both work right according to the new flags. This removes the TII::isReallySideEffectFree predicate, and adds TII::isInvariantLoad. It removes NeverHasSideEffects+MayHaveSideEffects and adds UnmodeledSideEffects as machine instr flags. Now the clients can decide everything they need. I think isRematerializable can be implemented in terms of the flags we have now, though I will let others tackle that. llvm-svn: 45843	2008-01-10 23:08:24 +00:00
Chris Lattner	8e60f2c996	IMPLICIT_USE and IMPLICIT_DEF are dead, remove them. llvm-svn: 45838	2008-01-10 19:27:54 +00:00
Chris Lattner	317332fc2a	Start inferring side effect information more aggressively, and fix many bugs in the x86 backend where instructions were not marked maystore/mayload, and perf issues where instructions were not marked neverHasSideEffects. It would be really nice if we could write patterns for copy instructions. I have audited all the x86 instructions down to MOVDQAmr. The flags on others and on other targets are probably not right in all cases, but no clients currently use this info that are enabled by default. llvm-svn: 45829	2008-01-10 07:59:24 +00:00
Chris Lattner	2e38f2458c	rename X86InstrX86-64.td -> X86Instr64bit.td llvm-svn: 45826	2008-01-10 05:50:42 +00:00
Chris Lattner	aca7ca3730	remove explicit sets of 'neverHasSideEffects' that can now be inferred from the instr patterns. llvm-svn: 45824	2008-01-10 05:45:39 +00:00
Chris Lattner	94de7bc3aa	get def use info more correct. llvm-svn: 45821	2008-01-10 05:12:37 +00:00
Chris Lattner	f171482a66	verify that the frame index is immutable before remat'ing (still disabled) or being side-effect free. llvm-svn: 45816	2008-01-10 04:16:31 +00:00
Evan Cheng	a26552493b	Mark byval parameter stack objects mutable for now. llvm-svn: 45813	2008-01-10 02:24:25 +00:00
Dale Johannesen	7ecb3b79c7	Emit unused EH frames for weak definitions on Darwin, because assembler/linker can't cope with weak absolutes. PR 1880. llvm-svn: 45811	2008-01-10 02:03:30 +00:00
Evan Cheng	fead113fe0	Do not use the stack pointer directly, issue a copyfromreg instead. Otherwise we can end up with something like ADD32ri %esp, x which two-address pass won't like. llvm-svn: 45798	2008-01-10 00:37:26 +00:00
Evan Cheng	73d1017871	Remove comments that do not correspond to anything after recent refactoring. llvm-svn: 45792	2008-01-10 00:09:10 +00:00
Chris Lattner	9129f51f9b	add a testcase llvm-svn: 45768	2008-01-09 00:37:18 +00:00
Duncan Sands	bb956ca730	Use size_t to store Pos, avoid truncating value on 64-bit builds. Analysis and original patch by Török Edwin. Code audit found another place with the same problem, also fixed here. llvm-svn: 45746	2008-01-08 10:06:15 +00:00
Evan Cheng	00300ddff1	Minor fix to enable x86-64 pic jit (still fails for other reasons). llvm-svn: 45734	2008-01-08 02:07:10 +00:00
Evan Cheng	4951da49aa	Fix a x86-64 static codegen bug. This fixes a lot of x86-64 jit failures. llvm-svn: 45733	2008-01-08 02:06:11 +00:00
Bill Wendling	3b6fe5fa8d	Silence warning about loss of precision. llvm-svn: 45731	2008-01-08 00:52:29 +00:00
Evan Cheng	8242168ef4	Unbreak x86-64. llvm-svn: 45725	2008-01-07 23:08:23 +00:00
Chris Lattner	ef81aa75f6	add a note that is important for some fp apps. llvm-svn: 45723	2008-01-07 21:59:58 +00:00
Duncan Sands	d8d4170f84	Unbreak x86-32 darwin long double! llvm-svn: 45703	2008-01-07 16:36:38 +00:00
Duncan Sands	28bf7ac219	Fix long double support on x86-32 linux. llvm-svn: 45701	2008-01-07 13:44:22 +00:00
Bill Wendling	a3bdad153f	Operand 1 should be a register. We don't care if it's a preg, vreg, or 0. llvm-svn: 45699	2008-01-07 08:05:29 +00:00
Chris Lattner	03ad885039	rename TargetInstrDescriptor -> TargetInstrDesc. Make MachineInstr::getDesc return a reference instead of a pointer, since it can never be null. llvm-svn: 45695	2008-01-07 07:27:27 +00:00
Chris Lattner	f376c99ea0	rename hasVariableOperands() -> isVariadic(). Add some comments. Evan, please review the comments I added to getNumDefs to make sure that they are accurate, thx. llvm-svn: 45687	2008-01-07 05:19:29 +00:00
Chris Lattner	b0d06b4381	Move a bunch more accessors from TargetInstrInfo to TargetInstrDescriptor llvm-svn: 45680	2008-01-07 03:13:06 +00:00
Chris Lattner	f0f438a517	remove MachineOpCode typedef. llvm-svn: 45679	2008-01-07 02:48:55 +00:00
Chris Lattner	e55e115616	Add predicates methods to TargetOperandInfo, and switch all clients over to using them, instead of diddling Flags directly. Change the various flags from const variables to enums. llvm-svn: 45677	2008-01-07 02:39:19 +00:00
Chris Lattner	a98c679de0	Rename MachineInstr::getInstrDescriptor -> getDesc(), which reflects that it is cheap and efficient to get. Move a variety of predicates from TargetInstrInfo into TargetInstrDescriptor, which makes it much easier to query a predicate when you don't have TII around. Now you can use MI->getDesc()->isBranch() instead of going through TII, and this is much more efficient anyway. Not all of the predicates have been moved over yet. Update old code that used MI->getInstrDescriptor()->Flags to use the new predicates in many places. llvm-svn: 45674	2008-01-07 01:56:04 +00:00
Owen Anderson	2a3be7bb6c	Move even more functionality from MRegisterInfo into TargetInstrInfo. Some day I'll get it all moved over... llvm-svn: 45672	2008-01-07 01:35:02 +00:00
Chris Lattner	b296b0f1c1	The pic base can't be duplicated. llvm-svn: 45668	2008-01-06 23:49:32 +00:00
Chris Lattner	a4ce4f6987	rename isLoad -> isSimpleLoad due to evan's desire to have such a predicate. llvm-svn: 45667	2008-01-06 23:38:27 +00:00
Bill Wendling	5fa2c64b78	Fix comment. llvm-svn: 45638	2008-01-05 23:30:51 +00:00
Nate Begeman	22950d26f5	Remove an incorrect optimization that is performed correctly by the target independent legalizer. llvm-svn: 45631	2008-01-05 20:51:30 +00:00
Gordon Henriksen	9231958391	Refactoring the x86 and x86-64 calling convention implementations, unifying the copied algorithms and saving over 500 LOC. There should be no functionality change, but please test on your favorite x86 target. llvm-svn: 45627	2008-01-05 16:56:59 +00:00
Bill Wendling	be984cf10b	Chris and Evan noticed that this check was compleatly fubared. I was checking that there was a from a global instead of a load from the stub for a global, which is the one that's safe to hoist. Consider this program: volatile char G[100]; int B(char *F, int N) { for (; N > 0; --N) F[N] = G[N]; } In static mode, we shouldn't be hoisting the load from G: $ llc -relocation-model=static -o - a.bc -march=x86 -machine-licm LBB1_1: # bb.preheader leal -1(%eax), %edx testl %edx, %edx movl $1, %edx cmovns %eax, %edx xorl %esi, %esi LBB1_2: # bb movb _G(%eax), %bl movb %bl, (%ecx,%eax) llvm-svn: 45626	2008-01-05 09:18:04 +00:00
Chris Lattner	e0f0c4aa02	enable sinking and licm of loads from the argument area. I'd like to enable this for remat, but can't due to an RA bug. llvm-svn: 45623	2008-01-05 06:10:42 +00:00
Chris Lattner	d4738ee8e1	simplify some code by using shorter accessors. llvm-svn: 45622	2008-01-05 05:28:30 +00:00
Chris Lattner	86f4a2e4f2	revert my previous patch. llvm-svn: 45621	2008-01-05 05:26:26 +00:00
Chris Lattner	69d7902cf1	factor some code better to avoid redundancy between isReallySideEffectFree and isReallyTriviallyReMaterializable. Why is a load from a global considered side-effect-free but not rematable? llvm-svn: 45620	2008-01-05 05:19:56 +00:00
Chris Lattner	9fa8ae6c6b	getting the pic base has no side effects. llvm-svn: 45618	2008-01-05 03:54:32 +00:00
Evan Cheng	880b080887	X86 JIT PIC jumptable support. llvm-svn: 45616	2008-01-05 02:26:58 +00:00
Evan Cheng	f55b7381af	Combine MovePCtoStack + POP32r into one instruction MOVPC32r so it can be moved if needed. llvm-svn: 45605	2008-01-05 00:41:47 +00:00
Owen Anderson	6bb0c52628	Move some more functionality from MRegisterInfo to TargetInstrInfo. llvm-svn: 45603	2008-01-04 23:57:37 +00:00
Evan Cheng	c1d1e54fc4	Unbreak tailcall opt in JIT. llvm-svn: 45576	2008-01-04 10:50:28 +00:00
Evan Cheng	49ff8ecd03	X86 PIC JIT support fixes: encoding bugs, add lazy pointer stubs support. llvm-svn: 45575	2008-01-04 10:46:51 +00:00
Gordon Henriksen	f066fc477c	First steps in in X86 calling convention cleanup. llvm-svn: 45536	2008-01-03 16:47:34 +00:00
Evan Cheng	563fcc3428	Change MachineRelocation::DoesntNeedFnStub to NeedStub. This fields will be used for non-function GV relocations that require function address stubs (e.g. Mac OS X in non-static mode). llvm-svn: 45527	2008-01-03 02:56:28 +00:00
Evan Cheng	96334b4e3b	X86 PIC JIT bug fix: relocations for constantpool and jumptable. llvm-svn: 45515	2008-01-02 23:38:59 +00:00
Bill Wendling	e1f28e7871	Machine LICM will check that operands are defined outside of the loop. Also check that register isn't 0 before going further. llvm-svn: 45498	2008-01-02 21:10:40 +00:00
Chris Lattner	cce79c67ca	darwin9 and above support aligned common symbols. llvm-svn: 45494	2008-01-02 19:44:55 +00:00
Owen Anderson	eee14601b1	Move some more instruction creation methods from RegisterInfo into InstrInfo. llvm-svn: 45484	2008-01-01 21:11:32 +00:00
Chris Lattner	b0fb17fd34	Fix a bug in my previous patch: refer to the impl not the pure virtual version. It's unclear why gcc would ever compile this... llvm-svn: 45476	2008-01-01 01:05:34 +00:00
Chris Lattner	25568e4cef	Fix a problem where lib/Target/TargetInstrInfo.h would include and use a header file from libcodegen. This violates a layering order: codegen depends on target, not the other way around. The fix to this is to split TII into two classes, TII and TargetInstrInfoImpl, which defines stuff that depends on libcodegen. It is defined in libcodegen, where the base is not. llvm-svn: 45475	2008-01-01 01:03:04 +00:00
Owen Anderson	7a73ae9a86	Move copyRegToReg from MRegisterInfo to TargetInstrInfo. This is part of the Machine-level API cleanup instigated by Chris. llvm-svn: 45470	2007-12-31 06:32:00 +00:00
Chris Lattner	a10fff51d9	Rename SSARegMap -> MachineRegisterInfo in keeping with the idea that "machine" classes are used to represent the current state of the code being compiled. Given this expanded name, we can start moving other stuff into it. For now, move the UsedPhysRegs and LiveIn/LoveOuts vectors from MachineFunction into it. Update all the clients to match. This also reduces some needless #includes, such as MachineModuleInfo from MachineFunction. llvm-svn: 45467	2007-12-31 04:13:23 +00:00
Chris Lattner	a5bb370aa4	Add new shorter predicates for testing machine operands for various types: e.g. MO.isMBB() instead of MO.isMachineBasicBlock(). I don't plan on switching everything over, so new clients should just start using the shorter names. Remove old long accessors, switching everything over to use the short accessor: getMachineBasicBlock() -> getMBB(), getConstantPoolIndex() -> getIndex(), setMachineBasicBlock -> setMBB(), etc. llvm-svn: 45464	2007-12-30 23:10:15 +00:00
Chris Lattner	6005589faf	More cleanups for MachineOperand: - Eliminate the static "print" method for operands, moving it into MachineOperand::print. - Change various set* methods for register flags to take a bool for the value to set it to. Remove unset* methods. - Group methods more logically by operand flavor in MachineOperand.h llvm-svn: 45461	2007-12-30 21:56:09 +00:00
Chris Lattner	5c4637816e	Use MachineOperand::getImm instead of MachineOperand::getImmedValue. Likewise setImmedValue -> setImm llvm-svn: 45453	2007-12-30 20:49:49 +00:00
Bill Wendling	7749a9014b	If we have a load of a global address that's not modified during the function, then go ahead and hoist it out of the loop. This is the result: $ cat a.c volatile int G; int A(int N) { for (; N > 0; --N) G++; } $ llc -o - -relocation-model=pic _A: ... LBB1_2: # bb movl L_G$non_lazy_ptr-"L1$pb"(%eax), %esi incl (%esi) incl %edx cmpl %ecx, %edx jne LBB1_2 # bb ... $ llc -o - -relocation-model=pic -machine-licm _A: ... movl L_G$non_lazy_ptr-"L1$pb"(%eax), %eax LBB1_2: # bb incl (%eax) incl %edx cmpl %ecx, %edx jne LBB1_2 # bb ... I'm limiting this to the MOV32rm x86 instruction for now. llvm-svn: 45444	2007-12-30 03:18:58 +00:00
Chris Lattner	4b762496a9	Shrinkify the machine operand creation method names. llvm-svn: 45433	2007-12-30 00:45:46 +00:00
Chris Lattner	f3ebc3f3d2	Remove attribution from file headers, per discussion on llvmdev. llvm-svn: 45418	2007-12-29 20:36:04 +00:00
Chris Lattner	a087a8d2ce	remove attribution from lib Makefiles. llvm-svn: 45415	2007-12-29 20:09:26 +00:00
Chris Lattner	d2b8a36f0e	One readme entry is done, one is really easy (Evan, want to investigate eliminating the llvm.x86.sse2.loadl.pd intrinsic?), one shuffle optzn may be done (if shufps is better than pinsw, Evan, please review), and we already know about LICM of simple instructions. llvm-svn: 45407	2007-12-29 19:31:47 +00:00
Chris Lattner	3b6a82118b	Fold comparisons against a constant nan, and optimize ORD/UNORD comparisons with a constant. This allows us to compile isnan to: _foo: fcmpu cr7, f1, f1 mfcr r2 rlwinm r3, r2, 0, 31, 31 blr instead of: LCPI1_0: ; float .space 4 _foo: lis r2, ha16(LCPI1_0) lfs f0, lo16(LCPI1_0)(r2) fcmpu cr7, f1, f0 mfcr r2 rlwinm r3, r2, 0, 31, 31 blr llvm-svn: 45405	2007-12-29 08:37:08 +00:00
Chris Lattner	33de0c6e92	this xform is implemented. llvm-svn: 45404	2007-12-29 08:19:39 +00:00
Chris Lattner	07ccbfa64a	Codegen: as: _bar: pushl %esi subl $8, %esp movl 16(%esp), %esi call L_foo$stub fstps (%esi) addl $8, %esp popl %esi #FP_REG_KILL ret instead of: _bar: pushl %esi subl $8, %esp movl 16(%esp), %esi call L_foo$stub fstpl (%esi) cvtsd2ss (%esi), %xmm0 movss %xmm0, (%esi) addl $8, %esp popl %esi #FP_REG_KILL ret llvm-svn: 45401	2007-12-29 06:57:38 +00:00
Chris Lattner	8013bd339b	avoid going through a stack slot to convert from fpstack to xmm reg if we are just going to store it back anyway. This improves things like: double foo(); void bar(double P) { P = foo(); } llvm-svn: 45399	2007-12-29 06:41:28 +00:00
Chris Lattner	62ba67c0f7	add a note llvm-svn: 45397	2007-12-29 05:51:58 +00:00
Chris Lattner	d798002401	add a note. llvm-svn: 45387	2007-12-28 21:50:40 +00:00
Chris Lattner	6c234bf58f	add a simple hack llvm-svn: 45343	2007-12-24 19:27:46 +00:00
Anton Korobeynikov	1b1250770c	Erm, really disable :) llvm-svn: 45319	2007-12-22 20:46:24 +00:00
Anton Korobeynikov	d34112a4ee	Disable, until we'll really need it llvm-svn: 45318	2007-12-22 20:41:12 +00:00
Evan Cheng	345a00ba05	Preliminary PIC JIT support for X86 (32-bit) / Darwin. llvm-svn: 45313	2007-12-22 09:40:20 +00:00
Evan Cheng	db33a0211b	Oops. llvm-svn: 45312	2007-12-22 09:14:34 +00:00
Evan Cheng	f4f52dbc8c	Fix JIT code emission of X86::MovePCtoStack. llvm-svn: 45307	2007-12-22 02:26:46 +00:00
Evan Cheng	ac134551c6	Allow JIT with non-static relocation model. llvm-svn: 45304	2007-12-22 01:12:14 +00:00
Evan Cheng	78c460c8c4	New entry. llvm-svn: 45280	2007-12-21 01:31:58 +00:00
Evan Cheng	01c7c198ee	Fix JIT encoding for CMPSD as well. llvm-svn: 45268	2007-12-20 19:57:09 +00:00
Chris Lattner	2583a66295	add an obvious load folding missed optzn. llvm-svn: 45161	2007-12-18 16:48:14 +00:00
Bill Wendling	b3d85a5d4b	Add "mayHaveSideEffects" and "neverHasSideEffects" flags to some instructions. I based what flag to set on whether it was already marked as "isRematerializable". If there was a further check to determine if it's "really" rematerializable, then I marked it as "mayHaveSideEffects" and created a check in the X86 back-end similar to the remat one. llvm-svn: 45132	2007-12-17 23:07:56 +00:00
Bill Wendling	a2401be121	LD_Fp64m should have "isRematerializable" set. llvm-svn: 45128	2007-12-17 22:17:14 +00:00
Chris Lattner	e3b05fe31b	fix a questionable cast, thanks to Mike Stump for pointing this out. llvm-svn: 45075	2007-12-16 20:26:54 +00:00
Chris Lattner	dab6bd902e	Fix the JIT encoding of cmpss, which aborts with this assertion currently: X86CodeEmitter.cpp:378: failed assertion `0 && "Immediate size not set!"' I think* this is right, but Evan, please verify. It also looks like CMPSDrr and maybe others are missing this info. Evan, plz investigate. llvm-svn: 45074	2007-12-16 20:12:41 +00:00
Evan Cheng	23d2d4dc6c	Make better use of instructions that clear high bits; fix various 2-wide shuffle bugs. llvm-svn: 45058	2007-12-15 03:00:47 +00:00
Evan Cheng	9556729128	Actually, MOVPQIto64mr is a dup of MOVPQI2QImr, MOV64toPQIrm is a dup of MOVQI2PQIrm. llvm-svn: 45041	2007-12-14 20:08:14 +00:00
Evan Cheng	e28372c0d6	Fix (mem) <-> low 64-bits of xmm bugs pointed out by David Greene. Mac OS X Leopard assembler recognizes movq. llvm-svn: 45040	2007-12-14 19:54:07 +00:00
Dale Johannesen	f7cefdd5f0	x86-32 long doubles are 4-byte aligned on the stack for parameter passing (only for that, on Darwin). llvm-svn: 45038	2007-12-14 19:25:34 +00:00
Evan Cheng	a56e6ff9a7	Fix bsf / bsr jit encoding. llvm-svn: 45037	2007-12-14 18:49:43 +00:00
Evan Cheng	f28c810036	Oops. Forgot these. llvm-svn: 45036	2007-12-14 18:25:34 +00:00
Dan Gohman	9d2e9e376f	Fix Intel asm syntax for the bsr and bsf instructions. llvm-svn: 45030	2007-12-14 15:10:00 +00:00
Evan Cheng	0e6408124e	Fix ctlz and cttz. llvm definition requires them to return number of bits in of the src type when value is zero. llvm-svn: 45029	2007-12-14 08:30:15 +00:00
Evan Cheng	e9fbc3f014	Implement ctlz and cttz with bsr and bsf. llvm-svn: 45024	2007-12-14 02:13:44 +00:00
Evan Cheng	827d30db19	Fold some and + shift in x86 addressing mode. llvm-svn: 44970	2007-12-13 00:43:27 +00:00
Evan Cheng	6e68381e02	Implicit def instructions, e.g. X86::IMPLICIT_DEF_GR32, are always re-materializable and they should not be spilled. llvm-svn: 44960	2007-12-12 23:12:09 +00:00
Dan Gohman	7a7742c2fe	Allow vector integer constants to be created with SelectionDAG::getConstant, in the same way as vector floating-point constants. This allows the legalize expansion code for @llvm.ctpop and friends to be usable with vector types. llvm-svn: 44954	2007-12-12 22:21:26 +00:00
Evan Cheng	0f42730722	Use shuffles to implement insert_vector_elt for i32, i64, f32, and f64. llvm-svn: 44929	2007-12-12 07:55:34 +00:00
Evan Cheng	2a98956796	Lower a build_vector with all constants into a constpool load unless it can be done with a move to low part. llvm-svn: 44921	2007-12-12 06:45:40 +00:00
Scott Michel	4a8bc7e105	Correct typo for Linux: s/esp/%rsp/ llvm-svn: 44904	2007-12-12 02:38:28 +00:00
Nate Begeman	6dc8b4ed13	Allow the JIT to encode MMX instructions llvm-svn: 44869	2007-12-11 18:06:14 +00:00
Evan Cheng	4fbf459549	- Improved v8i16 shuffle lowering. It now uses pshuflw and pshufhw as much as possible before resorting to pextrw and pinsrw. - Better codegen for v4i32 shuffles masquerading as v8i16 or v16i8 shuffles. - Improves (i16 extract_vector_element 0) codegen by recognizing (i32 extract_vector_element 0) does not require a pextrw. llvm-svn: 44836	2007-12-11 01:46:18 +00:00
Nate Begeman	a55a67ae91	x86 doesn't actually want to custom lower v3i32 llvm-svn: 44835	2007-12-11 01:41:33 +00:00
Anton Korobeynikov	21ade5880b	Hey, English is not my native language :) llvm-svn: 44820	2007-12-10 23:10:20 +00:00
Anton Korobeynikov	77eb5e649d	Clarify the need of CFI() stuff llvm-svn: 44819	2007-12-10 23:08:35 +00:00
Anton Korobeynikov	a6b0f7e244	Provide convenient way to disable CFI stuff for old/broken assemblers. Use it for Darwin. llvm-svn: 44818	2007-12-10 23:04:38 +00:00
Chris Lattner	8a72a7d586	Disable cfi directives for now, darwin does't support them. These should probably be something like: CFI(".cfi_def_cfa_offset 16\n") where CFI is defined to a noop on darwin and other platforms that don't support those directives. llvm-svn: 44803	2007-12-10 19:10:18 +00:00
Anton Korobeynikov	657be86229	And finally annotate X86-64 version of callback. All bad stuff from SSE version is implicitely inherited :) llvm-svn: 44794	2007-12-10 15:27:07 +00:00
Anton Korobeynikov	88e9d082d8	Provide annotation for SSE version of callback. It's even more broken, because doesn't mark xmm regs properly llvm-svn: 44793	2007-12-10 15:13:55 +00:00
Anton Korobeynikov	81e9dc4af7	Annotate JIT callback function with call frame infromation. This will allow us (theoretically) to unwind through JITer. The code wasn't verified, so I'm pretty sure offsets are wrong :) llvm-svn: 44792	2007-12-10 14:54:42 +00:00
Bill Wendling	3f19dfe794	Reverting 44702. It wasn't correct to rename them. llvm-svn: 44727	2007-12-08 23:58:46 +00:00
Chris Lattner	ff87f05e43	aesthetic changes, no functionality change. Evan, it's not clear what 'Available' is, please add a comment near it and rename it if appropriate. llvm-svn: 44703	2007-12-08 07:22:58 +00:00
Bill Wendling	2b07d8c5a0	Renaming: isTriviallyReMaterializable -> hasNoSideEffects isReallyTriviallyReMaterializable -> isTriviallyReMaterializable llvm-svn: 44702	2007-12-08 07:17:56 +00:00
Evan Cheng	b41d838d28	Add comment. llvm-svn: 44686	2007-12-07 21:30:01 +00:00
Evan Cheng	bfd373a53e	Much improved v8i16 shuffles. (Step 1). llvm-svn: 44676	2007-12-07 08:07:39 +00:00
Evan Cheng	c829e5cdf0	Remove a bogus optimization. It's not possible to do a move to low element to a <8 x i16> or <16 x i8> vector. llvm-svn: 44669	2007-12-06 22:14:22 +00:00
Chris Lattner	ad05e17491	add a note llvm-svn: 44637	2007-12-05 22:58:19 +00:00
Evan Cheng	bb26301864	Add a argument to storeRegToStackSlot and storeRegToAddr to specify whether the stored register is killed. llvm-svn: 44600	2007-12-05 03:14:33 +00:00
Evan Cheng	f45a1d623c	Remove redundant foldMemoryOperand variants and other code clean up. llvm-svn: 44517	2007-12-02 08:30:39 +00:00
Evan Cheng	69fda0a716	Allow some reloads to be folded in multi-use cases. Specifically testl r, r -> cmpl [mem], 0. llvm-svn: 44479	2007-12-01 02:07:52 +00:00
Nate Begeman	6f026a654c	Support returning non-power-of-2 vectors to unblock some work llvm-svn: 44371	2007-11-27 19:28:48 +00:00
Duncan Sands	ad0ea2d430	Fix PR1146: parameter attributes are longer part of the function type, instead they belong to functions and function calls. This is an updated and slightly corrected version of Reid Spencer's original patch. The only known problem is that auto-upgrading of bitcode files doesn't seem to work properly (see test/Bitcode/AutoUpgradeIntrinsics.ll). Hopefully a bitcode guru (who might that be? :) ) will fix it. llvm-svn: 44359	2007-11-27 13:23:08 +00:00
Chris Lattner	5728bdd4db	Fix a long standing deficiency in the X86 backend: we would sometimes emit "zero" and "all one" vectors multiple times, for example: _test2: pcmpeqd %mm0, %mm0 movq %mm0, _M1 pcmpeqd %mm0, %mm0 movq %mm0, _M2 ret instead of: _test2: pcmpeqd %mm0, %mm0 movq %mm0, _M1 movq %mm0, _M2 ret This patch fixes this by always arranging for zero/one vectors to be defined as v4i32 or v2i32 (SSE/MMX) instead of letting them be any random type. This ensures they get trivially CSE'd on the dag. This fix is also important for LegalizeDAGTypes, as it gets unhappy when the x86 backend wants BUILD_VECTOR(i64 0) to be legal even when 'i64' isn't legal. This patch makes the following changes: 1) X86TargetLowering::LowerBUILD_VECTOR now lowers 0/1 vectors into their canonical types. 2) The now-dead patterns are removed from the SSE/MMX .td files. 3) All the patterns in the .td file that referred to immAllOnesV or immAllZerosV in the wrong form now use *_bc to match them with a bitcast wrapped around them. 4) X86DAGToDAGISel::SelectScalarSSELoad is generalized to handle bitcast'd zero vectors, which simplifies the code actually. 5) getShuffleVectorZeroOrUndef is updated to generate a shuffle that is legal, instead of generating one that is illegal and expecting a later legalize pass to clean it up. 6) isZeroShuffle is generalized to handle bitcast of zeros. 7) several other minor tweaks. This patch is definite goodness, but has the potential to cause random code quality regressions. Please be on the lookout for these and let me know if they happen. llvm-svn: 44310	2007-11-25 00:24:49 +00:00
Chris Lattner	f72ad16263	remove bogus assertion that broke CodeGen/Generic/cast-fp.ll on x86 among others. llvm-svn: 44302	2007-11-24 18:37:20 +00:00
Chris Lattner	f81d5886c6	Several changes: 1) Change the interface to TargetLowering::ExpandOperationResult to take and return entire NODES that need a result expanded, not just the value. This allows us to handle things like READCYCLECOUNTER, which returns two values. 2) Implement (extremely limited) support in LegalizeDAG::ExpandOp for MERGE_VALUES. 3) Reimplement custom lowering in LegalizeDAGTypes in terms of the new ExpandOperationResult. This makes the result simpler and fully general. 4) Implement (fully general) expand support for MERGE_VALUES in LegalizeDAGTypes. 5) Implement ExpandOperationResult support for ARM f64->i64 bitconvert and ARM i64 shifts, allowing them to work with LegalizeDAGTypes. 6) Implement ExpandOperationResult support for X86 READCYCLECOUNTER and FP_TO_SINT, allowing them to work with LegalizeDAGTypes. LegalizeDAGTypes now passes several more X86 codegen tests when enabled and when type legalization in LegalizeDAG is ifdef'd out. llvm-svn: 44300	2007-11-24 07:07:01 +00:00
Chris Lattner	ab98c41337	add a note llvm-svn: 44299	2007-11-24 06:13:33 +00:00
Dale Johannesen	763e110a9f	Fix .eh table linkage issues on Darwin. Some EH support for Darwin PPC, but it's not fully working yet. llvm-svn: 44258	2007-11-20 23:24:42 +00:00
Nate Begeman	d4d45c268c	Add support for vectors to int <-> float casts. llvm-svn: 44204	2007-11-17 03:58:34 +00:00
Anton Korobeynikov	91460e43f1	Implement codegen for flt_rounds on x86 llvm-svn: 44183	2007-11-16 01:31:51 +00:00
Evan Cheng	0cbe920d7c	Oops. Debugging code shouldn't have been checked in. llvm-svn: 44128	2007-11-14 19:08:32 +00:00
Anton Korobeynikov	2c6387803e	Fix PIC jump table codegen on x86-32/linux. In fact, such thing should be applied to all targets uses GOT-relative offsets for PIC (Alpha?) llvm-svn: 44108	2007-11-14 09:18:41 +00:00
Duncan Sands	e2287ed552	Eliminate the recently introduced CCAssignToStackABISizeAlign in favour of teaching CCAssignToStack that size 0 and/or align 0 means to use the ABI values. This seems a neater solution. It is safe since no legal value type has size 0. llvm-svn: 44107	2007-11-14 08:29:13 +00:00
Evan Cheng	7f02cfa599	Clean up sub-register implementation by moving subReg information back to MachineOperand auxInfo. Previous clunky implementation uses an external map to track sub-register uses. That works because register allocator uses a new virtual register for each spilled use. With interval splitting (coming soon), we may have multiple uses of the same register some of which are of using different sub-registers from others. It's too fragile to constantly update the information. llvm-svn: 44104	2007-11-14 07:59:08 +00:00
Dale Johannesen	7904708369	Revert previous; these files aren't ready to go in yet. llvm-svn: 44057	2007-11-13 19:16:02 +00:00
Dale Johannesen	7a7085f6d3	Add parameter to getDwarfRegNum to permit targets to use different mappings for EH and debug info; no functional change yet. Fix warning in X86CodeEmitter. llvm-svn: 44056	2007-11-13 19:13:01 +00:00
Evan Cheng	c891ae92dc	Fix x86-64 jit: remove reliance on Dwarf numbers. llvm-svn: 44048	2007-11-13 17:54:34 +00:00
Bill Wendling	77b13af9a6	Unifacalize the CALLSEQ{START,END} stuff. llvm-svn: 44045	2007-11-13 09:19:02 +00:00
Bill Wendling	f359fed9f9	Unify CALLSEQ_{START,END}. They take 4 parameters: the chain, two stack adjustment fields, and an optional flag. If there is a "dynamic_stackalloc" in the code, make sure that it's bracketed by CALLSEQ_START and CALLSEQ_END. If not, then there is the potential for the stack to be changed while the stack's being used by another instruction (like a call). This can only result in tears... llvm-svn: 44037	2007-11-13 00:44:25 +00:00
Owen Anderson	933b5b7e62	Add a flag for indirect branch instructions. Target maintainers: please check that the instructions for your target are correctly marked. llvm-svn: 44012	2007-11-12 07:39:39 +00:00
Anton Korobeynikov	4edfea438a	Use TableGen to emit information for dwarf register numbers. This makes DwarfRegNum to accept list of numbers instead. Added three different "flavours", but only slightly tested on x86-32/linux. Please check another subtargets if possible, llvm-svn: 43997	2007-11-11 19:50:10 +00:00
Dale Johannesen	b988e7e8cd	Add CCAssignToStackABISizeAlign for convenience in dealing with types whose size & alignment are different on different subtargets. Use it for x86 f80. llvm-svn: 43988	2007-11-10 22:07:15 +00:00
Arnold Schwaighofer	d2c16ff905	Update tailcall code to include inline attribute operand for memcpy. llvm-svn: 43978	2007-11-10 10:48:01 +00:00
Evan Cheng	fb13fd6f93	Unbreak x86-64 jumptable. llvm-svn: 43955	2007-11-09 19:11:23 +00:00
Dale Johannesen	dfb85c7831	Revert previous rewrite per chris's comments. llvm-svn: 43950	2007-11-09 18:07:11 +00:00
Evan Cheng	797d56ff17	Much improved pic jumptable codegen: Then: call "L1$pb" "L1$pb": popl %eax ... LBB1_1: # entry imull $4, %ecx, %ecx leal LJTI1_0-"L1$pb"(%eax), %edx addl LJTI1_0-"L1$pb"(%ecx,%eax), %edx jmpl %edx .align 2 .set L1_0_set_3,LBB1_3-LJTI1_0 .set L1_0_set_2,LBB1_2-LJTI1_0 .set L1_0_set_5,LBB1_5-LJTI1_0 .set L1_0_set_4,LBB1_4-LJTI1_0 LJTI1_0: .long L1_0_set_3 .long L1_0_set_2 Now: call "L1$pb" "L1$pb": popl %eax ... LBB1_1: # entry addl LJTI1_0-"L1$pb"(%eax,%ecx,4), %eax jmpl %eax .align 2 .set L1_0_set_3,LBB1_3-"L1$pb" .set L1_0_set_2,LBB1_2-"L1$pb" .set L1_0_set_5,LBB1_5-"L1$pb" .set L1_0_set_4,LBB1_4-"L1$pb" LJTI1_0: .long L1_0_set_3 .long L1_0_set_2 llvm-svn: 43924	2007-11-09 01:32:10 +00:00
Dale Johannesen	04fd82088e	Rewrite Dwarf number handling per review comments. llvm-svn: 43918	2007-11-09 00:47:10 +00:00
Dale Johannesen	1b9de4dd6f	Complete conditionalization of Dwarf reg numbers. Would somebody not on Darwin please make sure this doesn't break anything. Exception handling failures would be the most likely symptom. llvm-svn: 43844	2007-11-07 21:48:35 +00:00
Dale Johannesen	fbe69d2cd6	Interchange Dwarf numbers of ESP and EBP on x86 Darwin. Much improvement in exception handling. llvm-svn: 43794	2007-11-07 00:25:05 +00:00
Rafael Espindola	fa0df55bdd	Move the LowerMEMCPY and LowerMEMCPYCall to a common place. Thanks for the suggestions Bill :-) llvm-svn: 43742	2007-11-05 23:12:20 +00:00
Evan Cheng	9337929aae	Use movups to spill / restore SSE registers on targets where stacks alignment is less than 16. This is a temporary solution until dynamic stack alignment is implemented. llvm-svn: 43703	2007-11-05 07:30:01 +00:00
Duncan Sands	283207a71c	Eliminate the remaining uses of getTypeSize. This should only effect x86 when using long double. Now 12/16 bytes are output for long double globals (the exact amount depends on the alignment). This brings globals in line with the rest of LLVM: the space reserved for an object is now always the ABI size. One tricky point is that only 10 bytes should be output for long double if it is a field in a packed struct, which is the reason for the additional argument to EmitGlobalConstant. llvm-svn: 43688	2007-11-05 00:04:43 +00:00
Chris Lattner	9329e780cd	Fix PR1761 by not printing (rip) suffix when in -static mode. Evan, please review this. llvm-svn: 43680	2007-11-04 19:23:28 +00:00
Chris Lattner	296160d443	Fix PR1763 by allowing the 'q' constraint to work with 64-bit regs on x86-64. llvm-svn: 43669	2007-11-04 06:51:12 +00:00
Evan Cheng	2b93a20b09	Unbreak tailcall opt. llvm-svn: 43646	2007-11-02 17:45:40 +00:00
Chris Lattner	389d430c49	add a note llvm-svn: 43642	2007-11-02 17:04:20 +00:00
Evan Cheng	e453ff4913	Missing a getNumOperands check. llvm-svn: 43630	2007-11-02 01:26:22 +00:00
Bill Wendling	b7cabbe295	Silence, accersed warning llvm-svn: 43609	2007-11-01 08:51:44 +00:00
Rafael Espindola	419b6d7ce4	Make ARM and X86 LowerMEMCPY identical by moving the isThumb check into getMaxInlineSizeThreshold and by restructuring the X86 version. New I just have to move this to a common place :-) llvm-svn: 43554	2007-10-31 14:39:58 +00:00
Rafael Espindola	063f177300	Make ARM an X86 memcpy expansion more similar to each other. Now both subtarget define getMaxInlineSizeThreshold and the expansion uses it. This should not change generated code. llvm-svn: 43552	2007-10-31 11:52:06 +00:00
Dale Johannesen	b066c1f216	Make i64=expand_vector_elt(v2i64) work in 32-bit mode. llvm-svn: 43535	2007-10-31 00:32:36 +00:00
Dale Johannesen	d50c8bcef6	Add missing SSE builtins: CVTPD2PI, CVTPS2PI, CVTTPD2PI, CVTTPS2PI, CVTPI2PD, CVTPI2PS. llvm-svn: 43523	2007-10-30 22:15:38 +00:00
Duncan Sands	b508c53c63	Fix for visibility warnings generated by gcc-4.2. llvm-svn: 43500	2007-10-30 13:14:37 +00:00
Dale Johannesen	6aa304e529	Add missing MMX PSUBQ. llvm-svn: 43488	2007-10-30 01:18:38 +00:00
Evan Cheng	e106e2f142	Enable more fold (sext (load x)) -> (sext (truncate (sextload x))) transformation. Previously, it's restricted by ensuring the number of load uses is one. Now the restriction is loosened up by allowing setcc uses to be "extended" (e.g. setcc x, c, eq -> setcc sext(x), sext(c), eq). llvm-svn: 43465	2007-10-29 19:58:20 +00:00
Evan Cheng	7b3f7feaea	Avoid doing something dumb like rewriting using a 64-bit iv in 32-bit mode. llvm-svn: 43446	2007-10-29 07:57:50 +00:00
Chris Lattner	909a54ccd4	add a note. llvm-svn: 43444	2007-10-29 06:19:48 +00:00
Chris Lattner	5e99fd8c0d	Add support for the x86-64 'q' regigster modifier, and add support for the b/h/w/k/q inline asm memory modifiers, which are just ignored. This fixes PR1748 and CodeGen/X86/2007-10-28-inlineasm-q-modifier.ll llvm-svn: 43430	2007-10-29 03:09:07 +00:00
Evan Cheng	c826ac533b	New entry. llvm-svn: 43420	2007-10-28 04:01:09 +00:00
Anton Korobeynikov	d07d6a411c	Fix off-by-one stack offset computations (dwarf information) for callee-saved registers in case, when FP pointer was eliminated. This should fixes misc. random EH-related crahses, when stuff is compiled with -fomit-frame-pointer. Thanks Duncan for nailing this bug! llvm-svn: 43381	2007-10-26 09:13:24 +00:00
Evan Cheng	7f3d02471d	Loosen up iv reuse to allow reuse of the same stride but a larger type when truncating from the larger type to smaller type is free. e.g. Turns this loop: LBB1_1: # entry.bb_crit_edge xorl %ecx, %ecx xorw %dx, %dx movw %dx, %si LBB1_2: # bb movl L_X$non_lazy_ptr, %edi movw %si, (%edi) movl L_Y$non_lazy_ptr, %edi movw %dx, (%edi) addw $4, %dx incw %si incl %ecx cmpl %eax, %ecx jne LBB1_2 # bb into LBB1_1: # entry.bb_crit_edge xorl %ecx, %ecx xorw %dx, %dx LBB1_2: # bb movl L_X$non_lazy_ptr, %esi movw %cx, (%esi) movl L_Y$non_lazy_ptr, %esi movw %dx, (%esi) addw $4, %dx incl %ecx cmpl %eax, %ecx jne LBB1_2 # bb llvm-svn: 43375	2007-10-26 01:56:11 +00:00
Dan Gohman	bf474959a3	Fix the folding of multiplication into addresses on x86, which was broken by the recent {U,S}MUL_LOHI changes. llvm-svn: 43230	2007-10-22 20:22:24 +00:00
Evan Cheng	c92446af1f	Fix an unfolding bug. llvm-svn: 43212	2007-10-22 03:03:20 +00:00
Dale Johannesen	8ee70112ea	Allow for copysign having f80 second argument. Fixes 5550319. llvm-svn: 43205	2007-10-21 01:07:44 +00:00
Evan Cheng	45e096c77e	Resolve unfold tables ambiguity. llvm-svn: 43194	2007-10-19 23:50:58 +00:00
Evan Cheng	35ff79370b	Local spiller optimization: Turn a store folding instruction into a load folding instruction. e.g. xorl %edi, %eax movl %eax, -32(%ebp) movl -36(%ebp), %eax orl %eax, -32(%ebp) => xorl %edi, %eax orl -36(%ebp), %eax mov %eax, -32(%ebp) This enables the unfolding optimization for a subsequent instruction which will also eliminate the newly introduced store instruction. llvm-svn: 43192	2007-10-19 21:23:22 +00:00
Rafael Espindola	846c19dd70	Add support for byval function whose argument is not 32 bit aligned. To do this it is necessary to add a "always inline" argument to the memcpy node. For completeness I have also added this node to memmove and memset. I have also added getMem* functions, because the extra argument makes it cumbersome to use getNode and because I get confused by it :-) llvm-svn: 43172	2007-10-19 10:41:11 +00:00
Evan Cheng	463e2ab0ac	- Added getOpcodeAfterMemoryUnfold(). It doesn't unfold an instruction, but only returns the opcode of the instruction post unfolding. - Fix some copy+paste bugs. llvm-svn: 43153	2007-10-18 22:40:57 +00:00
Evan Cheng	aa9a225699	Use SmallVectorImpl instead of SmallVector with hardcoded size in MRegister public interface. llvm-svn: 43150	2007-10-18 21:29:24 +00:00
Christopher Lamb	7f68cf0d57	Fix a typo llvm-svn: 43144	2007-10-18 19:28:55 +00:00
Chris Lattner	12d5da49d3	Change fp to sint legalization on x86-32 to do 2 x i32 loads instead of 1 x i64 loads. This doesn't change any functionality yet. llvm-svn: 43068	2007-10-17 06:17:29 +00:00
Chris Lattner	693cbeadff	fix some funny indentation, add comments. llvm-svn: 43066	2007-10-17 06:02:13 +00:00
Dale Johannesen	e5530a35d4	Check for invalid cc's in f80 select. llvm-svn: 43033	2007-10-16 18:09:08 +00:00
Arnold Schwaighofer	b3d58b98d0	Correction to tail call optimization code. The new return address was stored to the acutal stack slot before the parameters were lowered to their stack slot. This could cause arguments to be overwritten by the return address if the called function had less parameters than the caller function. The update should remove the last failing test case of llc-beta: SPASS. llvm-svn: 43027	2007-10-16 09:05:00 +00:00
Evan Cheng	7bcfd8f880	LowerFP_TO_SINT must not create a stack object if it's not needed. llvm-svn: 43004	2007-10-15 20:11:21 +00:00
Evan Cheng	4099f4f91a	Unbreak x86-64. llvm-svn: 42962	2007-10-14 10:09:39 +00:00
Evan Cheng	cdf3609130	Revert 42908 for now. llvm-svn: 42960	2007-10-14 05:57:21 +00:00
Duncan Sands	29af26f147	Clarify that fastcc has a problem with nested function trampolines, rather than with nested functions themselves. llvm-svn: 42955	2007-10-13 07:38:37 +00:00
Evan Cheng	7082dcf605	Change unfoldMemoryOperand(). User is now responsible for passing in the register used by the unfolded instructions. User can also specify whether to unfold the load, the store, or both. llvm-svn: 42946	2007-10-13 02:35:06 +00:00
Arnold Schwaighofer	e8d0bf2669	Correcting the corrections. Bad bad baaad emacs! llvm-svn: 42935	2007-10-12 21:53:12 +00:00
Arnold Schwaighofer	1f0da1fefb	Corrected many typing errors. And removed 'nest' parameter handling for fastcc from X86CallingConv.td. This means that nested functions are not supported for calling convention 'fastcc'. llvm-svn: 42934	2007-10-12 21:30:57 +00:00
Duncan Sands	a6286bd502	Due to the new tail call optimization, trampolines can no longer be created for fastcc functions. llvm-svn: 42925	2007-10-12 19:37:31 +00:00
Evan Cheng	409fa443fc	Update. llvm-svn: 42922	2007-10-12 18:22:55 +00:00
Dan Gohman	dc35bd79ca	Change the names used for internal labels to use the current function symbol name instead of a codegen-assigned function number. Thanks Evan! :-) llvm-svn: 42908	2007-10-12 14:53:36 +00:00
Dan Gohman	8d978da3b0	Mark vector ctpop, cttz, and ctlz as Expand on x86. llvm-svn: 42905	2007-10-12 14:09:42 +00:00
Evan Cheng	09c0fe0a7f	Fold load / store into MOV32to32_ and MOV16to16_. llvm-svn: 42895	2007-10-12 08:38:01 +00:00
Evan Cheng	f8c23f074b	Flag MOV32to32_ with EXTRACT_SUBREG. They should not be scheduled apart. llvm-svn: 42894	2007-10-12 07:55:53 +00:00
Dan Gohman	482732af9d	Set ISD::FPOW to Expand. llvm-svn: 42881	2007-10-11 23:21:31 +00:00
Dale Johannesen	62f65edc32	Add missing argument to PALIGNR llvm-svn: 42874	2007-10-11 20:58:37 +00:00
Arnold Schwaighofer	9ccea99165	Added tail call optimization to the x86 back end. It can be enabled by passing -tailcallopt to llc. The optimization is performed if the following conditions are satisfied: * caller/callee are fastcc * elf/pic is disabled OR elf/pic enabled + callee is in module + callee has visibility protected or hidden llvm-svn: 42870	2007-10-11 19:40:01 +00:00
Dan Gohman	e8c8ef5234	LowerIntegerDivOrRem no longer exists. llvm-svn: 42787	2007-10-09 15:45:13 +00:00
Dan Gohman	51554bf30e	Fix grammar in a comment. llvm-svn: 42786	2007-10-09 15:44:37 +00:00
Dan Gohman	6d28778bfd	This is done. llvm-svn: 42785	2007-10-09 15:42:21 +00:00
Evan Cheng	82bc90ac60	Under 64-bit mode use LEA64_32r instead of LEA64r to save a byte. llvm-svn: 42783	2007-10-09 07:14:53 +00:00
Evan Cheng	f5ec10b64c	Bug fix. X86 was emitting redundant setcc and test instructions before a conditional move. llvm-svn: 42774	2007-10-08 22:16:29 +00:00
Dan Gohman	a160361c85	Migrate X86 and ARM from using X86ISD::{,I}DIV and ARMISD::MULHILO{U,S} to use ISD::{S,U}DIVREM and ISD::{S,U}MUL_HIO. Move the lowering code associated with these operators into target-independent in LegalizeDAG.cpp and TargetLowering.cpp. llvm-svn: 42762	2007-10-08 18:33:35 +00:00
Evan Cheng	18109c88c3	Allow x86 compare to be commutable by default. llvm-svn: 42761	2007-10-08 18:27:46 +00:00
Chris Lattner	b20757d578	disable this entirely: it is causing use of invalidated iterators and infinite looping. llvm-svn: 42739	2007-10-07 22:00:31 +00:00
Chris Lattner	8dd66ab3b2	Fix many regressions on x86 by avoiding dereferencing the end iterator. llvm-svn: 42738	2007-10-07 21:53:12 +00:00
Anton Korobeynikov	67ac2de8bf	Oops, I really wanted to commit this part also :) llvm-svn: 42700	2007-10-06 16:39:43 +00:00
Anton Korobeynikov	c59496f737	Move merge code into new helper function. llvm-svn: 42699	2007-10-06 16:17:49 +00:00
Evan Cheng	f4b5d491df	Added DAG xforms. e.g. (vextract (v4f32 s2v (f32 load $addr)), 0) -> (f32 load $addr) (vextract (v4i32 bc (v4f32 s2v (f32 load $addr))), 0) -> (i32 load $addr) Remove x86 specific patterns. llvm-svn: 42677	2007-10-06 02:46:29 +00:00
Evan Cheng	1151ffde70	Commute x86 cmove instructions by swapping the operands and change the condition to its inverse. Testing this as llcbeta llvm-svn: 42661	2007-10-05 23:13:21 +00:00
Evan Cheng	42a13757de	This is done. llvm-svn: 42656	2007-10-05 22:34:59 +00:00
Evan Cheng	484cab7a2f	Enable convertToThreeAddress for X86 by default. llvm-svn: 42655	2007-10-05 22:31:10 +00:00
Evan Cheng	d3ccf00870	INC64_32r -> LEA64_32r is better than INC64_32r -> LEA32r, but it still can cause performance degradation. llvm-svn: 42653	2007-10-05 21:55:32 +00:00
Evan Cheng	fa2c828687	In 64-bit mode, avoid using leal with 32-bit 32-bit address size, e.g. leal 1(%ecx), %edi, which requires 67H prefix. llvm-svn: 42647	2007-10-05 20:34:26 +00:00
Evan Cheng	aac0f8e351	Add support to convert more 64-bit instructions to 3-address instructions. llvm-svn: 42642	2007-10-05 18:20:36 +00:00
Evan Cheng	97eba74a52	ADC and SBB uses EFLAGS. llvm-svn: 42640	2007-10-05 17:59:57 +00:00
Dan Gohman	43c29dce18	Change a few more spaces to tabs in assembly output. llvm-svn: 42638	2007-10-05 15:58:41 +00:00
Dan Gohman	b074f23dff	Change a space to a tab in the assembly output of a .globl directive for consistency. llvm-svn: 42637	2007-10-05 15:54:58 +00:00
Evan Cheng	a8a9c15e30	Testing convertToThreeeAddress as X86 llcbeta. llvm-svn: 42630	2007-10-05 08:04:01 +00:00
Evan Cheng	a851e2b92e	Added storeRegToAddr, loadRegFromAddr, and unfoldMemoryOperand's. llvm-svn: 42624	2007-10-05 01:34:55 +00:00
Evan Cheng	6912b50958	Not needed any more. llvm-svn: 42623	2007-10-05 01:34:14 +00:00
Chris Lattner	1f2b5f0e13	add a note. llvm-svn: 42607	2007-10-04 15:47:27 +00:00
Dan Gohman	c731c97fac	Use empty() member functions when that's what's being tested for instead of comparing begin() and end(). llvm-svn: 42585	2007-10-03 19:26:29 +00:00
Chris Lattner	4bdb84fe53	add a note llvm-svn: 42579	2007-10-03 17:10:03 +00:00
Chris Lattner	21ba176c4b	Bill's example is still not enough to repro this, but it has other issues that seem significant as well. llvm-svn: 42564	2007-10-03 03:40:24 +00:00
Bill Wendling	3efc0758ae	Another micro-opt. llvm-svn: 42554	2007-10-02 21:49:31 +00:00
Bill Wendling	f214ff8701	Another missed optimization with LICM. llvm-svn: 42552	2007-10-02 21:43:06 +00:00
Bill Wendling	855011e5c6	Small label changes. llvm-svn: 42549	2007-10-02 21:02:53 +00:00
Bill Wendling	4eb7ca4b4c	Now with source code. llvm-svn: 42548	2007-10-02 21:01:16 +00:00
Bill Wendling	96ed3bb2d4	Now with LL code! llvm-svn: 42547	2007-10-02 20:54:32 +00:00
Bill Wendling	9c4d61b523	Another missed optimization. llvm-svn: 42546	2007-10-02 20:42:59 +00:00
Bill Wendling	88ea107fdb	Micro-optimization -- missed LICM opportunity. llvm-svn: 42542	2007-10-02 19:55:05 +00:00
Evan Cheng	1f79ba6fe6	Refactor code to add load / store folded instructions -> register only instructions reverse map. llvm-svn: 42509	2007-10-01 23:44:33 +00:00
Evan Cheng	a1b7e95039	Typo. X86comi doesn't read / write chain's. llvm-svn: 42492	2007-10-01 18:12:48 +00:00
Gordon Henriksen	0b7cf862bc	AsmPrinters overriding getAnalysisUsage should call super. And not super's super, either. llvm-svn: 42482	2007-09-30 13:39:29 +00:00

... 4 5 6 7 8 ...

3274 Commits