llvm-project

Commit Graph

Author	SHA1	Message	Date
Craig Topper	04b4e83cf7	Fix bad comment. No functional change. llvm-svn: 164000	2012-09-16 16:48:25 +00:00
Craig Topper	8dcaf4998c	Add 'virtual' keywoards to output file for overridden functions. llvm-svn: 163999	2012-09-16 16:35:22 +00:00
Nadav Rotem	ae6809b19a	Fix the testcase to work on all platforms. llvm-svn: 163997	2012-09-16 07:58:47 +00:00
Craig Topper	f4214c0f85	Add explicit virtual keywords for methods that override base class. llvm-svn: 163996	2012-09-16 07:39:55 +00:00
Nadav Rotem	37521aa89c	The PMOVZXWD family of functions had patterns extends narrow vector types to wide vector types. It had patterns for zext-loading and extending. This commit adds patterns for loading a wide type, performing a bitcast, and extending. This is an odd pattern, but it is commonly used when writing code with intrinsics. rdar://11897677 llvm-svn: 163995	2012-09-16 07:39:07 +00:00
Andrew Trick	04b38b0d97	Guard fields by NDEBUG until they get used in the release build. llvm-svn: 163993	2012-09-16 05:55:04 +00:00
Craig Topper	b5319a0acc	Tidy up formatting of some elses on a separate line from preceding bracing. No functional change. llvm-svn: 163992	2012-09-16 03:00:03 +00:00
Jakob Stoklund Olesen	17e2185543	Add alternative coalescing algorithm under a flag. The live range of an SSA value forms a sub-tree of the dominator tree. That means the live ranges of two values overlap if and only if the def of one value lies within the live range of the other. This can be used to simplify the interference checking a bit: Visit each def in the two registers about to be joined. Check for interference against the value that is live in the other register at the def point only. It is not necessary to scan the set of overlapping live ranges, this interference check can be done while computing the value mapping required for the final live range join. The new algorithm is prepared to handle more complicated conflict resolution - We can allow overlapping live ranges with different values as long as the differing lanes are undef or unused in the other register. The implementation in this patch doesn't do that yet, it creates code that is nearly identical to the old algorithm's, except: - The new stripCopies() function sees through multiple copies while the old RegistersDefinedFromSameValue() only can handle one. - There are a few rare cases where the new algorithm can erase an IMPLICIT_DEF instuction that RegistersDefinedFromSameValue() couldn't handle. llvm-svn: 163991	2012-09-16 02:15:36 +00:00
Jakob Stoklund Olesen	932e3d7e2d	Fix problem when using LiveRangeQuery with block entries. A value that is live in to a basic block should be returned by valueIn() in LiveRangeQuery(getMBBStartIdx(MBB)), unless it is a PHI-def which should be returned by valueDefined() instead. Current code isn't using this functionality. Future code will. llvm-svn: 163990	2012-09-16 02:15:33 +00:00
Craig Topper	1c0fcdab2e	Tidy up trailing whitespace. llvm-svn: 163988	2012-09-16 01:20:35 +00:00
Craig Topper	4ee5a2d53e	Remove unneeded header. llvm-svn: 163987	2012-09-16 01:18:51 +00:00
Dmitri Gribenko	8d30240939	Fix Doxygen issues: wrap code examples in \code and use \p to refer to parameters. llvm-svn: 163984	2012-09-15 20:22:05 +00:00
Craig Topper	bc40d7e023	Fix includes of llvm files that used angle brackets. llvm-svn: 163979	2012-09-15 18:45:38 +00:00
Craig Topper	53d08e4f0c	Fix a couple include directives that used angle brackets for llvm files. llvm-svn: 163978	2012-09-15 18:41:37 +00:00
Craig Topper	a60c0f1163	Use LLVM_DELETED_FUNCTION in place of 'DO NOT IMPLEMENT' comments. llvm-svn: 163974	2012-09-15 17:09:36 +00:00
Craig Topper	2ed23ce767	Remove unused private fields to silence -Wunused-private-field. llvm-svn: 163973	2012-09-15 17:08:51 +00:00
Jakob Stoklund Olesen	b7d27a3dd7	Don't depend on kill flags in removeCopyByCommutingDef(). Kill flags are removed more and more aggressively during the register allocation passes, it is better to get information from LiveIntervals. llvm-svn: 163972	2012-09-15 16:32:11 +00:00
Jakob Stoklund Olesen	d98e6c9fb5	Make LiveRangeQuery work for PHIDefs as well. If a PHI value happens to be live out from the layout predecessor of its def block, the def slot index will be in the middle of the segment: %vreg11 = [192r,240B:0)[352r,416B:2)[416B,496r:1) 0@192r 1@480B-phi %2@352r A LiveRangeQuery for 480 should return NULL from valueIn() since the PHI value is defined at the block entry, not live in to the block. No test case, future code depends on this functionality. llvm-svn: 163971	2012-09-15 16:29:49 +00:00
Craig Topper	2e6644c260	Use LLVM_DELETED_FUNCTION in place of 'DO NOT IMPLEMENT' comments. llvm-svn: 163970	2012-09-15 16:23:52 +00:00
Craig Topper	da386573c7	Use LLVM_DELETED_FUNCTION in place of 'DO NOT IMPLEMENT' comments. llvm-svn: 163969	2012-09-15 16:22:27 +00:00
Benjamin Kramer	ed11e35e57	Disable new sroa now that all buildbots have tested it. What we have so far: - Some clang test failures (these were known already) - Perf results are mixed, some big regressions http://llvm.org/perf/db_default/v4/nts/3844 http://llvm.org/perf/db_default/v4/nts/3845 bullet suffers a lot. matmul is interesting: slower scalar code, faster with -vectorize. - Some dragonegg selfhost bots crash in SROA during selfhost now http://lab.llvm.org:8011/builders/dragonegg-x86_64-linux-gcc-4.6-self-host-checks/builds/1632 http://lab.llvm.org:8011/builders/dragonegg-x86_64-linux-gcc-4.5-self-host/builds/1891 llvm-svn: 163968	2012-09-15 15:11:10 +00:00
Benjamin Kramer	ece434252c	X86: Emitting x87 fsin/fcos for sinf/cosf is not safe without unsafe fp math. This was only an issue if sse is disabled. llvm-svn: 163967	2012-09-15 12:44:27 +00:00
Chandler Carruth	70b44c5ccf	Port the SSAUpdater-based promotion logic from the old SROA pass to the new one, and add support for running the new pass in that mode and in that slot of the pass manager. With this the new pass can completely replace the old one within the pipeline. The strategy for enabling or disabling the SSAUpdater logic is to do it by making the requirement of the domtree analysis optional. By default, it is required and we get the standard mem2reg approach. This is usually the desired strategy when run in stand-alone situations. Within the CGSCC pass manager, we disable requiring of the domtree analysis and consequentially trigger fallback to the SSAUpdater promotion. In theory this would allow the pass to re-use a domtree if one happened to be available even when run in a mode that doesn't require it. In practice, it lets us have a single pass rather than two which was simpler for me to wrap my head around. There is a hidden flag to force the use of the SSAUpdater code path for the purpose of testing. The primary testing strategy is just to run the existing tests through that path. One notable difference is that it has custom code to handle lifetime markers, and one of the tests has been enhanced to exercise that code. This has survived a bootstrap and the test suite without serious correctness issues, however my run of the test suite produced very alarming performance numbers. I don't entirely understand or trust them though, so more investigation is on-going. To aid my understanding of the performance impact of the new SROA now that it runs throughout the optimization pipeline, I'm enabling it by default in this commit, and will disable it again once the LNT bots have picked up one iteration with it. I want to get those bots (which are much more stable) to evaluate the impact of the change before I jump to any conclusions. NOTE: Several Clang tests will fail because they run -O3 and check the result's order of output. They'll go back to passing once I disable it again. llvm-svn: 163965	2012-09-15 11:43:14 +00:00
Akira Hatanaka	3e7ba76157	Remove aligned/unaligned load/store fragments defined in MipsInstrInfo.td and use load/store fragments defined in TargetSelectionDAG.td in place of them. Unaligned loads/stores are either expanded or lowered to target-specific nodes, so instruction selection should see only aligned load/store nodes. No changes in functionality. llvm-svn: 163960	2012-09-15 01:52:08 +00:00
Craig Topper	f8f0a23ce7	Revert r163878 as it breaks on targets with alternate register names. Such targets do not exist in the main tree so this was not noticed. llvm-svn: 163959	2012-09-15 01:22:42 +00:00
Akira Hatanaka	189d0adde9	Handled unaligned load/stores properly in Mips16 Patch by Reed Kotler. llvm-svn: 163956	2012-09-15 01:02:03 +00:00
Manman Ren	bfb9d435e4	PGO: preserve branch-weight metadata when simplifying two branches with a common destination. Updated previous implementation to fix a case not covered: // PBI: br i1 %x, TrueDest, BB // BI: br i1 %y, TrueDest, FalseDest The other case was handled correctly. // PBI: br i1 %x, BB, FalseDest // BI: br i1 %y, TrueDest, FalseDest Also tried to use 64-bit arithmetic instead of APInt with scale to simplify the computation. Let me know if you have other opinions about this. llvm-svn: 163954	2012-09-15 00:39:57 +00:00
Andrew Trick	1e46d48814	TableGen subtarget parser. Handle new machine model. Collect processor resources from the subtarget defs. llvm-svn: 163953	2012-09-15 00:20:02 +00:00
Andrew Trick	33401e8469	TableGen subtarget parser. Handle new machine model. Infer SchedClasses from variants defined by the target or subtarget. llvm-svn: 163952	2012-09-15 00:19:59 +00:00
Andrew Trick	766864963b	TableGen subtarget parser. Handle new machine model. Collect SchedClasses and SchedRW types from the subtarget defs. llvm-svn: 163951	2012-09-15 00:19:57 +00:00
Daniel Dunbar	b93a2ceba5	cmake: Fix file path. llvm-svn: 163950	2012-09-14 23:36:56 +00:00
Daniel Dunbar	9affb245a5	formatted_raw_ostream: Fix a serious bug in tell(). - The current_pos function is supposed to return all the written bytes, not the current position of the underlying stream. - This caused tell() to be broken whenever the underlying stream had buffered content. llvm-svn: 163948	2012-09-14 23:15:56 +00:00
Bill Wendling	25cc99fa44	Some small reorganization to get read for Attributes overhaul. llvm-svn: 163947	2012-09-14 23:05:52 +00:00
Bill Wendling	8d26bc38f5	Remove comment. llvm-svn: 163945	2012-09-14 22:35:49 +00:00
David Blaikie	21e27ce264	Fix up erroneous alignas usage while making this portable to GCC 4.7 Review by Chandler Carruth. llvm-svn: 163944	2012-09-14 22:26:11 +00:00
Manman Ren	8691e5220b	PGO: preserve branch-weight metadata when simplifying a switch with a single case to a conditional branch and when removing dead cases. llvm-svn: 163942	2012-09-14 21:53:06 +00:00
Evan Cheng	71be12b35b	Stylistic and 80-col fixes llvm-svn: 163940	2012-09-14 21:25:34 +00:00
Andrew Trick	46846fcdc8	comment typo llvm-svn: 163935	2012-09-14 20:27:25 +00:00
Andrew Trick	d2a19da1b8	TargetSchedModel interface. To be implemented... llvm-svn: 163934	2012-09-14 20:26:46 +00:00
Andrew Trick	ac36af470c	Define MC data tables for the new scheduling machine model. llvm-svn: 163933	2012-09-14 20:26:41 +00:00
Andrew Trick	39cf40a2ae	whitespace llvm-svn: 163932	2012-09-14 20:26:39 +00:00
Alex Rosenberg	af2808cb72	Review feedback from Duncan Sands. Alphabetize includes and simplify lit config. llvm-svn: 163928	2012-09-14 19:19:57 +00:00
Manman Ren	5e5049d9a6	Try to fix the bots by detecting inconsistant branch-weight metadata. llvm-svn: 163926	2012-09-14 19:05:19 +00:00
Andrew Trick	2ac6f7d6f6	Implement getNumLDMAddresses and expose through ARMBaseInstrInfo. llvm-svn: 163922	2012-09-14 18:48:46 +00:00
Andrew Trick	985dc0dd64	Cortex-A9 instruction-level scheduling machine model. This models the A9 processor at the level of instruction operands, as opposed to the itinerary, which models each operation at the level of pipeline stages. The two primary motivations are: 1) Allow MachineScheduler to model A9 as an out-of-order processor. It can now distinguish between hazards that force interlocking vs. buffered resources. 2) Reduce long-term maintenance by allowing the itinerary and target hooks to eventually be removed. Note that almost all of the complexity in the new model exists to model instruction variants, which the itinerary cannot handle. Instead the scheduler previously relied on processor-specific target hooks which are incomplete and buggy. llvm-svn: 163921	2012-09-14 18:31:58 +00:00
Manman Ren	d81b8e88e3	PGO: preserve branch-weight metadata when merging two switches where the default target of the first switch is not the basic block the second switch is in (PredDefault != BB). llvm-svn: 163916	2012-09-14 17:29:56 +00:00
Andrew Trick	a2733e9549	misched: add a hook for custom DAG postprocessing. llvm-svn: 163915	2012-09-14 17:22:42 +00:00
Micah Villmow	e2da82afe3	Add in comments that explain what the indexing and the size of the arrays is about. llvm-svn: 163904	2012-09-14 15:36:50 +00:00
Sergei Larin	2db64a7031	DAG post-process for Hexagon MI scheduler This patch introduces a possibility for Hexagon MI scheduler to perform some target specific post- processing on the scheduling DAG prior to scheduling. llvm-svn: 163903	2012-09-14 15:07:59 +00:00
Dmitri Gribenko	5485acd440	Fix Doxygen issues: * wrap code blocks in \code ... \endcode; * refer to parameter names in paragraphs correctly (\arg is not what most people want -- it starts a new paragraph); * use \param instead of \arg to document parameters in order to be consistent with the rest of the codebase. llvm-svn: 163902	2012-09-14 14:57:36 +00:00
Benjamin Kramer	4622cd7edd	SROA: Silence unused variable warnings in Release builds. The NDEBUG hack is ugly, but I see no better solution. llvm-svn: 163900	2012-09-14 13:08:09 +00:00
Benjamin Kramer	61f6708eee	Remove redundant private field. clang warned about this being unused in Release builds. llvm-svn: 163899	2012-09-14 12:19:58 +00:00
Chandler Carruth	054a40a4ff	Rework the computation of a sub-structure natural type. There were pointless checks in here, bad asserts, and just confusing code. I've also added a bit more to the comment to clarify what this function is really trying to do as it was not obvious to Duncan when studying it. Thanks to Duncan for helping me dig through the issue. No real functionality changed here in practical cases, and certainly no test case. This is just cleanup spotted by inspection. llvm-svn: 163897	2012-09-14 11:08:31 +00:00
Chandler Carruth	0cc59250d5	Rely on the recursive check for pointer types rather than adding an explicit check before recursing. A simplification requested by Duncan during review. llvm-svn: 163896	2012-09-14 10:30:44 +00:00
Chandler Carruth	cabd96cbaa	Be a bit more aggressive in bailing out of this routine. Spotted by inspection by Duncan during review. My suspicion is that we would still have returned 0 anyways in this case, but doing it sooner is better. llvm-svn: 163895	2012-09-14 10:30:42 +00:00
Chandler Carruth	dd3cea898f	Add some comments clarifying that the GEP analysis for vector GEPs is deeply suspicious and likely to go away eventually. Also fix a bogus comment about one of the checks in the vector GEP analysis. Based on review from Duncan. llvm-svn: 163894	2012-09-14 10:30:40 +00:00
Chandler Carruth	19450da9e6	Move an instance variable to a local variable based on review by Duncan. Originally I had anticipated needing to thread this through more bits of the SROA pass itself, but that ended up not happening. In the end, this is a much simpler way to manange the variable. llvm-svn: 163893	2012-09-14 10:26:38 +00:00
Chandler Carruth	4b40e008bd	Add a comment about debug intrinsics that I really don't want to forget from Duncan's review as a FIXME. llvm-svn: 163892	2012-09-14 10:26:36 +00:00
Chandler Carruth	b0de6ddbe0	Add two asserts that Duncan thought would help ensure things don't rot unexpectedly in the future. More fixes from his code review. llvm-svn: 163891	2012-09-14 10:26:34 +00:00
Chandler Carruth	6ba9824c2b	Actually keep the flag default-off for now. =/ That's what I get for being busy testing this... llvm-svn: 163890	2012-09-14 10:18:54 +00:00
Chandler Carruth	796de48459	Remove some dead, commented out code Duncan spotted in review. llvm-svn: 163889	2012-09-14 10:18:53 +00:00
Chandler Carruth	25fb23d687	Wrap the dumping and printing routines in NDEBUG and LLVM_ENABLE_DUMP macros. llvm-svn: 163888	2012-09-14 10:18:51 +00:00
Chandler Carruth	93a21e7aaf	Lots of comment fixes and cleanups from Duncan's review. llvm-svn: 163887	2012-09-14 10:18:49 +00:00
NAKAMURA Takumi	4bbca0bb6c	SROA.cpp: Unbreak gcc, sorry! llvm-svn: 163886	2012-09-14 10:06:10 +00:00
NAKAMURA Takumi	f4619d169d	SROA.cpp: Appease msvc. LLVM_ATTRIBUTE(s) should come front of "const". llvm-svn: 163885	2012-09-14 09:55:22 +00:00
Chandler Carruth	9a447db9fc	Speculative change to try to fix older GCC versions that can't handle the injected class name of a dependent base class here. llvm-svn: 163884	2012-09-14 09:30:33 +00:00
Chandler Carruth	1b398ae0ae	Introduce a new SROA implementation. This is essentially a ground up re-think of the SROA pass in LLVM. It was initially inspired by a few problems with the existing pass: - It is subject to the bane of my existence in optimizations: arbitrary thresholds. - It is overly conservative about which constructs can be split and promoted. - The vector value replacement aspect is separated from the splitting logic, missing many opportunities where splitting and vector value formation can work together. - The splitting is entirely based around the underlying type of the alloca, despite this type often having little to do with the reality of how that memory is used. This is especially prevelant with unions and base classes where we tail-pack derived members. - When splitting fails (often due to the thresholds), the vector value replacement (again because it is separate) can kick in for preposterous cases where we simply should have split the value. This results in forming i1024 and i2048 integer "bit vectors" that tremendously slow down subsequnet IR optimizations (due to large APInts) and impede the backend's lowering. The new design takes an approach that fundamentally is not susceptible to many of these problems. It is the result of a discusison between myself and Duncan Sands over IRC about how to premptively avoid these types of problems and how to do SROA in a more principled way. Since then, it has evolved and grown, but this remains an important aspect: it fixes real world problems with the SROA process today. First, the transform of SROA actually has little to do with replacement. It has more to do with splitting. The goal is to take an aggregate alloca and form a composition of scalar allocas which can replace it and will be most suitable to the eventual replacement by scalar SSA values. The actual replacement is performed by mem2reg (and in the future SSAUpdater). The splitting is divided into four phases. The first phase is an analysis of the uses of the alloca. This phase recursively walks uses, building up a dense datastructure representing the ranges of the alloca's memory actually used and checking for uses which inhibit any aspects of the transform such as the escape of a pointer. Once we have a mapping of the ranges of the alloca used by individual operations, we compute a partitioning of the used ranges. Some uses are inherently splittable (such as memcpy and memset), while scalar uses are not splittable. The goal is to build a partitioning that has the minimum number of splits while placing each unsplittable use in its own partition. Overlapping unsplittable uses belong to the same partition. This is the target split of the aggregate alloca, and it maximizes the number of scalar accesses which become accesses to their own alloca and candidates for promotion. Third, we re-walk the uses of the alloca and assign each specific memory access to all the partitions touched so that we have dense use-lists for each partition. Finally, we build a new, smaller alloca for each partition and rewrite each use of that partition to use the new alloca. During this phase the pass will also work very hard to transform uses of an alloca into a form suitable for promotion, including forming vector operations, speculating loads throguh PHI nodes and selects, etc. After splitting is complete, each newly refined alloca that is a candidate for promotion to a scalar SSA value is run through mem2reg. There are lots of reasonably detailed comments in the source code about the design and algorithms, and I'm going to be trying to improve them in subsequent commits to ensure this is well documented, as the new pass is in many ways more complex than the old one. Some of this is still a WIP, but the current state is reasonbly stable. It has passed bootstrap, the nightly test suite, and Duncan has run it successfully through the ACATS and DragonEgg test suites. That said, it remains behind a default-off flag until the last few pieces are in place, and full testing can be done. Specific areas I'm looking at next: - Improved comments and some code cleanup from reviews. - SSAUpdater and enabling this pass inside the CGSCC pass manager. - Some datastructure tuning and compile-time measurements. - More aggressive FCA splitting and vector formation. Many thanks to Duncan Sands for the thorough final review, as well as Benjamin Kramer for lots of review during the process of writing this pass, and Daniel Berlin for reviewing the data structures and algorithms and general theory of the pass. Also, several other people on IRC, over lunch tables, etc for lots of feedback and advice. llvm-svn: 163883	2012-09-14 09:22:59 +00:00
Duncan Sands	291d47efdf	Remove silly dead store. Patch by Ettl Martin. llvm-svn: 163882	2012-09-14 09:00:11 +00:00
Craig Topper	06cec4ceac	Allow the second opcode info table to be 8, 16, or 32-bits as needed to represent additional fragments. This recovers some space on ATT X86 syntax and PowerPC which only need 40-bits instead of 48-bits. This also increases ARM to 64-bits to fully encode all of its operands. llvm-svn: 163880	2012-09-14 08:33:11 +00:00
Craig Topper	805e39beab	Reduce size of register name index tables by using uint16_t for all in tree targets. If more than 16-bits are needed for any out of tree targets, code will detect and use uint32_t instead. llvm-svn: 163878	2012-09-14 06:37:49 +00:00
Andrew Trick	b8804efc8e	misched: Generic tablegen classes for the new machine model. This is mostly documentation for the new machine model. It is designed to be flexible, easy to incrementally refine for a subtarget, and provide all the information that MachineScheduler will need. If all goes well, I will follow up with an example of the new model in use for ARM. llvm-svn: 163877	2012-09-14 06:18:55 +00:00
Andrew Trick	be5a54a363	comment llvm-svn: 163876	2012-09-14 06:18:52 +00:00
Andrew Trick	78276ea615	comment llvm-svn: 163875	2012-09-14 06:18:50 +00:00
Akira Hatanaka	0fbaec2246	mips16 fixes. 1. Add MoveR3216 2. Correct spelling for Move32R16 Patch by Reed Kotler. llvm-svn: 163869	2012-09-14 03:21:56 +00:00
Galina Kistanova	8201936f60	Patch by Sean Silva! The patch converts the "How to add a builder" document over to reStructuredText.. llvm-svn: 163860	2012-09-13 23:51:08 +00:00
Eric Christopher	b83dba2b84	Fix both the test for zero and what we do if we have a zero for umulo legalization. Fixes PR13839 llvm-svn: 163856	2012-09-13 23:24:02 +00:00
Eric Christopher	3bc248176c	Reformat, remove a couple unused variables and move some variables closer to where they're needed. llvm-svn: 163855	2012-09-13 23:23:58 +00:00
Jim Grosbach	b7b750d480	Assembler: Darwin variables defined via .set are no-dead-strip. For gas compatibility. rdar://12219394 llvm-svn: 163854	2012-09-13 23:11:31 +00:00
Jim Grosbach	d96ef194d9	MachO: Correctly mark symbol-difference variables as N_ABS. .set a, b - c + CONSTANT d = b - c + CONSTANT Both 'a' and 'd' should be marked as absolute symbols (N_ABS). rdar://12219394 llvm-svn: 163853	2012-09-13 23:11:25 +00:00
Dan Gohman	3f553c21eb	Handle the new !tbaa.struct metadata tags when converting a memcpy into scalar loads and stores. llvm-svn: 163844	2012-09-13 21:51:01 +00:00
Jim Grosbach	6d61397c73	Better const handling for RuntimeDyld and MCJIT. mapSectionAddress() wasn't consistent. llvm-svn: 163843	2012-09-13 21:50:06 +00:00
Richard Smith	d5b247b886	Fix some code which is invalid in C++11: an expression of enumeration type can't be used as a non-type template argument of type bool. llvm-svn: 163840	2012-09-13 21:18:18 +00:00
Manman Ren	4d9ae56a45	AsmWriterEmitter: OpInfo2 should be unsigned 16-bit. Fix an issue in r163814. llvm-svn: 163837	2012-09-13 20:47:48 +00:00
Michael Liao	8b48bf27b0	Fix comment llvm-svn: 163835	2012-09-13 20:30:16 +00:00
Dmitri Gribenko	737fc6c3c3	Fix documentation: parameter being documented was removed in r98220. llvm-svn: 163834	2012-09-13 20:28:31 +00:00
Michael Liao	137f8aedea	Add wider vector/integer support for PR12312 - Enhance the fix to PR12312 to support wider integer, such as 256-bit integer. If more than 1 fully evaluated vectors are found, POR them first followed by the final PTEST. llvm-svn: 163832	2012-09-13 20:24:54 +00:00
Michael Liao	460fc46e0f	Enhance type legalization on bitcast from vector to integer - Find a legal vector type before casting and extracting element from it. - As the new vector type may have more than 2 elements, build the final hi/lo pair by BFS pairing them from bottom to top. llvm-svn: 163830	2012-09-13 19:58:21 +00:00
Jakob Stoklund Olesen	32a56fa3ba	Fix test case to avoid PIC magic. llvm-svn: 163827	2012-09-13 19:47:45 +00:00
Jakob Stoklund Olesen	3cf3ffce24	Fix the TCRETURNmi64 bug differently. Add a PatFrag to match X86tcret using 6 fixed registers or less. This avoids folding loads into TCRETURNmi64 using 7 or more volatile registers. <rdar://problem/12282281> llvm-svn: 163819	2012-09-13 18:31:27 +00:00
Dan Gohman	d0080c45f9	Extract code for reducing a type to a single value type into a helper function. llvm-svn: 163817	2012-09-13 18:19:06 +00:00
Dan Gohman	3effe81bf7	Define an official slot for the new !tbaa.struct metadata tag. llvm-svn: 163815	2012-09-13 17:56:17 +00:00
Manman Ren	68cf9fc45d	AsmWriterEmitter: increase the number of bits for OpcodeInfo from 32-bit to 48-bit if necessary, in order to reduce the generated code size. We have 900 cases not covered by OpcodeInfo in ATT AsmWriter and more in Intel AsmWriter and ARM AsmWriter. This patch reduced the clang Release build size by 50k, running on a Mac Pro. llvm-svn: 163814	2012-09-13 17:43:46 +00:00
Akira Hatanaka	fcdd9b120d	mips16: When copying operands in a conditional branch instruction, allow for immediate operands to be copied. Patch by Reed Kotler. llvm-svn: 163811	2012-09-13 17:12:37 +00:00
Jakob Stoklund Olesen	78b9f8fc67	Revert r163761 "Don't fold indexed loads into TCRETURNmi64." The patch caused "Wrong topological sorting" assertions. llvm-svn: 163810	2012-09-13 16:52:17 +00:00
Benjamin Kramer	15a257dadd	MemCpyOpt: When forming a memset from stores also take GEP constexprs into account. This is common when storing to global variables. llvm-svn: 163809	2012-09-13 16:29:49 +00:00
Nadav Rotem	97d44349c9	Fix an 80 char line limit. llvm-svn: 163808	2012-09-13 16:27:32 +00:00
Nadav Rotem	77a09ebbeb	Rename the flag which protects from escaped allocas, which may come from bugs in user code or in the compiler. Also, dont assert if the protection is not enabled. llvm-svn: 163807	2012-09-13 15:46:30 +00:00
Micah Villmow	1cc19105e7	The current implementation does not allow more than 32 types to be properly handled with target lowering. This doubles the size to 64bit types and easily allows extension to more types. llvm-svn: 163806	2012-09-13 15:24:43 +00:00
Micah Villmow	857186d5e4	Unify the emission of the calling conventions into a single function to reduce code duplication. llvm-svn: 163805	2012-09-13 15:11:12 +00:00
Silviu Baranga	b47bb94f93	This patch introduces A15 as a target in LLVM. llvm-svn: 163803	2012-09-13 15:05:10 +00:00
Nadav Rotem	24a822a5cb	Fix a dagcombine optimization. The optimization attempts to optimize a bitcast of fneg to integers by xoring the high-bit. This fails if the source operand is a vector because we need to negate each of the elements in the vector. Fix rdar://12281066 PR13813. llvm-svn: 163802	2012-09-13 14:54:28 +00:00
Nadav Rotem	2bd25fed29	Fix a typo. llvm-svn: 163801	2012-09-13 14:51:00 +00:00
Bill Wendling	fb1f6681a3	Use Nick's suggestion of storing a large NULL into the GV instead of memset, which requires TargetData. llvm-svn: 163799	2012-09-13 14:32:30 +00:00
Nadav Rotem	4e9ad06617	Stack Coloring: We have code that checks that all of the uses of allocas are within the lifetime zone. Sometime legitimate usages of allocas are hoisted outside of the lifetime zone. For example, GEPS may calculate the address of a member of an allocated struct. This commit makes sure that we only check (abort regions or assert) for instructions that read and write memory using stack frames directly. Notice that by allowing legitimate usages outside the lifetime zone we also stop checking for instructions which use derivatives of allocas. We will catch less bugs in user code and in the compiler itself. llvm-svn: 163791	2012-09-13 12:38:37 +00:00
Dmitri Gribenko	2bc1d483fe	Fix Doxygen issues: * wrap code blocks in \code ... \endcode; * refer to parameter names in paragraphs correctly (\arg is not what most people want -- it starts a new paragraph). llvm-svn: 163790	2012-09-13 12:34:29 +00:00
Dmitri Gribenko	c1ca1fe255	Fix a doxygen issue: these examples are supposed to be displayed preformatted. llvm-svn: 163787	2012-09-13 11:42:30 +00:00
Craig Topper	5c8ba0fc83	Fix function name in comment. llvm-svn: 163783	2012-09-13 07:26:59 +00:00
Nick Lewycky	32feef063d	Fix typo in comment. llvm-svn: 163782	2012-09-13 07:01:25 +00:00
Craig Topper	963305b450	Add a new compression type to ModRM table that detects when the memory modRM byte represent 8 instructions and the reg modRM byte represents up to 64 instructions. Reduces modRM table from 43k entreis to 25k entries. Based on a patch from Manman Ren. llvm-svn: 163774	2012-09-13 05:45:42 +00:00
Jim Grosbach	6a465cd8ef	MCJIT: relocation addends encoded in the target aren't quite so easy. The assumption that the target address for the relocation will always be sizeof(intptr_t) and will always contain an addend for the relocation value is very wrong. Default to no addend for now. rdar://12157052 llvm-svn: 163765	2012-09-13 01:24:37 +00:00
Jim Grosbach	54289619bf	MCJIT: Make sure to mask off non-type-field bits. When comparing to the macho relocation type enum value, make sure we're only comparing against the bits in the RelType that correspond. llvm-svn: 163764	2012-09-13 01:24:35 +00:00
Jim Grosbach	fb7ee84839	MCJIT: Pass the i386 MachO relocation type properly. llvm-svn: 163763	2012-09-13 01:24:32 +00:00
Jakob Stoklund Olesen	bfacef45eb	Don't fold indexed loads into TCRETURNmi64. We don't have enough GR64_TC registers when calling a varargs function with 6 arguments. Since %al holds the number of vector registers used, only %r11 is available as a scratch register. This means that addressing modes using both base and index registers can't be folded into TCRETURNmi64. <rdar://problem/12282281> llvm-svn: 163761	2012-09-13 00:25:00 +00:00
Bill Wendling	2e6e86606a	Introduce the __llvm_gcov_flush function. This function writes out the current values of the counters and then resets them. This can be used similarly to the __gcov_flush function to sync the counters when need be. For instance, in a situation where the application doesn't exit. <rdar://problem/12185886> llvm-svn: 163757	2012-09-13 00:09:55 +00:00
Eric Christopher	e341776c1e	Recommit, with fixes: Add some support for dealing with an object pointer on arguments. Part of rdar://9797999 which now supports adding the object pointer attribute to the subprogram as it should. llvm-svn: 163754	2012-09-12 23:36:19 +00:00
Akira Hatanaka	92a96e1da4	Misc. 1. Remove RA from list of allocatable registers 2. Enable d,y,r constraint inline assembly instructions Patch by Reed Kotler. llvm-svn: 163753	2012-09-12 23:27:55 +00:00
Michael Liao	abb87d4857	Fix PR11985 - BlockAddress has no support of BA + offset form and there is no way to propagate that offset into machine operand; - Add BA + offset support and a new interface 'getTargetBlockAddress' to simplify target block address forming; - All targets are modified to use new interface and X86 backend is enhanced to support BA + offset addressing. llvm-svn: 163743	2012-09-12 21:43:09 +00:00
Dan Gohman	7c84dad80a	Detect overflow in the path count computation. rdar://12277446. llvm-svn: 163739	2012-09-12 20:45:17 +00:00
Owen Anderson	6f9dace01c	Remove an overly-aggressive assertion. The code following this assertion already knows how to handle the case where DstRC was NULL, so it's not actually protecting us from anything, and this pattern can come up when using unknown_class operands in the SelectionDAG. llvm-svn: 163736	2012-09-12 20:09:19 +00:00
Jakob Stoklund Olesen	5a3db551a8	Delete dead code. llvm-svn: 163735	2012-09-12 20:04:17 +00:00
Eric Christopher	c44e973a36	Revert "Add some support for dealing with an object pointer on arguments." This should be done on the subprogram, not the variable itself. llvm-svn: 163734	2012-09-12 18:42:31 +00:00
Chad Rosier	ab53b4f6d0	[ms-inline asm] Make the operand size directives case insensitive. llvm-svn: 163729	2012-09-12 18:24:26 +00:00
Jim Grosbach	f6cb1ee75a	TableGen: Convert an assert() to a proper diagnostic. llvm-svn: 163726	2012-09-12 17:40:25 +00:00
Manman Ren	49dbe255e6	PGO: preserve branch-weight metadata when removing a case which jumps to the default target. llvm-svn: 163724	2012-09-12 17:04:11 +00:00
Dmitri Gribenko	881929c1b6	Fix a couple of Doxygen comment issues pointed out by -Wdocumentation. llvm-svn: 163721	2012-09-12 16:59:47 +00:00
Roman Divacky	10a448d45a	Enable exceptions handling on PPC64 now that cr misaligned spilling was fixed in r163713. llvm-svn: 163715	2012-09-12 15:29:32 +00:00
Alexander Potapenko	e512fb78a4	Suppress the warnings about unused parameters in changeColor() llvm-svn: 163714	2012-09-12 15:01:33 +00:00
Roman Divacky	c9e23d93ae	This patch corrects logic in PPCFrameLowering for save and restore of nonvolatile condition register fields across calls under the SVR4 ABIs. * With the 64-bit ABI, the save location is at a fixed offset of 8 from the stack pointer. The frame pointer cannot be used to access this portion of the stack frame since the distance from the frame pointer may change with alloca calls. * With the 32-bit ABI, the save location is just below the general register save area, and is accessed via the frame pointer like the rest of the save areas. This is an optional slot, so it must only be created if any of CR2, CR3, and CR4 were modified. * For both ABIs, save/restore logic is generated only if one of the nonvolatile CR fields were modified. I also took this opportunity to clean up an extra FIXME in PPCFrameLowering.h. Save area offsets for 32-bit GPRs are meaningless for the 64-bit ABI, so I removed them for correctness and efficiency. Fixes PR13708 and partially also PR13623. It lets us enable exception handling on PPC64. Patch by William J. Schmidt! llvm-svn: 163713	2012-09-12 14:47:47 +00:00
Roman Divacky	fd69009419	Add support for AMD Geode. llvm-svn: 163710	2012-09-12 14:36:02 +00:00
Kristof Beyls	e6b876f4e5	Fix constant folding through bitcasts by no longer relying on undefined behaviour (converting NaN values between float and double). SelectionDAG::getConstantFP(double Val, EVT VT, bool isTarget); should not be used when Val is not a simple constant (as the comment in SelectionDAG.h indicates). This patch avoids using this function when folding an unknown constant through a bitcast, where it cannot be guaranteed that Val will be a simple constant. llvm-svn: 163703	2012-09-12 11:25:02 +00:00
Nadav Rotem	9566ca9af8	Add a flag to disable the code that looks for allocas which escaped the lifetime regions. This is useful for debugging. No testcase because without this check we fail on assertions when finding escaped allocas. llvm-svn: 163702	2012-09-12 11:06:26 +00:00
James Molloy	c747cdae24	Add a function computeRegisterLiveness() to MachineBasicBlock. This uses analyzePhysReg() from r163694 to heuristically try and determine the liveness state of a physical register upon arrival at a particular instruction in a block. The search for liveness is clipped to a specific number of instructions around the target MachineInstr, in order to avoid degenerating into an O(N^2) algorithm. It tries to use various clues about how instructions around (both before and after) a given MachineInstr use that register, to determine its state at the MachineInstr. llvm-svn: 163695	2012-09-12 10:18:23 +00:00
James Molloy	381fab93d5	Add an analyzePhysReg() function to MachineOperandIteratorBase that analyses an instruction's use of a physical register, analogous to analyzeVirtReg. Rename RegInfo to VirtRegInfo so as not to be confused with the new PhysRegInfo. llvm-svn: 163694	2012-09-12 10:03:31 +00:00
Duncan Sands	66fc0e63a0	When calling print directly on a global (eg from the debugger) it was printing a newline that doesn't occur when printing other kinds of LLVM values. Move the printing of that newline elsewhere, making globals print the same as other values while leaving the output when printing an entire module unchanged. Patch by Saša Tomić. llvm-svn: 163693	2012-09-12 09:55:51 +00:00
Nadav Rotem	b9e2202049	Enable stack-coloring, in hope that the recent fixes will enable correct dragonegg self-hosting. llvm-svn: 163687	2012-09-12 07:58:35 +00:00
Lang Hames	c3d9a3d881	Make findLastUseBefore handle reg-unit liveness. findLastUseBefore was previous considering virtreg liveness only, leading to incorrect live intervals for reg units when instrs with physreg operands were moved up. llvm-svn: 163685	2012-09-12 06:56:16 +00:00
Craig Topper	ad495964f1	Indentation fixes. No functional change. llvm-svn: 163682	2012-09-12 06:20:41 +00:00
Manman Ren	49d684e1e2	Release build: guard dump functions with "#if !defined(NDEBUG) \|\| defined(LLVM_ENABLE_DUMP)" No functional change. Update r163344. llvm-svn: 163679	2012-09-12 05:06:18 +00:00
Nadav Rotem	8ff00989fc	Stack coloring: remove lifetime intervals which contain escaped allocas. The input program may contain intructions which are not inside lifetime markers. This can happen due to a bug in the compiler or due to a bug in user code (for example, returning a reference to a local variable). This commit adds checks that all of the instructions in the function and invalidates lifetime ranges which do not contain all of the instructions. llvm-svn: 163678	2012-09-12 04:57:37 +00:00
Eric Christopher	97c0fdd116	Add some support for dealing with an object pointer on arguments. Part of rdar://9797999 llvm-svn: 163667	2012-09-12 00:26:55 +00:00
Owen Anderson	16ba4b2d83	Improve tblgen code cleanliness: create an unknown_class, from which the unknown def inherits. Make tblgen check for that class, rather than checking for the def itself. llvm-svn: 163664	2012-09-11 23:47:08 +00:00
Owen Anderson	ccd682c695	Compute a map from register names to registers, rather than scanning the list of registers every time we want to look up a register by name. llvm-svn: 163659	2012-09-11 23:32:17 +00:00
Chad Rosier	e65effe382	Add documentation. llvm-svn: 163658	2012-09-11 23:20:20 +00:00
Chad Rosier	03efc5e1ee	Add a few virtual functions to the abstract MCParsedAsmOperand class. llvm-svn: 163655	2012-09-11 23:03:44 +00:00
Chad Rosier	4109983fbc	Rename the isMemory() function to isMem(). No functional change intended. llvm-svn: 163654	2012-09-11 23:02:35 +00:00
Manman Ren	19f49ac624	Release build: guard dump functions with "#if !defined(NDEBUG) \|\| defined(LLVM_ENABLE_DUMP)" No functional change. Update r163339. llvm-svn: 163653	2012-09-11 22:23:19 +00:00
Chad Rosier	b6b8e966d6	StringSwitchify. llvm-svn: 163649	2012-09-11 21:10:25 +00:00
Chad Rosier	30888b176a	Simplify logic. No functional change intended. llvm-svn: 163648	2012-09-11 20:57:04 +00:00
Chad Rosier	1778831a3d	[ms-inline asm] Split the parsing of IR asm strings into GCC and MS variants. Add support in the EmitMSInlineAsmStr() function for handling integer consts. llvm-svn: 163645	2012-09-11 19:09:56 +00:00
Manman Ren	571d9e4b80	SimplifyCFG: preserve branch-weight metadata when creating a new switch from a pair of switch/branch where both depend on the value of the same variable and the default case of the first switch/branch goes to the second switch/branch. Code clean up and fixed a few issues: 1> handling the case where some cases of the 2nd switch are invalidated 2> correctly calculate the weight for the 2nd switch when it is a conditional eq Testing case is modified from Alastair's original patch. llvm-svn: 163635	2012-09-11 17:43:35 +00:00

1 2 3 4 5 ...

84994 Commits