llvm-project

Commit Graph

Author	SHA1	Message	Date
Hiroshi Inoue	9ff2380ea6	[NFC] fix trivial typos in comments and error message "is is" -> "is", "are are" -> "are" llvm-svn: 329546	2018-04-09 04:37:53 +00:00
Fangrui Song	c244a15801	[ADT] Simplify getMemory. NFC llvm-svn: 328334	2018-03-23 17:26:12 +00:00
Tim Shen	89337750a0	[APInt] Fix extractBits to correctly handle Result.isSingleWord() case. Summary: extractBits assumes that `!this->isSingleWord() implies !Result.isSingleWord()`, which may not necessarily be true. Handle both cases. Reviewers: RKSimon Subscribers: sanjoy, llvm-commits, hiraditya Differential Revision: https://reviews.llvm.org/D43363 llvm-svn: 325311	2018-02-16 01:44:36 +00:00
Craig Topper	03106bb40e	Recommit r318963 "[APInt] Don't print debug messages from the APInt knuth division algorithm by default" The previous commit had the condition in the do/while backwards. Debug builds currently print out low level details of the Knuth division algorithm when -debug is used. This information isn't useful in most cases and just adds noise to the log. This adds a new preprocessor flag to enable the prints in the knuth division code in APInt. Differential Revision: https://reviews.llvm.org/D40404 llvm-svn: 318966	2017-11-24 20:29:04 +00:00
Craig Topper	8375bec71e	Revert 318963 "[APInt] Don't print debug messages from the APInt knuth division algorithm by default" I seem to have botched the logic when switching to push_macro llvm-svn: 318964	2017-11-24 19:32:34 +00:00
Craig Topper	960c4e3bb4	[APInt] Don't print debug messages from the APInt knuth division algorithm by default Debug builds currently print out low level details of the Knuth division algorithm when -debug is used. This information isn't useful in most cases and just adds noise to the log. This adds a new preprocessor flag to enable the prints in the knuth division code in APInt. Differential Revision: https://reviews.llvm.org/D40404 llvm-svn: 318963	2017-11-24 19:13:24 +00:00
Aaron Ballman	615eb47035	Reverting r315590; it did not include changes for llvm-tblgen, which is causing link errors for several people. Error LNK2019 unresolved external symbol "public: void __cdecl `anonymous namespace'::MatchableInfo::dump(void)const " (?dump@MatchableInfo@?A0xf4f1c304@@QEBAXXZ) referenced in function "public: void __cdecl `anonymous namespace'::AsmMatcherEmitter::run(class llvm::raw_ostream &)" (?run@AsmMatcherEmitter@?A0xf4f1c304@@QEAAXAEAVraw_ostream@llvm@@@Z) llvm-tblgen D:\llvm\2017\utils\TableGen\AsmMatcherEmitter.obj 1 llvm-svn: 315854	2017-10-15 14:32:27 +00:00
Don Hinton	3e0199f7eb	[dump] Remove NDEBUG from test to enable dump methods [NFC] Summary: Add LLVM_FORCE_ENABLE_DUMP cmake option, and use it along with LLVM_ENABLE_ASSERTIONS to set LLVM_ENABLE_DUMP. Remove NDEBUG and only use LLVM_ENABLE_DUMP to enable dump methods. Move definition of LLVM_ENABLE_DUMP from config.h to llvm-config.h so it'll be picked up by public headers. Differential Revision: https://reviews.llvm.org/D38406 llvm-svn: 315590	2017-10-12 16:16:06 +00:00
Craig Topper	405165210b	[APInt] Move the single word cases of countTrailingZeros and countLeadingOnes inline for consistency with countTrailingOnes and countLeadingZeros. NFCI llvm-svn: 306153	2017-06-23 20:28:45 +00:00
Craig Topper	e6a2318573	[APInt] Use std::end to avoid mentioning the size of a local buffer repeatedly. llvm-svn: 303726	2017-05-24 07:00:55 +00:00
Craig Topper	8885f933b2	[APInt] Add support for dividing or remainder by a uint64_t or int64_t. Summary: This patch adds udiv/sdiv/urem/srem/udivrem/sdivrem methods that can divide by a uint64_t. This makes division consistent with all the other arithmetic operations. This modifies the interface of the divide helper method to work on raw arrays instead of APInts. This way we can pass the uint64_t in for the RHS without wrapping it in an APInt. This required moving all the Quotient and Remainder allocation handling up to the callers. For udiv/urem this was as simple as just creating the Quotient/Remainder with the right size when they were declared. For udivrem we have to rely on reallocate not changing the contents of the variable LHS or RHS is aliased with the Quotient or Remainder APInts. We also have to zero the upper bits of Remainder and Quotient that divide doesn't write to if lhsWords/rhsWords is smaller than the width. I've update the toString method to use the new udivrem. Reviewers: hans, dblaikie, RKSimon Reviewed By: RKSimon Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D33310 llvm-svn: 303431	2017-05-19 16:43:54 +00:00
Craig Topper	6a1d02024e	[APInt] Simplify a for loop initialization based on the fact that 'n' is known to be 1 by an earlier 'if'. llvm-svn: 303120	2017-05-15 22:01:03 +00:00
Craig Topper	2c9a70661c	[APInt] Use Lo_32/Hi_32/Make_64 in a few more places in the divide code. NFCI llvm-svn: 302983	2017-05-13 07:14:17 +00:00
Craig Topper	4b83b4d560	[APInt] Fix typo in comment. NFC llvm-svn: 302974	2017-05-13 00:35:30 +00:00
Craig Topper	b1a71cac4b	[APInt] Add early outs for a division by 1 to udiv/urem/udivrem We already counted the number of bits in the RHS so its pretty cheap to just check if the RHS is 1. Differential Revision: https://reviews.llvm.org/D33154 llvm-svn: 302953	2017-05-12 21:45:50 +00:00
Craig Topper	2579c7c69f	[APInt] In udivrem, remember the bit width in a local variable so we don't reread it from the LHS which might be aliased with Quotient or Remainder. This helped the compiler generate better code for the single word case. It was able to remember that the bit width was still a single word when it created the Remainder APInt and not create code for it possibly being multiword. llvm-svn: 302952	2017-05-12 21:45:44 +00:00
Craig Topper	4bdd621e93	[APInt] Add an assert to check for divide by zero in udivrem. NFC udiv and urem already had the same assert. llvm-svn: 302931	2017-05-12 18:19:01 +00:00
Craig Topper	06da0816fd	[APInt] Remove unnecessary checks of rhsWords==1 with lhsWords==1 from udiv and udivrem. NFC At this point in the code rhsWords is guaranteed to be non-zero and less than or equal to lhsWords. So if lhsWords is 1, rhsWords must also be 1. urem alread had the check removed so this makes all 3 consistent. llvm-svn: 302930	2017-05-12 18:18:57 +00:00
Craig Topper	8769403d49	[APInt] Fix a case where udivrem might delete and create a new allocation instead of reusing the original. llvm-svn: 302882	2017-05-12 07:21:09 +00:00
Craig Topper	a92fd0bebb	[APInt] Add a utility method to change the bit width and storage size of an APInt. Summary: This adds a resize method to APInt that manages deleting/allocating storage for an APInt and changes its bit width. Use this to simplify code in copy assignment and divide. The assignment code in particular was overly complicated. Treating every possible case as a separate implementation. I'm also pretty sure the clearUnusedBits code at the end was unnecessary. Since we always copying whole words from the source APInt. All unused bits should be clear in the source. Reviewers: hans, RKSimon Reviewed By: RKSimon Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D33073 llvm-svn: 302863	2017-05-12 01:46:01 +00:00
Craig Topper	dbd6219f81	[APInt] Remove an APInt copy from the return of APInt::multiplicativeInverse. llvm-svn: 302816	2017-05-11 18:40:53 +00:00
Craig Topper	3fbecadab6	[APInt] Fix typo in comment. NFC llvm-svn: 302815	2017-05-11 17:57:43 +00:00
Craig Topper	c59ced36aa	[APInt] Remove an unneeded extra temporary APInt from toString. Turns out udivrem can write its output to the same location as one of its inputs so the extra temporary isn't needed. llvm-svn: 302772	2017-05-11 07:10:43 +00:00
Craig Topper	b3c1f56737	[APInt] Use negate() instead of copying an APInt to negate it and then writing back over the original value. llvm-svn: 302770	2017-05-11 07:02:04 +00:00
Craig Topper	ef0114c4f0	[APInt] Add negate helper method to implement twos complement. Use it to shorten code. llvm-svn: 302716	2017-05-10 20:01:38 +00:00
Craig Topper	ecb97da108	[APInt] Make toString use udivrem instead of calling the divide helper method directly. Do a better job of reusing allocations while looping. NFCI This lets toString take advantage of the degenerate case checks in udivrem and is just generally cleaner. One minor downside of this is that the divisor APInt now needs to be the same size as Tmp which requires an additional allocation. But we were doing a poor job of reusing allocations before so the new code should still be an improvement. llvm-svn: 302704	2017-05-10 18:15:24 +00:00
Craig Topper	6271bc7146	[APInt] Use uint32_t instead of unsigned for the storage type throughout the divide code. Use Lo_32/Hi_32/Make_64 helpers instead of casts and shifts. NFCI llvm-svn: 302703	2017-05-10 18:15:20 +00:00
Craig Topper	f86b9d5063	[APInt] Use getRawData to slightly simplify some code. llvm-svn: 302702	2017-05-10 18:15:17 +00:00
Craig Topper	93eabae4aa	[APInt] Remove check for single word since single word was handled earlier in the function. NFC llvm-svn: 302701	2017-05-10 18:15:14 +00:00
Craig Topper	a584af5c8e	[APInt] Fix indentation of tcDivide. Combine variable declaration and initialization. llvm-svn: 302626	2017-05-10 07:50:17 +00:00
Craig Topper	62de039bd1	[APInt] Use getNumWords function in udiv/urem/udivrem instead of reimplementinging it. llvm-svn: 302625	2017-05-10 07:50:15 +00:00
Craig Topper	0acb6654d3	[APInt] Remove return value from tcFullMultiply. The description says it returns the number of words needed to represent the results. But the way it was coded it always returns (lhsWords + rhsWords) or (lhsWords + rhsWords - 1). But the result could be even smaller than that and it wouldn't tell you. No one uses the result today so rather than try to fix it, just remove it. llvm-svn: 302551	2017-05-09 16:47:33 +00:00
Craig Topper	3369f8cc4a	[APInt] Use default constructor instead of explicitly creating a 1-bit APInt in udiv and urem. NFC The default constructor does the same thing. llvm-svn: 302487	2017-05-08 23:49:54 +00:00
Craig Topper	24ae69515b	[APInt] Remove 'else' after 'return' in udiv and urem. NFC llvm-svn: 302486	2017-05-08 23:49:49 +00:00
Craig Topper	c96a84d813	[APInt] Modify tcMultiplyPart's overflow detection to not depend on 'i' from the earlier loop. NFC The value of 'i' is always the smaller of DstParts and SrcParts so we can just use that fact to write all the code in terms of SrcParts and DstParts. llvm-svn: 302408	2017-05-08 06:34:41 +00:00
Craig Topper	0cbab7cc7a	[APInt] Use std::min instead of writing the same thing with the ternary operator. NFC llvm-svn: 302407	2017-05-08 06:34:39 +00:00
Craig Topper	a6c142ab4d	[APInt] Remove 'else' after 'return' in tcMultiply methods. NFC llvm-svn: 302406	2017-05-08 06:34:36 +00:00
Craig Topper	f15bec5541	[APInt] Take advantage of new operator*=(uint64_t) to remove a temporary APInt. llvm-svn: 302403	2017-05-08 04:55:12 +00:00
Craig Topper	a51941f314	[APInt] Add support for multiplying by a uint64_t. This makes multiply similar to add, sub, xor, and, and or. llvm-svn: 302402	2017-05-08 04:55:09 +00:00
Craig Topper	93c68e1189	[APInt] Reduce number of allocations involved in multiplying. Reduce worst case multiply size Currently multiply is implemented in operator=. Operator makes a copy and uses operator= to modify the copy. Operator= itself allocates a temporary buffer to hold the multiply result as it computes it. Then copies it to the buffer in this. Operator= attempts to bound the size of the result based on the number of active bits in its inputs. It also has a couple special cases to handle 0 inputs without any memory allocations or multiply operations. The best case is that it calculates a single word regardless of input bit width. The worst case is that it calculates the a 2x input width result and drop the upper bits. Since operator* uses operator= it incurs two allocations, one for a copy of this and one for the temporary allocation. Neither of these allocations are kept after the method operation is done. The main usage in the backend appears to be ConstantRange::multiply which uses operator* rather than operator=. This patch moves the multiply operation to operator and implements operator= using it. This avoids the copy in operator. operator* now allocates a result buffer sized the same width as its inputs no matter what. This buffer will be used as the buffer for the returned APInt. Finally, we reuse tcMultiply to implement the multiply operation. This function is capable of not calculating additional upper words that will be discarded. This change does lose the special optimizations for the inputs using less words than their size implies. But it also removed the getActiveBits calls from all multiplies. If we think those optimizations are important we could look at providing additional bounds to tcMultiply to limit the computations. Differential Revision: https://reviews.llvm.org/D32830 llvm-svn: 302171	2017-05-04 17:00:41 +00:00
Craig Topper	b339c6dcc0	[APInt] Give the value union a name so we can remove assumptions on VAL being the larger member Currently several places assume the VAL member is always at least the same size as pVal. In particular for a memcpy in the move assignment operator. While this is a true assumption, it isn't good practice to assume this. This patch gives the union a name so we can write the memcpy in terms of the union itself. This also adds a similar memcpy to the move constructor where we previously just copied using VAL directly. This patch is mostly just a mechanical addition of the U in front of VAL and pVAL everywhere. But several constructors had to be modified since we can't directly initializer a field of named union from the initializer list. Differential Revision: https://reviews.llvm.org/D30629 llvm-svn: 302040	2017-05-03 15:46:24 +00:00
Craig Topper	9881bd9c1d	[APInt] Move APInt::getSplat out of line. I think this method is probably too complex to be inlined. llvm-svn: 301901	2017-05-02 06:32:27 +00:00
Craig Topper	1e91919ac1	[APInt] Move the setBit and clearBit methods inline. This makes setBit/clearBit more consistent with setBits which is already inlined. llvm-svn: 301900	2017-05-02 05:49:40 +00:00
Craig Topper	24e71017aa	[APInt] Use inplace shift methods where possible. NFCI llvm-svn: 301612	2017-04-28 03:36:24 +00:00
Craig Topper	1dec281104	[APInt] Simplify the zext and sext methods This replaces a hand written copy loop with a call to memcpy for both zext and sext. For sext, it replaces multiple if/else blocks propagating sign information forward. Now we just do a copy, a sign extension on the last copied word, a memset, and clearUnusedBits. Differential Revision: https://reviews.llvm.org/D32417 llvm-svn: 301201	2017-04-24 17:37:10 +00:00
Craig Topper	8b37326ae2	[APInt] Add ashrInPlace method and rewrite ashr to make a copy and then call ashrInPlace. This patch adds an in place version of ashr to match lshr and shl which were recently added. I've tried to make this similar to the lshr code with additions to handle the sign extension. I've also tried to do this with less if checks than the current ashr code by sign extending the original result to a word boundary before doing any of the shifting. This removes a lot of the complexity of determining where to fill in sign bits after the shifting. Differential Revision: https://reviews.llvm.org/D32415 llvm-svn: 301198	2017-04-24 17:18:47 +00:00
Craig Topper	c6b05684c6	[APInt] Fix repeated word in comments. NFC llvm-svn: 301192	2017-04-24 17:00:22 +00:00
Craig Topper	fc03d2d21f	[APInt] Make behavior of ashr by BitWidth consistent between single and multi word. Previously single word would always return 0 regardless of the original sign. Multi word would return all 0s or all 1s based on the original sign. Now single word takes into account the sign as well. llvm-svn: 301159	2017-04-24 05:38:26 +00:00
Craig Topper	652ca99622	[APInt] In sext single word case, use SignExtend64 and let the APInt constructor mask off any excess bits. The current code is trying to be clever with shifts to avoid needing to clear unused bits. But it looks like the compiler is unable to optimize out the unused bit handling in the APInt constructor. Given this its better to just use SignExtend64 and have more readable code. llvm-svn: 301133	2017-04-23 17:16:24 +00:00
Renato Golin	4abfb3d741	Revert "[APInt] Fix a few places that use APInt::getRawData to operate within the normal API." This reverts commit r301105, 4, 3 and 1, as a follow up of the previous revert, which broke even more bots. For reference: Revert "[APInt] Use operator<<= where possible. NFC" Revert "[APInt] Use operator<<= instead of shl where possible. NFC" Revert "[APInt] Use ashInPlace where possible." PR32754. llvm-svn: 301111	2017-04-23 12:15:30 +00:00
Renato Golin	cc4a9120f6	Revert "[APInt] Add ashrInPlace method and implement ashr using it. Also fix a bug in the shift by BitWidth handling." This reverts commit r301094, as it broke all ARM self-hosting bots. PR32754. llvm-svn: 301110	2017-04-23 12:02:07 +00:00
Craig Topper	5f68af0806	[APInt] Use operator<<= instead of shl where possible. NFC llvm-svn: 301103	2017-04-23 05:18:31 +00:00
Craig Topper	26af2a993a	[APInt] Add ashrInPlace method and implement ashr using it. Also fix a bug in the shift by BitWidth handling. For single word, shift by BitWidth was always returning 0, but for multiword it was based on original sign. Now single word matches multi word. llvm-svn: 301094	2017-04-22 22:00:03 +00:00
Craig Topper	3a29e3b8e7	[APInt] Remove unnecessary min with BitWidth from countTrailingOnesSlowCase. The unused upper bits are guaranteed to be 0 so we don't need to worry about accidentally counting them. llvm-svn: 301091	2017-04-22 19:59:11 +00:00
Craig Topper	5e113742e7	[APInt] Add WORD_MAX constant and use it instead of UINT64_MAX. NFC llvm-svn: 301069	2017-04-22 06:31:36 +00:00
Craig Topper	1dc8fc8bfa	[APInt] Add compare/compareSigned methods that return -1, 0, 1. Reimplement slt/ult and friends using them Currently sle and ule have to call slt/ult and eq to get the proper answer. This results in extra code for both calls and additional scans of multiword APInts. This patch replaces slt/ult with a compareSigned/compare that can return -1, 0, or 1 so we can cover all the comparison functions with a single call. While I was there I removed the activeBits calls and other checks at the start of the slow part of ult. Both of the activeBits calls potentially scan through each of the APInts separately. I can't imagine that's any better than just scanning them in parallel and doing the compares. Now we just share the code with tcCompare. These changes seem to be good for about a 7-8k reduction on the size of the opt binary on my local x86-64 build. Differential Revision: https://reviews.llvm.org/D32339 llvm-svn: 300995	2017-04-21 16:13:15 +00:00
Craig Topper	a8129a1122	[APInt] Add isSubsetOf method that can check if one APInt is a subset of another without creating temporary APInts This question comes up in many places in SimplifyDemandedBits. This makes it easy to ask without allocating additional temporary APInts. The BitVector class provides a similar functionality through its (IMHO badly named) test(const BitVector&) method. Though its output polarity is reversed. I've provided one example use case in this patch. I plan to do more as a follow up. Differential Revision: https://reviews.llvm.org/D32258 llvm-svn: 300851	2017-04-20 16:17:13 +00:00
Craig Topper	baa392e4e0	[APInt] Implement APInt::intersects without creating a temporary APInt in the multiword case Summary: This is a simple question we should be able to answer without creating a temporary to hold the AND result. We can also get an early out as soon as we find a word that intersects. Reviewers: RKSimon, hans, spatel, davide Reviewed By: hans, davide Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D32253 llvm-svn: 300812	2017-04-20 02:11:27 +00:00
Craig Topper	b3624e4f45	[APInt] Implement operator==(uint64_t) similar to ugt/ult(uint64_t) to remove one of the out of line EqualsSlowCase methods. llvm-svn: 300799	2017-04-19 23:57:51 +00:00
Craig Topper	c67fe57e1e	[APInt] Move the 'return this' from the slow cases of assignment operators inline. We should let the compiler see that the fast/slow cases both return this. I don't think we chain assignments together very often so this shouldn't matter much. llvm-svn: 300715	2017-04-19 17:01:58 +00:00
Craig Topper	ae8bd67d96	[APInt] Inline the single word case of lshrInPlace similar to what we do for <<=. llvm-svn: 300577	2017-04-18 19:13:27 +00:00
Craig Topper	fc947bcfba	[APInt] Use lshrInPlace to replace lshr where possible This patch uses lshrInPlace to replace code where the object that lshr is called on is being overwritten with the result. This adds an lshrInPlace(const APInt &) version as well. Differential Revision: https://reviews.llvm.org/D32155 llvm-svn: 300566	2017-04-18 17:14:21 +00:00
Craig Topper	9eaef07519	[APInt] Cleanup the reverseBits slow case a little. Use lshrInPlace. Use single bit extract and operator\|=(uint64_t) to avoid a few temporary APInts. llvm-svn: 300527	2017-04-18 05:02:21 +00:00
Craig Topper	a8a4f0db79	[APInt] Make operator<<= shift in place. Improve the implementation of tcShiftLeft and use it to implement operator<<=. llvm-svn: 300526	2017-04-18 04:39:48 +00:00
Craig Topper	9575d8ff36	[APInt] Merge the multiword code from lshrInPlace and tcShiftRight into a single implementation This merges the two different multiword shift right implementations into a single version located in tcShiftRight. lshrInPlace now calls tcShiftRight for the multiword case. I retained the memmove fast path from lshrInPlace and used a memset for the zeroing. The for loop is basically tcShiftRight's implementation with the zeroing and the intra-shift of 0 removed. Differential Revision: https://reviews.llvm.org/D32114 llvm-svn: 300503	2017-04-17 21:43:43 +00:00
Craig Topper	9edfb08d93	[APInt] Fix a bug in lshr by a value more than 64 bits above the bit width. This was throwing an assert because we determined the intra-word shift amount by subtracting the size of the full word shift from the total shift amount. But we failed to account for the fact that we clipped the full word shifts by total words first. To fix this just calculate the intra-word shift as the remainder of dividing by bits per word. llvm-svn: 300405	2017-04-16 01:03:51 +00:00
Richard Smith	55bd375b69	Remove all allocation and divisions from GreatestCommonDivisor Switch from Euclid's algorithm to Stein's algorithm for computing GCD. This avoids the (expensive) APInt division operation in favour of bit operations. Remove all memory allocation from within the GCD loop by tweaking our `lshr` implementation so it can operate in-place. Differential Revision: https://reviews.llvm.org/D31968 llvm-svn: 300252	2017-04-13 20:29:59 +00:00
Craig Topper	90377de972	[APInt] Reorder fields to avoid a hole in the middle of the class Summary: APInt is currently implemented with an unsigned BitWidth field first and then a uint_64/pointer union. Due to the 64-bit size of the union there is a hole after the bitwidth. Putting the union first allows the class to be packed. Making it 12 bytes instead of 16 bytes. An APSInt goes from 20 bytes to 16 bytes. This shows a 4k reduction on the size of the opt binary on my local x86-64 build. So this enables some other improvement to the code as well. Reviewers: dblaikie, RKSimon, hans, davide Reviewed By: davide Subscribers: davide, llvm-commits Differential Revision: https://reviews.llvm.org/D32001 llvm-svn: 300171	2017-04-13 04:59:11 +00:00
Craig Topper	92fc477292	[APInt] Generalize the implementation of tcIncrement to support adding a full 'word' by introducing tcAddPart. Use this to support tcIncrement, operator++ and operator+=(uint64_t). Do the same for subtract. NFCI. llvm-svn: 300169	2017-04-13 04:36:06 +00:00
Craig Topper	00b47eec0f	[APInt] Make use of whichWord and maskBit to simplify some code. NFC llvm-svn: 299342	2017-04-02 19:35:18 +00:00
Craig Topper	55229b780d	[APInt] Add a public typedef for the internal type of APInt use it instead of integerPart. Make APINT_BITS_PER_WORD and APINT_WORD_SIZE public. This patch is one step to attempt to unify the main APInt interface and the tc functions used by APFloat. This patch adds a WordType to APInt and uses that in all the tc functions. I've added temporary typedefs to APFloat to alias it to integerPart to keep the patch size down. I'll work on removing that in a future patch. In future patches I hope to reuse the tc functions to implement some of the main APInt functionality. I may remove APINT_ from BITS_PER_WORD and WORD_SIZE constants so that we don't have the repetitive APInt::APINT_ externally. Differential Revision: https://reviews.llvm.org/D31523 llvm-svn: 299341	2017-04-02 19:17:22 +00:00
Craig Topper	15e484aa2b	[X86] Use tcAdd/tcSubtract to implement the slow case of operator+=/operator-=. llvm-svn: 299326	2017-04-02 06:59:43 +00:00
Craig Topper	b8f1068765	[APInt] Combine declaration and initialization. NFC llvm-svn: 299325	2017-04-02 06:59:41 +00:00
Craig Topper	b7d8faa231	[APInt] Simplify some code by using operator+=(uint64_t) instead of doing a more complex assignment into a temporary APInt just to use the APInt operator+=. llvm-svn: 299324	2017-04-02 06:59:38 +00:00
Craig Topper	d7ed50de26	[APInt] Fix typo in comment. NFC llvm-svn: 299323	2017-04-02 06:59:36 +00:00
Craig Topper	68a3ed2e1b	[APInt] Use conditional operator to simplify some code. NFC llvm-svn: 299320	2017-04-01 21:50:10 +00:00
Craig Topper	a742cb5fc8	[APInt] Implement flipAllBitsSlowCase with tcComplement. NFCI llvm-svn: 299319	2017-04-01 21:50:08 +00:00
Craig Topper	99cfe4f99d	[APInt] Fix indentation. NFC llvm-svn: 299318	2017-04-01 21:50:06 +00:00
Craig Topper	b2aaa5da42	[APInt] Implement AndAssignSlowCase using tcAnd. Do the same for Or and Xor. NFCI llvm-svn: 299317	2017-04-01 21:50:03 +00:00
Craig Topper	278ebd2f98	[APInt] Allow GreatestCommonDivisor to take rvalue inputs efficiently. Use moves instead of copies in the loop. Summary: GreatestComonDivisor currently makes a copy of both its inputs. Then in the loop we do one move and two copies, plus any allocation the urem call does. This patch changes it to take its inputs by value so that we can do a move of any rvalue inputs instead of copying. Then in the loop we do 3 move assignments and no copies. This way the only possible allocations we have in the loop is from the urem call. Reviewers: dblaikie, RKSimon, hans Reviewed By: dblaikie Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D31572 llvm-svn: 299314	2017-04-01 20:30:57 +00:00
Craig Topper	9ab8d7f9c3	[APInt] Remove the mul/urem/srem/udiv/sdiv functions from the APIntOps namespace. Replace the few usages with calls to the class methods. NFC llvm-svn: 299292	2017-04-01 05:08:57 +00:00
Craig Topper	e7e3560288	[APInt] Rewrite getLoBits in a way that will do one less memory allocation in the multiword case. Rewrite getHiBits to use the class method version of lshr instead of the one in APIntOps. NFCI llvm-svn: 299243	2017-03-31 18:48:14 +00:00
Craig Topper	6a8518086a	[APInt] Reformat tc functions to put opening curly braces on the end of the previous line. NFC llvm-svn: 298900	2017-03-28 05:32:55 +00:00
Craig Topper	76f4246fac	[APInt] Remove an anonymous namespace around static functions. NFC llvm-svn: 298899	2017-03-28 05:32:53 +00:00
Craig Topper	b003816be7	[APInt] Combine variable declaration and initialization where possible in the tc functions. NFCI llvm-svn: 298898	2017-03-28 05:32:52 +00:00
Craig Topper	592b134aa1	[APInt] Use 'unsigned' instead of 'unsigned int' in the interface to the APInt tc functions. This is more consistent with the rest of the codebase. NFC llvm-svn: 298897	2017-03-28 05:32:48 +00:00
Craig Topper	f496f9a273	[APInt] Move the single word cases of the bitwise operators inline. llvm-svn: 298894	2017-03-28 04:00:47 +00:00
Craig Topper	6ebeb7041e	[APInt] Move operator=(uint64_t) inline as its pretty simple and is often used with small constants that the compiler can optimize. While there recognize that we only need to clearUnusedBits on the single word case. llvm-svn: 298881	2017-03-27 20:07:31 +00:00
Craig Topper	70d8ca9276	[APInt] Move operator&=(uint64_t) inline and use memset to clear the upper words. This method is pretty new and probably isn't use much in the code base so this should have a negligible size impact. The OR and XOR operators are already inline. llvm-svn: 298870	2017-03-27 18:16:17 +00:00
Craig Topper	afc9e35343	[APInt] Move the >64 bit case for flipAllBits out of line. This is more consistent with what we do for other operations. This shrinks the opt binary on my build by ~72k. llvm-svn: 298858	2017-03-27 17:10:21 +00:00
Craig Topper	0085ffb940	[APInt] Don't initialize VAL to 0 in APInt constructors. Push it down to the initSlowCase and other init methods. I'm not sure if zeroing VAL before writing pVal is really necessary, but at least one other place did it in code. But by taking the store out of line, this reduces the opt binary by about 20k on my local x86-64 build. llvm-svn: 298233	2017-03-20 01:29:52 +00:00
Simon Pilgrim	b02667c469	[APInt] Add APInt::insertBits() method to insert an APInt into a larger APInt We currently have to insert bits via a temporary variable of the same size as the target with various shift/mask stages, resulting in further temporary variables, all of which require the allocation of memory for large APInts (MaskSizeInBits > 64). This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::insertBits() helper method which avoids the temporary memory allocation and masks/inserts the raw bits directly into the target. Differential Revision: https://reviews.llvm.org/D30780 llvm-svn: 297458	2017-03-10 13:44:32 +00:00
Simon Pilgrim	0099beb51f	Fixed typos in comments. NFCI. llvm-svn: 297379	2017-03-09 13:57:04 +00:00
Craig Topper	b60a46fea1	[APInt] Add rvalue reference support to and, or, xor operations to allow their memory allocation to be reused when possible This extends an earlier change that did similar for add and sub operations. With this first patch we lose the fastpath for the single word case as operator&= and friends don't support it. This can be added there if we think that's important. I had to change some functions in the APInt class since the operator overloads were moved out of the class and can't be used inside the class now. The getBitsSet change collides with another outstanding patch to implement it with setBits. But I didn't want to make this patch dependent on that series. I've also removed the Or, And, Xor functions which were rarely or never used. I already commited two changes to remove the only uses of Or that existed. Differential Revision: https://reviews.llvm.org/D30612 llvm-svn: 297121	2017-03-07 05:36:19 +00:00
Craig Topper	bafdd03b55	[APInt] Add setLowBits/setHighBits methods to APInt. Summary: There are quite a few places in the code base that do something like the following to set the high or low bits in an APInt. KnownZero \|= APInt::getHighBitsSet(BitWidth, BitWidth - 1); For BitWidths larger than 64 this creates a short lived APInt with malloced storage. I think it might even call malloc twice. Its better to just provide methods that can set the necessary bits without the temporary APInt. I'll update usages that benefit in a separate patch. Reviewers: majnemer, MatzeB, davide, RKSimon, hans Reviewed By: hans Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30525 llvm-svn: 297111	2017-03-07 01:56:01 +00:00
Craig Topper	f78a6f084c	[APInt] Optimize APInt creation from uint64_t Summary: This patch moves the clearUnusedBits calls into the two different initialization paths for APInt from a uint64_t. This allows the compiler to better optimize the clearing of the unused bits for the single word case. And it puts the clearing for the multi word case into the initSlowCase function to save code. In the common case of initializing with 0 this allows the clearing to be completely optimized out for the single word case. On my local x86 build this is showing a ~45kb reduction in the size of the opt binary. Reviewers: RKSimon, hans, majnemer, davide, MatzeB Reviewed By: hans Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30486 llvm-svn: 296677	2017-03-01 21:06:18 +00:00
Simon Pilgrim	0f5fb5f549	[APInt] Add APInt::extractBits() method to extract APInt subrange (reapplied) The current pattern for extract bits in range is typically: Mask.lshr(BitOffset).trunc(SubSizeInBits); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation of memory for the temporary variable. This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::extractBits() helper method which avoids the temporary memory allocation. Differential Revision: https://reviews.llvm.org/D30336 llvm-svn: 296272	2017-02-25 20:01:58 +00:00
Simon Pilgrim	cdf2bd656a	Revert: r296141 [APInt] Add APInt::extractBits() method to extract APInt subrange The current pattern for extract bits in range is typically: Mask.lshr(BitOffset).trunc(SubSizeInBits); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation of memory for the temporary variable. This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::extractBits() helper method which avoids the temporary memory allocation. Differential Revision: https://reviews.llvm.org/D30336 llvm-svn: 296147	2017-02-24 18:31:04 +00:00
Simon Pilgrim	bd9fb2ae95	[APInt] Add APInt::extractBits() method to extract APInt subrange The current pattern for extract bits in range is typically: Mask.lshr(BitOffset).trunc(SubSizeInBits); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation of memory for the temporary variable. This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::extractBits() helper method which avoids the temporary memory allocation. Differential Revision: https://reviews.llvm.org/D30336 llvm-svn: 296141	2017-02-24 17:46:18 +00:00
Simon Pilgrim	aed352273e	[APInt] Add APInt::setBits() method to set all bits in range The current pattern for setting bits in range is typically: Mask \|= APInt::getBitsSet(MaskSizeInBits, LoPos, HiPos); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation memory for the temporary variable. This is one of the key compile time issues identified in PR32037. This patch adds the APInt::setBits() helper method which avoids the temporary memory allocation completely, this first implementation uses setBit() internally instead but already significantly reduces the regression in PR32037 (~10% drop). Additional optimization may be possible. I investigated whether there is need for APInt::clearBits() and APInt::flipBits() equivalents but haven't seen these patterns to be particularly common, but reusing the code would be trivial. Differential Revision: https://reviews.llvm.org/D30265 llvm-svn: 296102	2017-02-24 10:15:29 +00:00

1 2 3 4 5 ...

395 Commits