llvm-project

Commit Graph

Author	SHA1	Message	Date
Simon Pilgrim	7bbe7a2920	[X86][SSE] Add basic PACKUS support to X86TargetLowering::computeKnownBitsForTargetNode Helps improve analysis of saturation ops llvm-svn: 333995	2018-06-05 09:45:03 +00:00
Alexander Ivchenko	964b27fa21	[X86][CET] Shadow stack fix for setjmp/longjmp This is the new version of D46181, allowing setjmp/longjmp to work correctly with the Intel CET shadow stack by storing SSP on setjmp and fixing it on longjmp. The patch has been updated to use the cf-protection-return module flag instead of HasSHSTK, and the bug that caused D46181 to be reverted has been fixed with the test expanded to track that fix. patch by mike.dvoretsky Differential Revision: https://reviews.llvm.org/D47311 llvm-svn: 333990	2018-06-05 09:22:30 +00:00
Craig Topper	f17b33d6c6	[X86] Make all instructions that operate on MMX types, but were added after the initial MMX support via one of the SSE features flags make them require the MMX feature as well. Passing -mattr=-mmx needs to disable these instructions since the MMX register class won't have been set up. But we don't want -mattr=-mmx to disable SSE so we have to do it separately. llvm-svn: 333984	2018-06-05 06:20:06 +00:00
Vedant Kumar	800255f9f1	[Debugify] Don't insert debug values after terminating deopts As is the case with musttail calls, the IR does not allow for instructions inserted after a terminating deopt. llvm-svn: 333976	2018-06-05 00:56:07 +00:00
Francis Visoiu Mistrih	ca69b3bf6d	[ShrinkWrap] Add optimization remarks to the shrink-wrapping pass Start by emitting remarks for very basic unsupported cases such as irreducible CFGs and EHFunclets. The end goal is to be able to cover all the cases where we give up with an explanation. llvm-svn: 333972	2018-06-05 00:27:24 +00:00
Amara Emerson	d496cc8ffb	[MIRParser] Add parser support for 'true' and 'false' i1s. We already output true and false in the printer, but the parser isn't able to read it. Differential Revision: https://reviews.llvm.org/D47424 llvm-svn: 333970	2018-06-05 00:17:13 +00:00
Amaury Sechet	800ac42573	Remove various use of undef in the X86 test suite as patern involving undef can collapse them. NFC llvm-svn: 333961	2018-06-04 22:09:26 +00:00
Amaury Sechet	e2729faf52	Revert "Regenerate expected test results for test/CodeGen/X86/pr23103.ll . NFC" This reverts commit cf25dfc503c861845947f3e6a9d308811ebb9da3. llvm-svn: 333960	2018-06-04 21:49:23 +00:00
Amaury Sechet	f5db3a15bf	Revert "Remove various use of undef in the X86 test suite as patern involving undef can collapse them. NFC" This reverts commit f0e85c194ae5e87476bc767304470dec85b6774f. llvm-svn: 333953	2018-06-04 21:20:45 +00:00
Alexander Ivchenko	2f038c4094	[X86][ELF][CET] Adding the .note.gnu.property ELF section in X86 In preparation for the proposed linker ABI changes (https://github.com/hjl-tools/linux-abi/wiki/linux-abi-draft.pdf, https://github.com/hjl-tools/x86-psABI/wiki/x86-64-psABI-cet.pdf), this patch enables emission of the .note.gnu.property section to ELF object files when building CET-enabled modules. patch by mike.dvoretsky Differential Revision: https://reviews.llvm.org/D47145 llvm-svn: 333951	2018-06-04 21:07:35 +00:00
Amaury Sechet	87f1a240ba	Remove various use of undef in the X86 test suite as patern involving undef can collapse them. NFC llvm-svn: 333950	2018-06-04 20:57:27 +00:00
Amaury Sechet	1910090328	Regenerate expected test results for test/CodeGen/X86/pr23103.ll . NFC llvm-svn: 333949	2018-06-04 20:47:00 +00:00
Scott Linder	ba81d7f1eb	[CodeGen] Always update divergence in SelectionDAG::UpdateNodeOperands Some overloads failed to update divergence. Differential Revision: https://reviews.llvm.org/D47148 llvm-svn: 333947	2018-06-04 20:19:45 +00:00
Amaury Sechet	da661e9236	[DAGcombine] Teach the combiner about -a = ~a + 1 Summary: This include variant for add, uaddo and addcarry. usubo and subcarry require the carry to be flipped to preserve semantic, but we chose to do the transform anyway in that case as to push the transform down the carry chain. Reviewers: efriedma, spatel, RKSimon, zvi, bkramer Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D46505 llvm-svn: 333943	2018-06-04 19:23:22 +00:00
Mark Searles	f0b93f1e9e	[AMDGPU][Waitcnt] Fix handling of flat instrs On GFX9 and earlier, flat memory ops may decrement VMCNT out-of-order as well as LGKMCNT out-of-order. Differential Revision: https://reviews.llvm.org/D46616 llvm-svn: 333926	2018-06-04 16:51:59 +00:00
Andrea Di Biagio	39e5a5695f	[RFC][patch 3/3] Add support for variant scheduling classes in llvm-mca. This patch is the last of a sequence of three patches related to LLVM-dev RFC "MC support for variant scheduling classes". http://lists.llvm.org/pipermail/llvm-dev/2018-May/123181.html This fixes PR36672. The main goal of this patch is to teach llvm-mca how to solve variant scheduling classes. This patch does that, plus it adds new variant scheduling classes to the BtVer2 scheduling model to identify so-called zero-idioms (i.e. so-called dependency breaking instructions that are known to generate zero, and that are optimized out in hardware at register renaming stage). Without the BtVer2 change, this patch would not have had any meaningful tests. This patch is effectively the union of two changes: 1) a change that teaches llvm-mca how to resolve variant scheduling classes. 2) a change to the BtVer2 scheduling model that allows us to special-case packed XOR zero-idioms (this partially fixes PR36671). Differential Revision: https://reviews.llvm.org/D47374 llvm-svn: 333909	2018-06-04 15:43:09 +00:00
Simon Dardis	fb4dde1142	[mips] Restore the availablity of trap for microMIPS Reviewers: smaksimovic, atanasyan, abeserminji Differential Revision: https://reviews.llvm.org/D47584 llvm-svn: 333895	2018-06-04 12:50:32 +00:00
Craig Topper	9923eac358	[X86] Remove and autoupgrade masked avx512vnni intrinsics using the unmasked intrinsics and select instructions. llvm-svn: 333857	2018-06-03 23:24:17 +00:00
Vedant Kumar	77f4d4d8aa	[Debugify] Skip dbg.value placement for EH pads, musttail Placing meta-instructions into EH pads breaks certain IR invariants, as does placing instructions after a musttail call. llvm-svn: 333856	2018-06-03 22:50:22 +00:00
Simon Pilgrim	7c4446ce0c	[X86][TBM] Use realistic BEXTR control bits Avoid constant values that are guaranteed to give zero Found while investigating BEXTR optimizations for PR34042. llvm-svn: 333849	2018-06-03 18:15:06 +00:00
Simon Pilgrim	1f60e2b41b	[X86][AVX512] Cleanup intrinsics tests Ensure we test on 32-bit and 64-bit targets, and strip -mcpu usage. Part of ongoing work to ensure we test all intrinsic style tests on 32 and 64 bit targets where possible. llvm-svn: 333843	2018-06-03 14:56:04 +00:00
Simon Pilgrim	7d717fed0b	[X86][AVX512BW] Regenerate arithmetic tests using update_llc_test_checks.py script Require manual stripping of existing CHECKs as update_llc_test_checks doesn't remove them if they're outside the function llvm-svn: 333842	2018-06-03 14:31:30 +00:00
Simon Pilgrim	e370ade180	[X86][BMI1] Test i32 intrinsics on 32/64 bits + branch off i64 tests Further refactoring will wait until D47452 has landed. Part of ongoing work to ensure we test all intrinsic style tests on 32 and 64 bit targets where possible. llvm-svn: 333841	2018-06-03 14:11:34 +00:00
Simon Pilgrim	8dc43621ec	[X86][BMI] Remove CTTZ tests - this is fully covered in clz.ll llvm-svn: 333840	2018-06-03 13:55:17 +00:00
Simon Pilgrim	d4ef869e28	[X86][TBM] Branch off i32 intrinsics and test on 32/64 bits Part of ongoing work to ensure we test all intrinsic style tests on 32 and 64 bit targets where possible. llvm-svn: 333839	2018-06-03 13:38:52 +00:00
Amaury Sechet	99909e9308	Remove SETCCE use from Lanai's backend Summary: This creates a small perf regression, but after talking with Jacques Pienaar, he was good with it to get things moving toward removng SETCCE. Reviewers: jpienaar, bryant Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D47626 llvm-svn: 333838	2018-06-03 12:56:24 +00:00
Simon Pilgrim	2b55e751ce	[X86][SSE] Cleanup AVX1 intrinsics tests Ensure we cover 32/64-bit targets for SSE/AVX/AVX512 cases as necessary, strip -mcpu usage. llvm-svn: 333834	2018-06-02 21:35:48 +00:00
Simon Pilgrim	58ff2ecc4b	[X86][SSE] Cleanup SSE1 intrinsics tests Ensure we cover 32/64-bit targets for SSE/AVX/AVX512 cases as necessary llvm-svn: 333833	2018-06-02 20:25:56 +00:00
Simon Pilgrim	8790844848	[X86][SSE] Cleanup SSE2 intrinsics tests Ensure we cover 32/64-bit targets for SSE/AVX/AVX512 cases as necessary llvm-svn: 333832	2018-06-02 19:43:14 +00:00
Simon Pilgrim	8c5b33a085	[X86][SSE] Cleanup SSE3/SSSE3 intrinsics tests Ensure we cover 32/64-bit targets for SSE/AVX/AVX512 cases as necessary llvm-svn: 333831	2018-06-02 18:41:46 +00:00
Simon Pilgrim	1c0fa05397	[X86][SSE4] Tweak rL333828 sse41/sse42 cleanup to recover SKX/EVEX2VEX testing Just testing for avx512f was missing the tests for EVEX TO VEX Compression encoding etc. llvm-svn: 333830	2018-06-02 18:01:09 +00:00
Simon Pilgrim	dda8daec73	[X86][SSE] Cleanup SSE4A/SSE41/SSE42 intrinsics tests Ensure we cover 32/64-bit targets for SSE/AVX/AVX512 cases as necessary Added some missing encoding checks to SSE4A tests llvm-svn: 333828	2018-06-02 17:33:26 +00:00
Simon Pilgrim	d93157c1b3	[X86][BMI2] Test i32 intrinsics on 32/64 bits + branch off i64 tests I had to tweak the i32 tests so we check both reg-reg and reg-mem cases. I also added i64 load tests. Part of ongoing work to ensure we test all intrinsic style tests on 32 and 64 bit targets where possible. llvm-svn: 333827	2018-06-02 17:22:13 +00:00
Simon Pilgrim	6028dc451a	[X86][BMI1] Remove test for non-existent andn i16 instruction llvm-svn: 333826	2018-06-02 17:02:27 +00:00
Ivan A. Kosarev	60a991ed1a	[NEON] Support VLD1xN intrinsics in AArch32 mode (LLVM part) We currently support them only in AArch64. The NEON Reference, however, says they are 'ARMv7, ARMv8' intrinsics. Differential Revision: https://reviews.llvm.org/D47120 llvm-svn: 333825	2018-06-02 16:40:03 +00:00
Ivan A. Kosarev	73c5337a64	Revert r333819 "[NEON] Support VLD1xN intrinsics in AArch32 mode (Clang part)" The LLVM part was committed instead of the Clang part. Differential Revision: https://reviews.llvm.org/D47121 llvm-svn: 333824	2018-06-02 16:38:38 +00:00
Ivan A. Kosarev	51f19b9ee1	[NEON] Support VLD1xN intrinsics in AArch32 mode (Clang part) We currently support them only in AArch64. The NEON Reference, however, says they are 'ARMv7, ARMv8' intrinsics. Differential Revision: https://reviews.llvm.org/D47121 llvm-svn: 333819	2018-06-02 16:26:42 +00:00
Craig Topper	3828ce7eab	[X86] Do something sensible when an expand load intrinsic is passed a 0 mask. Previously we just returned undef, but really we should be returning the pass thru input. We also need to make sure we preserve the chain output that the original intrinsic node had to maintain connectivity in the DAG. So we should just return the incoming chain as the output chain. llvm-svn: 333804	2018-06-01 22:59:07 +00:00
Craig Topper	aa747412b1	[X86] Add isel patterns to use vexpand with zero masking when the passthru value is a zero vector. llvm-svn: 333800	2018-06-01 22:28:28 +00:00
Craig Topper	c45479c08e	[X86] Expand the testing of expand and compress intrinsics The avx512f intrinsic tests were in the avx512vl file. We were also missing some combinations of masking. This does show that we fail to use the zero masking form of expand loads when the passthru is zero. I'll try to get that fixed shortly. llvm-svn: 333795	2018-06-01 21:59:24 +00:00
Craig Topper	d7e11ee342	[X86] Add fast-isel tests for avx512vbmi2 instructions. llvm-svn: 333794	2018-06-01 21:59:22 +00:00
Krzysztof Parzyszek	aec2c0c9b6	[Hexagon] Select HVX code for vector CTPOP, CTLZ, and CTTZ llvm-svn: 333760	2018-06-01 14:52:58 +00:00
Krzysztof Parzyszek	0b6187c1a9	[SelectionDAG] Expand UADDO/USUBO into ADD/SUBCARRY if legal for target Additionally, implement handling of ADD/SUBCARRY on Hexagon, utilizing the UADDO/USUBO expansion. Differential Revision: https://reviews.llvm.org/D47559 llvm-svn: 333751	2018-06-01 14:00:32 +00:00
Alexander Ivchenko	b34afcec5d	[x86] NFC. Reautogenerate test/CodeGen/X86/vector-half-conversions.ll llvm-svn: 333750	2018-06-01 13:51:53 +00:00
Simon Pilgrim	ee7694442d	[Utils][X86] Help update_llc_test_checks.py to recognise retl/retq to reduce CHECK duplication (PR35003) This patch replaces the --x86_extra_scrub command line argument to automatically support a second level of regex-scrubbing if it improves the matching of nearly-identical code patterns. The argument '--extra_scrub' is there now to force extra matching if required. This is mostly useful to help us share 32-bit/64-bit x86 vector tests which only differs by retl/retq instructions, but any scrubber can now technically support this, meaning test checks don't have to be needlessly obfuscated. I've updated some of the existing checks that had been manually run with --x86_extra_scrub, to demonstrate the extra "ret{{[l\|q]}}" scrub now only happens when useful, and re-run the sse42-intrinsics file to show extra matches - most sse/avx intrinsics files should be able to now share 32/64 checks. Tested with the opt/analysis scripts as well which share common code - AFAICT the other update scripts use their own versions. Differential Revision: https://reviews.llvm.org/D47485 llvm-svn: 333749	2018-06-01 13:37:01 +00:00
Amara Emerson	5a3bb68e12	[AArch64][GlobalISel] Zero-extend s1 values when returning. Before we were relying on the any extend of the s1 to s32, but for AAPCS we need to zero-extend it to at least s8. Fixes PR36719 Differential Revision: https://reviews.llvm.org/D47425 llvm-svn: 333747	2018-06-01 13:20:32 +00:00
Simon Dardis	ee67dcb837	[mips] Select the correct instruction for computing frameindexes Reviewers: smaksimovic, atanasyan, abeserminji Differential Revision: https://reviews.llvm.org/D47582 llvm-svn: 333736	2018-06-01 10:07:10 +00:00
Matt Arsenault	72a9f52c87	AMDGPU: Switch some half using-tests to use amdhsa The default clover ABI weirdly promotes half to float, which should probably be fixed. llvm-svn: 333730	2018-06-01 07:06:03 +00:00
Dan Gohman	91ab25bbe3	[WebAssembly] Update to the new names for the memory intrinsics. The WebAssembly committee has decided on the names `memory.size` and `memory.grow` for the memory intrinsics, so update the LLVM intrinsics to follow those names, keeping both sets of old names in place for compatibility. llvm-svn: 333708	2018-05-31 22:35:25 +00:00
Dan Gohman	b17de645ea	[WebAssembly] Fix the signatures for the __mulo* libcalls. The __mulo* libcalls have an extra i32* to return the overflow value. Fixes PR37401. llvm-svn: 333706	2018-05-31 22:27:24 +00:00

1 2 3 4 5 ...

24683 Commits