llvm-project

Commit Graph

Author	SHA1	Message	Date
Alexey Bataev	59b81e51d3	[OPENMP][DEBUG] Set proper address space info if required by target. Arguments, passed to the outlined function, must have correct address space info for proper Debug info support. Patch sets global address space for arguments that are mapped and passed by reference. Also, cuda-gdb does not handle reference types correctly, so reference arguments are represented as pointers. llvm-svn: 310360	2017-08-08 14:25:14 +00:00
Nikolai Bozhenov	ce25d41b67	[libclang] Fix PR34055 (incompatible update of clang-c/Index.h) Fixes a regression introduced by r308218. llvm-svn: 310359	2017-08-08 14:13:50 +00:00
Nemanja Ivanovic	979dcb6f09	[PowerPC] Don't crash on larger splats achieved through 1-byte splats We've implemented a 1-byte splat using XXSPLTISB on P9. However, LLVM will produce a 1-byte splat even for wider element BUILD_VECTOR nodes. This patch prevents crashing in that situation. Differential Revision: https://reviews.llvm.org/D35650 llvm-svn: 310358	2017-08-08 13:52:45 +00:00
Daniel Sanders	75b84fc5ce	[globalisel][tablegen] Remove unnecessary ; to satisfy ubuntu-gcc7.1-werror. llvm-svn: 310357	2017-08-08 13:21:26 +00:00
Nemanja Ivanovic	bed7136eee	Appease compilers that have the -Wcovered-switch-default switch. llvm-svn: 310356	2017-08-08 12:41:56 +00:00
Siddharth Bhat	9aca1cb519	[NFC] [PPCGCodeGen] Add missing REQUIRES: pollyacc line. llvm-svn: 310354	2017-08-08 12:26:37 +00:00
Siddharth Bhat	83fe6b546d	[ScopInfo] [NFC] Typo fix. "to conservative" -> "too conservative". llvm-svn: 310353	2017-08-08 12:26:32 +00:00
Amjad Aboud	6fa6813aec	[X86] Improved X86::CMOV to Branch heuristic. Resolved PR33954. This patch contains two more constraints that aim to reduce the noise cases where we convert CMOV into branch for small gain, and end up spending more cycles due to overhead. Differential Revision: https://reviews.llvm.org/D36081 llvm-svn: 310352	2017-08-08 12:17:56 +00:00
Kamil Rytarowski	70a3511bd5	Reuse interception_linux for NetBSD Summary: Part of the code inspired by the original work on libsanitizer in GCC 5.4 by Christos Zoulas. Sponsored by <The NetBSD Foundation> Reviewers: joerg, kcc, vitalybuka, filcab Reviewed By: vitalybuka Subscribers: llvm-commits, #sanitizers Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D36321 llvm-svn: 310351	2017-08-08 12:10:08 +00:00
Siddharth Bhat	71dfb3eb07	[Polly] [PPCGCodeGeneration] Handle failing of invariant load hoisting gracefully. To do this, we replicate what `CodeGeneration` does. We expose `markNodeUnreachable` from `CodeGeneration` to `PPCGCodeGeneration`. Differential Revision: https://reviews.llvm.org/D36457 llvm-svn: 310350	2017-08-08 12:00:59 +00:00
Kamil Rytarowski	e528bd2193	Define OFF_T as 64-bit integer on NetBSD Summary: All 32 and 64 bit NetBSD platforms define off_t as 64-bit integer. Part of the code inspired by the original work on libsanitizer in GCC 5.4 by Christos Zoulas. Sponsored by <The NetBSD Foundation> Reviewers: joerg, filcab, kcc, vitalybuka Reviewed By: vitalybuka Subscribers: emaste, kubamracek, llvm-commits Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D35553 llvm-svn: 310349	2017-08-08 11:40:15 +00:00
Michael Kruse	27c010a22e	[DeLICM] Properly handle PHI writes becoming empty partial writes. It is possible that partial writes are empty (write is never executed). In this case, when in PHINode's incoming edge is never taken such that the incoming write becomes an empty partial write, if enabled. The issue is that when converting the union_map to an map, it's space cannot be derived from the union_map itself. Rather, we need to determine its space independently. This fixes test-suite's MultiSource/Benchmarks/ASC_Sequoia/CrystalMk. llvm-svn: 310348	2017-08-08 11:27:12 +00:00
Alex Lorenz	4003a98eec	Darwin's toolchain should be initialized before openmp offloading is processed This fixes an 'openmp-offload.c' test failure introduced by r310263. llvm-svn: 310347	2017-08-08 11:22:21 +00:00
Nemanja Ivanovic	809fbfa6a1	[PowerPC] Eliminate compares - add i32 sext/zext handling for SETLE/SETGE Adds handling for SETLE/SETGE comparisons on i32 values. Furthermore, it adds the handling for the special case where RHS == 0. Differential Revision: https://reviews.llvm.org/D34048 llvm-svn: 310346	2017-08-08 11:20:44 +00:00
Alex Lorenz	7e9c478cda	Revert r310291, r310300 and r310332 because of test failure on Darwin The commit r310291 introduced the failure. r310332 was a test fix commit and r310300 was a followup commit. I reverted these two to avoid merge conflicts when reverting. The 'openmp-offload.c' test is failing on Darwin because the following run lines: // RUN: touch %t1.o // RUN: touch %t2.o // RUN: %clang -### -no-canonical-prefixes -fopenmp=libomp -fopenmp-targets=nvptx64-nvidia-cuda -save-temps -no-canonical-prefixes %t1.o %t2.o 2>&1 \ // RUN: \| FileCheck -check-prefix=CHK-TWOCUBIN %s trigger the following assertion: Driver.cpp:3418: assert(CachedResults.find(ActionTC) != CachedResults.end() && "Result does not exist??"); llvm-svn: 310345	2017-08-08 11:20:17 +00:00
Simon Pilgrim	ef44228acb	[DAGCombiner] Simplify shuffle mask index if the referenced input element is UNDEF Fixes one of the cases in PR34041. Differential Revision: https://reviews.llvm.org/D36393 llvm-svn: 310344	2017-08-08 11:03:30 +00:00
Daniel Sanders	0554004698	[globalisel][tablegen] Add support for importing 'imm' operands. Summary: This patch enables the import of rules containing 'imm' operands that do not constrain the acceptable values using predicates. Support for ImmLeaf will arrive in a later patch. Depends on D35681 Reviewers: ab, t.p.northover, qcolombet, rovka, aditya_nandakumar Reviewed By: rovka Subscribers: kristof.beyls, javed.absar, igorb, llvm-commits Differential Revision: https://reviews.llvm.org/D35833 llvm-svn: 310343	2017-08-08 10:44:31 +00:00
Chandler Carruth	6e35c31d2d	[PM] Fix a likely more critical infloop bug in the CGSCC pass manager. This was just a bad oversight on my part. The code in question should never have worked without this fix. But it turns out, there are relatively few places that involve libfunctions that participate in a single SCC, and unless they do, this happens to not matter. The effect of not having this correct is that each time through this routine, the edge from write_wrapper to write was toggled between a call edge and a ref edge. First time through, it becomes a demoted call edge and is turned into a ref edge. Next time it is a promoted call edge from a ref edge. On, and on it goes forever. I've added the asserts which should have always been here to catch silly mistakes like this in the future as well as a test case that will actually infloop without the fix. The other (much scarier) infinite-inlining issue I think didn't actually occur in practice, and I simply misdiagnosed this minor issue as that much more scary issue. The other issue is still a real issue, but I'm somewhat relieved that so far it hasn't happened in real-world code yet... llvm-svn: 310342	2017-08-08 10:13:23 +00:00
Abhishek Aggarwal	95bd95c075	Checking in files accidentally missed in later diffs of revision r310261 -- 2 files were missing in this commit which should have been there. These files were submitted initially for review and were reviewed. However, while updating the revision with newer diffs, I accidentally forgot to include them in newer diffs. So commiting now. llvm-svn: 310341	2017-08-08 09:25:50 +00:00
Siddharth Bhat	8ff723dcf1	[NFC] [GPUJIT] Print line number & size information on allocateMemoryForDeviceCuda failure - It's useful to know the amount of memory asked for since, for example, asking for `0` bytes of memory is illegal. - Line number is helpful since we print the same message in the function at different points. llvm-svn: 310340	2017-08-08 09:03:27 +00:00
Craig Topper	8e351e9018	[InstCombine] Cast to BinaryOperator earlier in foldSelectIntoOp to simplify the code. We no longer need the explicit operand count check or the later dynamic cast. llvm-svn: 310339	2017-08-08 06:19:24 +00:00
Tobias Grosser	327e9ecb0d	[ScheduleOptimizer] Make matmul pattern detection work with delicm output In certain cases delicm might decide to not leave the original array write in the loop body, but to remove it and instead leave a transformed phi node as write access. This commit teached the matmul pattern detection to order the memory accesses according to when the access actually happens and use this information to detect the new pattern. This makes pattern based matmul optimization work for 2mm and 3mm in polybench 4 after polly-position=before-vectorizer has been enabled. llvm-svn: 310338	2017-08-08 06:15:15 +00:00
Tom Stellard	03aa3aee11	AMDGPU: Fix warnings introduced by r310336 llvm-svn: 310337	2017-08-08 05:52:00 +00:00
Tom Stellard	20287697f8	AMDGPU: Move R600 parts of AMDGPUISelDAGToDAG into their own class Summary: This refactoring is required in order to split the R600 and GCN tablegen files. Reviewers: arsenm Subscribers: kzhuravl, wdng, nhaehnle, yaxunl, dstuttard, tpr, llvm-commits, t-tye Differential Revision: https://reviews.llvm.org/D36286 llvm-svn: 310336	2017-08-08 04:57:55 +00:00
Konstantin Zhuravlyov	6cbcb27b5e	AMDGPU: Also remove SI from docs Differential Revision: https://reviews.llvm.org/D36424 llvm-svn: 310335	2017-08-08 04:28:31 +00:00
Chandler Carruth	e32ebcaa87	[PM] Relax the spelling of a pass name slightly in this test. I forgot that MSVC doesn't preserve this typedef, my bad. llvm-svn: 310334	2017-08-08 02:27:49 +00:00
Chandler Carruth	7c888dca46	[PM] Fix new LoopUnroll function pass by invalidating loop analysis results when a loop is completely removed. This is very hard to manifest as a visible bug. You need to arrange for there to be a subsequent allocation of a 'Loop' object which gets the exact same address as the one which the unroll deleted, and you need the LoopAccessAnalysis results to be significant in the way that they're stale. And you need a million other things to align. But when it does, you get a deeply mysterious crash due to actually finding a stale analysis result. This fixes the issue and tests for it by directly checking we successfully invalidate things. I have not been able to get any test case to reliably trigger this. Changes to LLVM itself caused the only test case I ever had to cease to crash. I've looked pretty extensively at less brittle ways of fixing this and they are actually very, very hard to do. This is a somewhat strange and unusual case as we have a pass which is deleting an IR unit, but is not running within that IR unit's pass framework (which is what handles this cleanly for the normal loop unroll). And where there isn't a definitive way to clear all of the stale cache entries. And where the pass is updating the core analysis that provides the IR units! For example, we don't have any of these problems with Function analyses because it is easy to clear out function analyses when the functions themselves may have been deleted -- we clear an entire module's worth! But that is too heavy of a hammer down here in the LoopAnalysisManager layer. A better long-term solution IMO is to require that AnalysisManager's make their keys durable to this kind of thing. Specifically, when caching an analysis for one IR unit that is conceptually "owned" by a higher level IR unit, the AnalysisManager should incorporate this into its data structures so that we can reliably clear these results without having to teach each and every pass to do so manually as we do here. But that is a change for another day as it will be a fairly invasive change to the AnalysisManager infrastructure. Until then, this fortunately seems to be quite rare. llvm-svn: 310333	2017-08-08 02:24:20 +00:00
Reid Kleckner	908ac3916d	Fix openmp-offload.c test on Windows llvm-svn: 310332	2017-08-08 01:36:16 +00:00
Reid Kleckner	59d1220cfd	[codeview] Fix class name formatting In particular, removes spaces between template arguments of class templates to better match VS type visualizers. llvm-svn: 310331	2017-08-08 01:33:53 +00:00
Vitaly Buka	4bc6c466b8	[asan] Restore dead-code-elimination optimization for Fuchsia Summary: r310244 fixed a bug introduced by r309914 for non-Fuchsia builds. In doing so it also reversed the intended effect of the change for Fuchsia builds, which was to allow all the AllocateFromLocalPool code and its variables to be optimized away entirely. This change restores that optimization for Fuchsia builds, but doesn't have the original change's bug because the comparison arithmetic now takes into account the size of the elements. Submitted on behalf of Roland McGrath. Reviewers: vitalybuka, alekseyshl Reviewed By: alekseyshl Subscribers: llvm-commits, kubamracek Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D36430 llvm-svn: 310330	2017-08-08 01:01:59 +00:00
Shoaib Meenai	1285013dbe	[libc++abi] Use proper calling convention for TLS destructor This is needed when using Windows threading. llvm-svn: 310329	2017-08-08 00:54:33 +00:00
Eugene Zelenko	59e128266c	[AMDGPU] Fix some Clang-tidy modernize-use-using and Include What You Use warnings; other minor fixes (NFC). llvm-svn: 310328	2017-08-08 00:47:13 +00:00
Petr Hosek	7ec1a56baf	[CMake] Allow overriding lib dir suffix independently from LLVM This matches the options already supported by libc++ and libc++abi. Differential Revision: https://reviews.llvm.org/D36383 llvm-svn: 310327	2017-08-08 00:37:59 +00:00
Kostya Serebryany	e863796dca	[libFuzzer] simplify code, NFC llvm-svn: 310326	2017-08-08 00:17:20 +00:00
Kostya Serebryany	22e5f9a16a	[libFuzzer] remove stale code llvm-svn: 310325	2017-08-08 00:14:49 +00:00
Kostya Serebryany	854be98c93	[libFuzzer] simplify the implementation of -print_coverage=1 llvm-svn: 310324	2017-08-08 00:12:09 +00:00
Kamil Rytarowski	1b39be7867	Fix asan_test.cc build on NetBSD Summary: Include <stdarg.h> for variable argument list macros (va_list, va_start etc). Add fallback definition of _LIBCPP_GET_C_LOCALE, this is required for GNU libstdc++ compatibility. Define new macro SANITIZER_GET_C_LOCALE. This value is currently required for FreeBSD and NetBSD for printf_l(3) tests. Sponsored by <The NetBSD Foundation> Reviewers: joerg, kcc, vitalybuka, filcab, fjricci Reviewed By: vitalybuka Subscribers: llvm-commits, emaste, kubamracek, #sanitizers Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D36406 llvm-svn: 310323	2017-08-07 23:38:14 +00:00
Kamil Rytarowski	123f62d515	Add NetBSD support in asan_stack.h Summary: Part of the code inspired by the original work on libsanitizer in GCC 5.4 by Christos Zoulas. Sponsored by <The NetBSD Foundation> Reviewers: joerg, kcc, vitalybuka, filcab, fjricci Reviewed By: vitalybuka Subscribers: davide, kubamracek, llvm-commits, #sanitizers Tags: #sanitizers Differential Revision: https://reviews.llvm.org/D36377 llvm-svn: 310322	2017-08-07 23:34:45 +00:00
Craig Topper	0ff0d74187	[KnownBits] Fix copy pasto in comment. NFC llvm-svn: 310320	2017-08-07 22:35:55 +00:00
Tobias Grosser	50206d8f57	Change Polly's position to "before-vectorizer" Polly has traditionally always been executed at the beginning of the pass pipeline as LLVM's inliner and DeLICM passes introduced plenty of scalar dependences which prevented any kind of useful high-level loop optimizations later in the pass pipeline. With DeLICM now being available, Polly can also run optimizations when folded into the pass pipeline. This has the benefit that Polly should now be more effective on C++ code and as an additional bonus, no additional early canonicalization phase must be run. As a result, Polly touches the code only if it applies a transformation. Code that does not benefit from Polly is not touched and consequently will have the very same execution time as without Polly enabled. Random performance changes, as could sometimes be observed with polly-position=early are consequently not possible any more. If performance is changed, this is due to Polly is choosing to perform a transformation. If this choice is wrong, it can be fixed directly in Polly. http://polly.llvm.org/docs/Architecture.html#polly-in-the-llvm-pass-pipeline llvm-svn: 310319	2017-08-07 22:33:34 +00:00
Sean Callanan	2b3a54bafc	This adds the argument --dump-ir to clang-import-test, which allows viewing of the final IR. This is useful for confirming that structure layout was correct. I've added two tests: - A test that checks that structs in top-level code are completed correctly during struct layout (they are) - A test that checks that structs defined in function bodies are cpmpleted correctly during struct layout (currently they are not, so this is XFAIL). The second test fails because LookupSameContext() (ExternalASTMerger.cpp) can't find the struct. This is an issue I intend to resolve separately. Differential Revision: https://reviews.llvm.org/D36429 llvm-svn: 310318	2017-08-07 22:27:30 +00:00
Simon Pilgrim	8c1167df5c	[X86][AVX] Added test for broadcast shuffle from binary sources with undefs (D36393) llvm-svn: 310317	2017-08-07 22:20:06 +00:00
Tobias Grosser	736c44c848	[test] Add some missing options that become necessary after the recent default changes llvm-svn: 310315	2017-08-07 22:10:23 +00:00
Tobias Grosser	32f64ed22b	[DeLICM] Enable partial writes This allows us to remove more scalar dependences. While this feature is still rather experimental, we want to give it sufficient test coverage. llvm-svn: 310314	2017-08-07 22:06:07 +00:00
Tobias Grosser	ad73f6a7b3	Enable delicm to automatically remove scalar loop carried dependences While this code is still rather we enable it by default to get better test coverage. llvm-svn: 310313	2017-08-07 22:04:20 +00:00
Tobias Grosser	a98081c9f5	[test] Add one more test case for the previous commit llvm-svn: 310312	2017-08-07 22:02:06 +00:00
Tobias Grosser	2ef378120d	[ZoneAlgo] Allow two writes that write identical values into same array slot Two write statements which write into the very same array slot generally are conflicting. However, in case the value that is written is identical, this does not cause any problem. Hence, allow such write pairs in this specific situation. llvm-svn: 310311	2017-08-07 22:01:29 +00:00
Matt Arsenault	bd57cea6e4	AMDGPU: Implement getMinimumNopSize llvm-svn: 310310	2017-08-07 22:00:58 +00:00
Reid Kleckner	ad7dc6e31f	[Object] Initialize LoadConfig member to null Executables may not contain a load config, and clients should be able to test for nullability. Previously we'd return uninitialized memory. Now getLoadConfig32/64 return valid pointers or null. Fixes PR34108 llvm-svn: 310308	2017-08-07 21:23:38 +00:00
Gheorghe-Teodor Bercea	ef5e106fc1	[OpenMP] Error when trying to offload to an unsupported architecture Summary: Throw an error when offloading is unsupported for a particular target architecture. Reviewers: sfantao, caomhin, carlo.bertolli, ABataev, Hahnfeld Reviewed By: ABataev Subscribers: cfe-commits, rengolin Differential Revision: https://reviews.llvm.org/D32035 llvm-svn: 310307	2017-08-07 21:11:10 +00:00

1 2 3 4 5 ...

269030 Commits All Branches Search

269030 Commits

All Branches