bcm5719-llvm - Project Ortega BCM5719 LLVM

	Commit message (Collapse)	Author	Age	Files	Lines
*	Revert "[Hexagon] Start using regmasks on calls"	Rafael Espindola	2017-02-17	18	-270/+115
\| \| \| \| \| \| \| \| \| \|	This reverts commit r295371. It broke windows bots: http://bb.pgr.jp/builders/ninja-clang-i686-msc19-R/builds/11402/steps/test-llvm/logs/stdio llvm-svn: 295402
*	[XRAY] [x86_64] Adding a Flight Data filetype reader to the llvm-xray Trace ↵	Dean Michael Berris	2017-02-17	1	-19/+304
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	implementation. Summary: The file type packs function trace data onto disk from potentially multiple threads that are aggregated and flushed during the course of an instrumented program's runtime. It is named FDR mode or Flight Data recorder as an analogy to plane blackboxes, which instrument a running system without access to IO. The writer code is defined in compiler-rt in xray_fdr_logging.h/cc Reviewers: rSerge, kcc, dberris Reviewed By: dberris Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D29697 llvm-svn: 295397
*	Bug 31948: Fix assertion when bitcasting constantexpr pointers	Matt Arsenault	2017-02-17	1	-0/+6
\| \| \| \|	llvm-svn: 295387
*	Handle link of NoDebug CU with a CU that has debug emission enabled	Teresa Johnson	2017-02-17	1	-0/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Summary: This is an issue both with regular and Thin LTO. When we link together a DICompileUnit that is marked NoDebug (e.g when compiling with -g0 but applying an AutoFDO profile, which requires location tracking in the compiler) and a DICompileUnit with debug emission enabled, we can have failures during dwarf debug generation. Specifically, when we have inlined from the NoDebug compile unit into the debug compile unit, we can fail during construction of the abstract and inlined scope DIEs. This is because the SPMap does not include NoDebug CUs (they are skipped in the debug_compile_units_iterator). This patch fixes the failures by skipping locations from NoDebug CUs when extracting lexical scopes. Reviewers: dblaikie, aprantl Subscribers: mehdi_amini, llvm-commits Differential Revision: https://reviews.llvm.org/D29765 llvm-svn: 295384
*	[IR] Fix some Clang-tidy modernize and Include What You Use warnings; other ↵	Eugene Zelenko	2017-02-17	6	-59/+115
\| \| \| \| \| \|	minor fixes (NFC). llvm-svn: 295383
*	[pdb] Add the ability to resolve TypeServer PDBs.	Zachary Turner	2017-02-16	9	-9/+213
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Some PDBs or object files can contain references to other PDBs where the real type information lives. When this happens, all type indices in the original PDB are meaningless because their records are not there. With this patch we add the ability to pull type info from those secondary PDBs. Differential Revision: https://reviews.llvm.org/D29973 llvm-svn: 295382
*	[LSR] Prevent formula with SCEVAddRecExpr type of Reg from Sibling loops	Wei Mi	2017-02-16	1	-0/+7
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	In rL294814, we allow formula with SCEVAddRecExpr type of Reg from loops other than current loop. This is good for the case when induction variable of outerloop being used in expr in innerloop. But it is very bad to allow such Reg from sibling loop because we may need to add lsr.iv in other sibling loops when scev expanding those SCEVAddRecExpr type exprs. For the testcase below, one loop can be inserted with a bunch of lsr.iv because of LSR for other loops. // The induction variable j from a loop in the middle will have initial // value generated from previous sibling loop and exit value used by its // next sibling loop. void goo(long i, long j); long cond; void foo(long N) { long i = 0; long j = 0; i = 0; do { goo(i, j); i++; j++; } while (cond); i = 0; do { goo(i, j); i++; j++; } while (cond); i = 0; do { goo(i, j); i++; j++; } while (cond); i = 0; do { goo(i, j); i++; j++; } while (cond); i = 0; do { goo(i, j); i++; j++; } while (cond); i = 0; do { goo(i, j); i++; j++; } while (cond); } The fix is to only allow formula with SCEVAddRecExpr type of Reg from current loop or its parents. Differential Revision: https://reviews.llvm.org/D30021 llvm-svn: 295378
*	Fix -Wunused-lambda-capture by removing some unused lambda captures	David Blaikie	2017-02-16	1	-2/+2
\| \| \| \|	llvm-svn: 295373
*	[MachinePipeliner] Remove redundant destructor. NFC.	Benjamin Kramer	2017-02-16	1	-8/+1
\| \| \| \|	llvm-svn: 295372
*	[Hexagon] Start using regmasks on calls	Krzysztof Parzyszek	2017-02-16	18	-115/+270
\| \| \| \| \| \|	All the cool targets are doing it... llvm-svn: 295371
*	Change default TimerGroup singleton to use magic statics	Erich Keane	2017-02-16	1	-16/+3
\| \| \| \| \| \| \| \| \| \| \|	TimerGroup was showing up on a leak in valigrind, and used some pretty complex code to implement a singleton. This patch replaces the implementation with a vastly simpler one. Differential Revision: https://reviews.llvm.org/D28367 llvm-svn: 295370
*	[RDF] Aggregate shadow phi uses into one cluster when propagating live info	Krzysztof Parzyszek	2017-02-16	2	-70/+68
\| \| \| \|	llvm-svn: 295366
*	AMDGPU: Remove llvm.AMDGPU.cube intrinsic	Matt Arsenault	2017-02-16	3	-25/+1
\| \| \| \|	llvm-svn: 295359
*	AMDGPU: Remove llvm.AMDGPU.rsq intrinsic	Matt Arsenault	2017-02-16	2	-6/+0
\| \| \| \|	llvm-svn: 295358
*	Re-apply r282920 "X86: Allow conditional tail calls in Win64 "leaf" ↵	Hans Wennborg	2017-02-16	2	-6/+6
\| \| \| \| \| \| \| \| \| \|	functions (PR26302)" The original commit was reverted in r283329 due to a miscompile in Chromium. That turned out to be the same issue as PR31257, which was fixed in r295262. llvm-svn: 295357
*	[RDF] Differentiate between defining and clobbering nodes	Krzysztof Parzyszek	2017-02-16	4	-13/+88
\| \| \| \| \| \| \| \| \| \|	Defining nodes should not alias with one another, while clobbering nodes can. When pushing defs on stacks, push clobbers first, link non-clobbering defs, then push the defs. The data flow in a statement is now: uses -> clobbers -> defs. llvm-svn: 295356
*	Refactor DebugHandlerBase a bit to common non-debug-having-function filtering	David Blaikie	2017-02-16	6	-54/+60
\| \| \| \|	llvm-svn: 295354
*	InstCombine: Canonicalize fast fmuladd to fmul + fadd	Matt Arsenault	2017-02-16	1	-1/+14
\| \| \| \|	llvm-svn: 295353
*	[RDF] Move normalize(RegisterRef) to PhysicalRegisterInfo	Krzysztof Parzyszek	2017-02-16	6	-45/+36
\| \| \| \| \| \|	Remove the duplicate from DFG and make some members of PRI private. llvm-svn: 295351
*	x86 interrupt calling convention: only save xmm registers if the target ↵	Andrea Di Biagio	2017-02-16	2	-2/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	supports SSE The existing code always saves the xmm registers for 64-bit targets even if the target doesn't support SSE (which is common for kernels). Thus, the compiler inserts movaps instructions which lead to CPU exceptions when an interrupt handler is invoked. This commit fixes this bug by returning a register set without xmm registers from getCalleeSavedRegs and getCallPreservedMask for such targets. Patch by Philipp Oppermann. Differential Revision: https://reviews.llvm.org/D29959 llvm-svn: 295347
*	[DAGCombiner] Support {a\|s}ext, {a\|z\|s}ext load nodes in load combine	Artur Pilipenko	2017-02-16	1	-8/+19
\| \| \| \| \| \| \| \| \| \| \| \|	Resubmit -r295314 with PowerPC and AMDGPU tests updated. Support {a\|s}ext, {a\|z\|s}ext load nodes as a part of load combine patters. Reviewed By: filcab Differential Revision: https://reviews.llvm.org/D29591 llvm-svn: 295336
*	[AArch64] AArch64AsmParser clean up of isImmediate functions. NFC	Sjoerd Meijer	2017-02-16	2	-144/+11
\| \| \| \| \| \| \| \| \| \| \|	Regression test neon-diagnostics.s needed changing because it now produces a more specific diagnostic about the immediate ranges. One change in the expected error message is not obvious, but there multiple candidate and it happens to pick the immediate diagnostic. Differential Revision: https://reviews.llvm.org/D29939 llvm-svn: 295331
*	[WebAssembly] Add a cast to void to fix an unused private member warning, ↵	Dan Gohman	2017-02-16	1	-1/+3
\| \| \| \| \| \|	for now. llvm-svn: 295327
*	[X86] Remove local areOnlyUsersOf helper and use SDNode::areOnlyUsersOf instead.	Simon Pilgrim	2017-02-16	1	-9/+1
\| \| \| \|	llvm-svn: 295326
*	[ARM] GlobalISel: Select floating point loads	Diana Picus	2017-02-16	1	-10/+31
\| \| \| \|	llvm-svn: 295321
*	Rever -r295314 "[DAGCombiner] Support {a\|s}ext, {a\|z\|s}ext load nodes in ↵	Artur Pilipenko	2017-02-16	1	-19/+8
\| \| \| \| \| \| \| \|	load combine" This change causes some of AMDGPU and PowerPC tests to fail. llvm-svn: 295316
*	[DAGCombiner] Support {a\|s}ext, {a\|z\|s}ext load nodes in load combine	Artur Pilipenko	2017-02-16	1	-8/+19
\| \| \| \| \| \| \| \| \| \|	Support {a\|s}ext, {a\|z\|s}ext load nodes as a part of load combine patters. Reviewed By: filcab Differential Revision: https://reviews.llvm.org/D29591 llvm-svn: 295314
*	[ARM] GlobalISel: Select G_SEQUENCE and G_EXTRACT	Diana Picus	2017-02-16	1	-0/+78
\| \| \| \| \| \| \| \|	Since they're only used for passing around double precision floating point values into the general purpose registers, we'll lower them to VMOVDRR and VMOVRRD. llvm-svn: 295310
*	[ARM] GlobalISel: Select double G_FADD and copies	Diana Picus	2017-02-16	1	-6/+29
\| \| \| \| \| \|	Just use VADDD if available, bail out if not. llvm-svn: 295309
*	[ARM] GlobalISel: Assert that we don't use the FPR bank if we don't have VFP	Diana Picus	2017-02-16	1	-0/+12
\| \| \| \|	llvm-svn: 295308
*	[ARM] GlobalISel: Add reg bank mappings for G_SEQUENCE and G_EXTRACT	Diana Picus	2017-02-16	1	-0/+26
\| \| \| \| \| \| \|	Support G_SEQUENCE and G_EXTRACT as needed for passing double precision floating point values in the soft-fp float mode. llvm-svn: 295306
*	[ARM] GlobalISel: Make the FPR bank 64-bit wide	Diana Picus	2017-02-16	2	-5/+22
\| \| \| \| \| \| \|	Also add mappings for single and double precision FP, and use them for G_FADD and G_LOAD. llvm-svn: 295302
*	[ARM] GlobalISel: Legalize 64-bit G_FADD and G_LOAD	Diana Picus	2017-02-16	1	-0/+7
\| \| \| \| \| \| \| \|	For now we just mark them as legal all the time and let the other passes bail out if they can't handle it. In the future, we'll want to move more of the brains into the legalizer. llvm-svn: 295300
*	[ARM] GlobalISel: Lower double precision FP args	Diana Picus	2017-02-16	2	-8/+85
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	For the hard float calling convention, we just use the D registers. For the soft-fp calling convention, we use the R registers and move values to/from the D registers by means of G_SEQUENCE/G_EXTRACT. While doing so, we make sure to honor the endianness of the target, since the CCAssignFn doesn't do that for us. For pure soft float targets, we still bail out because we don't support the libcalls yet. llvm-svn: 295295
*	[AVX-512][InstCombine] Teach InstCombine to optimize 512-bit packss/packus ↵	Craig Topper	2017-02-16	2	-4/+9
\| \| \| \| \| \|	intrinsics like it does 128/256-bit. llvm-svn: 295294
*	[AVX-512] Remove masked packss/packus intrinsics and autoupgrade to unmasked ↵	Craig Topper	2017-02-16	2	-12/+44
\| \| \| \| \| \| \| \|	intrinsics with select instructions. For 512-bit add new unmasked intrinsics. The new 512-bit unmasked intrinsics will make it easy to handle these with the SSE/AVX intrinsics in InstCombine where we currently have a TODO. llvm-svn: 295290
*	Split WinCOFFObjectWriter::writeSection.	Rui Ueyama	2017-02-16	1	-28/+39
\| \| \| \|	llvm-svn: 295276
*	Split WinCOFFObjectWriter::writeObject function.	Rui Ueyama	2017-02-16	1	-160/+183
\| \| \| \|	llvm-svn: 295273
*	AMDGPU: Remove llvm.SI.sendmsg	Matt Arsenault	2017-02-16	2	-6/+3
\| \| \| \|	llvm-svn: 295270
*	AMDGPU: Remove SI_fs_constant and SI_fs_interp intrinsics	Matt Arsenault	2017-02-16	3	-50/+3
\| \| \| \| \| \|	Update test uses with expansion in terms of new intrinsics. llvm-svn: 295269
*	Remove useless local variable.	Rui Ueyama	2017-02-16	1	-9/+4
\| \| \| \|	llvm-svn: 295268
*	Rename variables to match the LLVM style.	Rui Ueyama	2017-02-16	1	-94/+97
\| \| \| \|	llvm-svn: 295265
*	[X86] Re-enable conditional tail calls and fix PR31257.	Hans Wennborg	2017-02-16	6	-2/+193
\| \| \| \| \| \| \| \| \| \| \|	This reverts r294348, which removed support for conditional tail calls due to the PR above. It fixes the PR by marking live registers as implicitly used and defined by the now predicated tailcall. This is similar to how IfConversion predicates instructions. Differential Revision: https://reviews.llvm.org/D29856 llvm-svn: 295262
*	PMB: Add an importing WPD pass to the start of the ThinLTO backend pipeline.	Peter Collingbourne	2017-02-15	1	-1/+15
\| \| \| \| \| \|	Differential Revision: https://reviews.llvm.org/D30008 llvm-svn: 295260
*	GlobalISel: legalize va_arg on AArch64.	Tim Northover	2017-02-15	4	-0/+95
\| \| \| \| \| \| \| \|	Uses a Custom implementation because the slot sizes being a multiple of the pointer size isn't really universal, even for the architectures that do have a simple "void *" va_list. llvm-svn: 295255
*	GlobalISel: support translating va_arg	Tim Northover	2017-02-15	1	-0/+12
\| \| \| \| \| \| \|	Since (say) i128 and [16 x i8] map to the same type in generic MIR, we also need to attach the required alignment info. llvm-svn: 295254
*	Implement intrinsic mangling for literal struct types.	Daniel Berlin	2017-02-15	2	-6/+29
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Fixes PR 31921 Summary: Predicateinfo requires an ugly workaround to try to avoid literal struct types due to the intrinsic mangling not being implemented. This workaround actually does not work in all cases (you can hit the assert by bootstrapping with -print-predicateinfo), and can't be made to work without DFS'ing the type (IE copying getMangledStr and using a version that detects if it would crash). Rather than do that, i just implemented the mangling. It seems simple, since they are unified structurally. Looking at the overloaded-mangling testcase we have, it actually turns out the gc intrinsics will also crash if you try to use a literal struct. Thus, the testcase added fails before this patch, and works after, without needing to resort to predicateinfo. Reviewers: chandlerc, davide Subscribers: llvm-commits, sanjoy Differential Revision: https://reviews.llvm.org/D29925 llvm-svn: 295253
*	AMDGPU: Remove dead node definitions	Matt Arsenault	2017-02-15	1	-10/+0
\| \| \| \|	llvm-svn: 295247
*	Fix typos	Matt Arsenault	2017-02-15	2	-2/+2
\| \| \| \|	llvm-svn: 295246
*	AMDGPU: Consolidate sendmsg/sendmsghalt handling and tests	Matt Arsenault	2017-02-15	1	-7/+4
\| \| \| \|	llvm-svn: 295244