bcm5719-llvm - Project Ortega BCM5719 LLVM

	Commit message (Collapse)	Author	Age	Files	Lines
*	Access the TargetLoweringInfo from the TargetMachine object instead of ↵	Bill Wendling	2013-06-19	1	-11/+9
\| \| \| \| \| \|	caching it. The TLI may change between functions. No functionality change. llvm-svn: 184352
*	Second part of pr16069	Rafael Espindola	2013-06-04	1	-4/+9
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The problem this time seems to be a thinko. We were assuming that in the CFG A \| \ \| B \| / C speculating the basic block B would cause only the phi value for the B->C edge to be speculated. That is not true, the phi's are semantically in the edges, so if the A->B->C path is taken, any code needed for A->C is not executed and we have to consider it too when deciding to speculate B. llvm-svn: 183226
*	Typo: s/caes/cases/ in SimplifyCFG	Hans Wennborg	2013-06-04	1	-1/+1
\| \| \| \|	llvm-svn: 183219
*	SimplifyCFG: Do not transform PHI to select if doing so would be unsafe	David Majnemer	2013-06-03	1	-2/+19
\| \| \| \| \| \| \| \| \| \| \| \| \|	PR16069 is an interesting case where an incoming value to a PHI is a trap value while also being a 'ConstantExpr'. We do not consider this case when performing the 'HoistThenElseCodeToIf' optimization. Instead, make our modifications more conservative if we detect that we cannot transform the PHI to a select. llvm-svn: 183152
*	SimplifyCFG: Small cleanup, use ICmpInst::isEquality()	David Majnemer	2013-06-03	1	-3/+1
\| \| \| \|	llvm-svn: 183151
*	SimplifyCFG: Fix typo in comment for ComputeSpeculationCost	David Majnemer	2013-06-01	1	-1/+1
\| \| \| \|	llvm-svn: 183078
*	Extend RemapInstruction and friends to take an optional new parameter, a ↵	James Molloy	2013-05-28	2	-12/+22
\| \| \| \| \| \| \| \|	ValueMaterializer. Extend LinkModules to pass a ValueMaterializer to RemapInstruction and friends to lazily create Functions for lazily linked globals. This is a big win when linking small modules with large (mostly unused) library modules. llvm-svn: 182776
*	More symbols that should be static.	Benjamin Kramer	2013-05-23	1	-2/+2
\| \| \| \|	llvm-svn: 182590
*	Rename LoopSimplify.h to LoopUtils.h	Hal Finkel	2013-05-20	1	-1/+1
\| \| \| \| \| \|	As discussed, LoopUtils.h is a better name. llvm-svn: 182314
*	Expose InsertPreheaderForLoop from LoopSimplify to other passes	Hal Finkel	2013-05-20	1	-11/+12
\| \| \| \| \| \| \| \| \| \| \|	Other passes, PPC counter-loop formation for example, also need to add loop preheaders outside of the regular loop simplification pass. This makes InsertPreheaderForLoop a global function so that it can be used by other passes. No functionality change intended. llvm-svn: 182299
*	Add ArrayRef constructor from None, and do the cleanups that this ↵	Dmitri Gribenko	2013-05-05	1	-1/+1
\| \| \| \| \| \| \| \|	constructor enables Patch by Robert Wilhelm. llvm-svn: 181138
*	This patch breaks up Wrap.h so that it does not have to include all of	Filip Pizlo	2013-05-01	1	-1/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	the things, and renames it to CBindingWrapping.h. I also moved CBindingWrapping.h into Support/. This new file just contains the macros for defining different wrap/unwrap methods. The calls to those macros, as well as any custom wrap/unwrap definitions (like for array of Values for example), are put into corresponding C++ headers. Doing this required some #include surgery, since some .cpp files relied on the fact that including Wrap.h implicitly caused the inclusion of a bunch of other things. This also now means that the C++ headers will include their corresponding C API headers; for example Value.h must include llvm-c/Core.h. I think this is harmless, since the C API headers contain just external function declarations and some C types, so I don't believe there should be any nasty dependency issues here. llvm-svn: 180881
*	Fix a use after free. RI is freed before the call to getDebugLoc(). To	Richard Trieu	2013-04-30	1	-4/+5
\| \| \| \| \| \|	prevent this, capture the location before RI is freed. llvm-svn: 180824
*	Spelling. Thanks, Eric.	Adrian Prantl	2013-04-30	1	-1/+1
\| \| \| \|	llvm-svn: 180794
*	Set debug locations for branch instructions created during inlining, even	Adrian Prantl	2013-04-30	1	-2/+10
\| \| \| \| \| \| \| \|	the inlined function has multiple returns. rdar://problem/12415623 llvm-svn: 180793
*	SimplifyCFG: If convert single conditional stores	Arnold Schwaighofer	2013-04-29	1	-4/+90
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This resurrects r179957, but adds code that makes sure we don't touch atomic/volatile stores: This transformation will transform a conditional store with a preceeding uncondtional store to the same location: a[i] = may-alias with a[i] load if (cond) a[i] = Y into an unconditional store. a[i] = X may-alias with a[i] load tmp = cond ? Y : X; a[i] = tmp We assume that on average the cost of a mispredicted branch is going to be higher than the cost of a second store to the same location, and that the secondary benefits of creating a bigger basic block for other optimizations to work on outway the potential case where the branch would be correctly predicted and the cost of the executing the second store would be noticably reflected in performance. hmmer's execution time improves by 30% on an imac12,2 on ref data sets. With this change we are on par with gcc's performance (gcc also performs this transformation). There was a 1.2 % performance improvement on a ARM swift chip. Other tests in the test-suite+external seem to be mostly uninfluenced in my experiments: This optimization was triggered on 41 tests such that the executable was different before/after the patch. Only 1 out of the 40 tests (dealII) was reproducable below 100% (by about .4%). Given that hmmer benefits so much I believe this to be a fair trade off. llvm-svn: 180731
*	fix a typo that due to cu&paste quadrupled itself	Adrian Prantl	2013-04-26	1	-2/+2
\| \| \| \| \| \|	rdar://problem/13056109 llvm-svn: 180618
*	Bugfix for the debug intrinsic handling in InstCombiner:	Adrian Prantl	2013-04-26	1	-2/+27
\| \| \| \| \| \| \| \| \| \| \|	Since we can't guarantee that the original dbg.declare instrinsic is removed by LowerDbgDeclare(), we need to make sure that we are not inserting the same dbg.value intrinsic over and over. This removes tons of redundant DIEs when compiling optimized code. rdar://problem/13056109 llvm-svn: 180615
*	Make sure the instruction right after an inlined function has a	Adrian Prantl	2013-04-23	1	-4/+10
\| \| \| \| \| \| \| \| \| \|	debug location. This solves a problem where range of an inlined subroutine is emitted wrongly. Patch by Manman Ren. Fixes rdar://problem/12415623 llvm-svn: 180140
*	Move C++ code out of the C headers and into either C++ headers	Eric Christopher	2013-04-22	1	-0/+1
\| \| \| \| \| \| \|	or the C++ files themselves. This enables people to use just a C compiler to interoperate with LLVM. llvm-svn: 180063
*	Revert "SimplifyCFG: If convert single conditional stores"	Arnold Schwaighofer	2013-04-21	1	-88/+4
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	There is the temptation to make this tranform dependent on target information as it is not going to be beneficial on all (sub)targets. Therefore, we should probably do this in MI Early-Ifconversion. This reverts commit r179957. Original commit message: "SimplifyCFG: If convert single conditional stores This transformation will transform a conditional store with a preceeding uncondtional store to the same location: a[i] = may-alias with a[i] load if (cond) a[i] = Y into an unconditional store. a[i] = X may-alias with a[i] load tmp = cond ? Y : X; a[i] = tmp We assume that on average the cost of a mispredicted branch is going to be higher than the cost of a second store to the same location, and that the secondary benefits of creating a bigger basic block for other optimizations to work on outway the potential case were the branch would be correctly predicted and the cost of the executing the second store would be noticably reflected in performance. hmmer's execution time improves by 30% on an imac12,2 on ref data sets. With this change we are on par with gcc's performance (gcc also performs this transformation). There was a 1.2 % performance improvement on a ARM swift chip. Other tests in the test-suite+external seem to be mostly uninfluenced in my experiments: This optimization was triggered on 41 tests such that the executable was different before/after the patch. Only 1 out of the 40 tests (dealII) was reproducable below 100% (by about .4%). Given that hmmer benefits so much I believe this to be a fair trade off. I am going to watch performance numbers across the builtbots and will revert this if anything unexpected comes up." llvm-svn: 179980
*	SimplifyCFG: If convert single conditional stores	Arnold Schwaighofer	2013-04-20	1	-4/+88
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This transformation will transform a conditional store with a preceeding uncondtional store to the same location: a[i] = may-alias with a[i] load if (cond) a[i] = Y into an unconditional store. a[i] = X may-alias with a[i] load tmp = cond ? Y : X; a[i] = tmp We assume that on average the cost of a mispredicted branch is going to be higher than the cost of a second store to the same location, and that the secondary benefits of creating a bigger basic block for other optimizations to work on outway the potential case were the branch would be correctly predicted and the cost of the executing the second store would be noticably reflected in performance. hmmer's execution time improves by 30% on an imac12,2 on ref data sets. With this change we are on par with gcc's performance (gcc also performs this transformation). There was a 1.2 % performance improvement on a ARM swift chip. Other tests in the test-suite+external seem to be mostly uninfluenced in my experiments: This optimization was triggered on 41 tests such that the executable was different before/after the patch. Only 1 out of the 40 tests (dealII) was reproducable below 100% (by about .4%). Given that hmmer benefits so much I believe this to be a fair trade off. I am going to watch performance numbers across the builtbots and will revert this if anything unexpected comes up. llvm-svn: 179957
*	Do not optimise fprintf() calls if its return value is used.	Peter Collingbourne	2013-04-17	1	-9/+12
\| \| \| \| \| \|	Differential Revision: http://llvm-reviews.chandlerc.com/D620 llvm-svn: 179661
*	simplifycfg: Fix integer overflow converting switch into icmp.	Hans Wennborg	2013-04-16	1	-1/+6
\| \| \| \| \| \| \| \| \| \| \|	If a switch instruction has a case for every possible value of its type, with the same successor, SimplifyCFG would replace it with an icmp ult, but the computation of the bound overflows in that case, which inverts the test. Patch by Jed Davis! llvm-svn: 179587
*	Change CloneFunctionInto to always clone Argument attributes induvidually,	Joey Gouly	2013-04-10	1	-22/+19
\| \| \| \| \| \| \|	rather than checking if the source and destination have the same number of arguments and copying the attributes over directly. llvm-svn: 179169
*	Add all clauses when merging the landing pads. Duplicates will be handled ↵	Bill Wendling	2013-03-22	1	-24/+14
\| \| \| \| \| \|	later on. llvm-svn: 177757
*	Don't use the removed API.	Bill Wendling	2013-03-22	1	-5/+2
\| \| \| \|	llvm-svn: 177749
*	Fix llvm::removeUnreachableBlocks to handle unreachable loops.	Evgeniy Stepanov	2013-03-22	1	-12/+7
\| \| \| \|	llvm-svn: 177713
*	Always forward 'resume' instructions to the outter landing pad.	Bill Wendling	2013-03-21	1	-16/+39
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	How did this ever work? Basically, if you have a function that's inlined into the caller, it may not have any 'call' instructions, but any 'resume' instructions it may have should still be forwarded to the outer (caller's) landing pad. This requires that all of the 'landingpad' instructions in the callee have their clauses merged with the caller's outer 'landingpad' instruction (hence the bit of ugly code in the `forwardResume' method). Testcase in a follow commit to the test-suite repository. <rdar://problem/13360379> & PR15555 llvm-svn: 177680
*	LibCallSimplifier: optimize speed for short-lived instances	Meador Inge	2013-03-12	1	-177/+225
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Nadav reported a performance regression due to the work I did to merge the library call simplifier into instcombine [1]. The issue is that a new LibCallSimplifier object is being created whenever InstCombiner::runOnFunction is called. Every time a LibCallSimplifier object is used to optimize a call it creates a hash table to map from a function name to an object that optimizes functions of that name. For short-lived LibCallSimplifier instances this is quite inefficient. Especially for cases where no calls are actually simplified. This patch fixes the issue by dropping the hash table and implementing an explicit lookup function to correlate the function name to the object that optimizes functions of that name. This avoids the cost of always building and destroying the hash table in cases where the LibCallSimplifier object is short-lived and avoids the cost of building the table when no simplifications are actually preformed. On a benchmark containing 100,000 calls where none of them are simplified I noticed a 30% speedup. On a benchmark containing 100,000 calls where all of them are simplified I noticed an 8% speedup. [1] http://lists.cs.uiuc.edu/pipermail/llvm-commits/Week-of-Mon-20130304/167639.html llvm-svn: 176840
*	Don't remove a landing pad if the invoke requires a table entry.	Bill Wendling	2013-03-11	1	-3/+17
\| \| \| \| \| \| \| \|	An invoke may require a table entry. For instance, when the function it calls is expected to throw. <rdar://problem/13360379> llvm-svn: 176827
*	Fixed a crash when cloning a function into a function with	Pekka Jaaskelainen	2013-03-07	1	-3/+6
\| \| \| \| \| \| \|	different size argument list and without attributes in the arguments. llvm-svn: 176632
*	SimplifyCFG fix for volatile load/store.	Andrew Trick	2013-03-07	1	-2/+4
\| \| \| \| \| \| \| \| \| \| \| \| \|	Fixes rdar:13349374. Volatile loads and stores need to be preserved even if the language standard says they are undefined. "volatile" in this context means "get out of the way compiler, let my platform handle it". Additionally, this is the only way I know of with llvm to write to the first page (when hardware allows) without dropping to assembly. llvm-svn: 176599
*	Bypass Slow Divides	Preston Gurd	2013-03-04	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \| \| \|	* Only apply divide bypass optimization when not optimizing for size. * Fixed bug caused by constant for 0 value of type Int32, used dividend type to generate the constant instead. * For atom x86-64 apply the divide bypass to use 16-bit divides instead of 64-bit divides when operand values are small enough. * Added lit tests for 64-bit divide bypass. Patch by Tyler Nowicki! llvm-svn: 176442
*	Modify {Call,Invoke}Inst::addAttribute to take an AttrKind.	Peter Collingbourne	2013-03-02	1	-2/+1
\| \| \| \|	llvm-svn: 176397
*	For each function that we optimize we initialize a new list of lib ↵	Nadav Rotem	2013-02-27	1	-1/+2
\| \| \| \| \| \|	functions. For each function name we malloc memory. This patch changes the Libcall map to use BumpPtrAllocator. Now we malloc only once. This speeds up instcombine by a few % on a large c++ program. llvm-svn: 176170
*	Enhance integer division emulation support to handle types smaller than 32 bits,	Pedro Artigas	2013-02-26	1	-0/+104
\| \| \| \| \| \| \| \|	enhancement done the trivial way; by extending inputs and truncating outputs which is addequate for targets with little or no support for integer arithmetic on integer types less than 32 bits. llvm-svn: 176139
*	Implement the NoBuiltin attribute.	Bill Wendling	2013-02-22	1	-0/+1
\| \| \| \| \| \| \| \|	The 'nobuiltin' attribute is applied to call sites to indicate that LLVM should not treat the callee function as a built-in function. I.e., it shouldn't try to replace that function with different code. llvm-svn: 175835
*	Temporarily revert r175470 for more review.	Bill Wendling	2013-02-19	1	-3/+0
\| \| \| \|	llvm-svn: 175476
*	Check to see if the 'no-builtin' attribute is set before simplifying a ↵	Bill Wendling	2013-02-18	1	-0/+3
\| \| \| \| \| \|	library call. llvm-svn: 175470
*	Remove #includes from the commonly used LoopInfo.h.	Jakub Staszak	2013-02-09	1	-0/+1
\| \| \| \|	llvm-svn: 174786
*	[SimplifyLibCalls] Library call simplification doen't work if the call site	Chad Rosier	2013-02-08	1	-1/+7
\| \| \| \| \| \| \| \|	isn't using the default calling convention. However, if the transformation is from a call to inline IR, then the calling convention doesn't matter. rdar://13157990 llvm-svn: 174724
*	[SjLj Prepare] When demoting an invoke instructions to the stack, if the normal	Chad Rosier	2013-02-05	1	-5/+15
\| \| \| \| \| \| \|	edge is critical, then split it so we can insert the store. rdar://13126179 llvm-svn: 174418
*	Linker: correctly link in dbg.declare	Manman Ren	2013-01-31	1	-2/+17
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This is a re-worked version of r174048. Given source IR: call void @llvm.dbg.declare(metadata !{i32* %argc.addr}, metadata !14), !dbg !15 we used to generate call void @llvm.dbg.declare(metadata !27, metadata !28), !dbg !29 !27 = metadata !{null} With this patch, we will correctly generate call void @llvm.dbg.declare(metadata !{i32* %argc.addr}, metadata !27), !dbg !28 Looking up %argc.addr in ValueMap will return null, since %argc.addr is already correctly set up, we can use identity mapping. rdar://problem/13089880 llvm-svn: 174093
*	Revert r173946. This breaks compilation of googletest with Clang	Alexey Samsonov	2013-01-31	1	-11/+2
\| \| \| \|	llvm-svn: 174048
*	Remove addRetAttributes and addFnAttributes, which aren't useful abstractions.	Bill Wendling	2013-01-30	1	-4/+6
\| \| \| \|	llvm-svn: 173992
*	Linker: correctly link in dbg.declare	Manman Ren	2013-01-30	1	-2/+11
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Given source IR: call void @llvm.dbg.declare(metadata !{i32* %argc.addr}, metadata !14), !dbg !15 we used to generate call void @llvm.dbg.declare(metadata !27, metadata !28), !dbg !29 !27 = metadata !{null} With this patch, we will correctly generate call void @llvm.dbg.declare(metadata !{i32* %argc.addr}, metadata !27), !dbg !28 Looking up %argc.addr in ValueMap will return null, since %argc.addr is already correctly set up, we can use identity mapping. llvm-svn: 173946
*	Re-revert r173342, without losing the compile time improvements, flat	Chandler Carruth	2013-01-27	1	-27/+12
\| \| \| \| \| \|	out bug fixes, or functionality preserving refactorings. llvm-svn: 173610
*	Convert BuildLibCalls.cpp to using the AttributeSet methods instead of ↵	Bill Wendling	2013-01-26	1	-66/+66
\| \| \| \| \| \|	AttributeWithIndex. llvm-svn: 173536
*	Switch this code away from Value::isUsedInBasicBlock. That code either	Chandler Carruth	2013-01-25	1	-7/+29
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	loops over instructions in the basic block or the use-def list of the value, neither of which are really efficient when repeatedly querying about values in the same basic block. What's more, we already know that the CondBB is small, and so we can do a much more efficient test by counting the uses in CondBB, and seeing if those account for all of the uses. Finally, we shouldn't blanket fail on any such instruction, instead we should conservatively assume that those instructions are part of the cost. Note that this actually fixes a bug in the pass because isUsedInBasicBlock has a really terrible bug in it. I'll fix that in my next commit, but the fix for it would make this code suddenly take the compile time hit I thought it already was taking, so I wanted to go ahead and migrate this code to a faster & better pattern. The bug in isUsedInBasicBlock was also causing other tests to test the wrong thing entirely: for example we weren't actually disabling speculation for floating point operations as intended (and tested), but the test passed because we failed to speculate them due to the isUsedInBasicBlock failure. llvm-svn: 173417