bcm5719-llvm - Project Ortega BCM5719 LLVM

	Commit message (Collapse)	Author	Age	Files	Lines
*	DebugInfo: Support debug_loc under fission	David Blaikie	2014-03-25	11	-22/+172
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Implement debug_loc.dwo, as well as llvm-dwarfdump support for dumping this section. Outlined in the DWARF5 spec and http://gcc.gnu.org/wiki/DebugFission the debug_loc.dwo section has more variation than the standard debug_loc, allowing 3 different forms of entry (plus the end of list entry). GCC seems to, and Clang certainly, only use one form, so I've just implemented dumping support for that for now. It wasn't immediately obvious that there was a good refactoring to share the implementation of dumping support between debug_loc and debug_loc.dwo, so they're separate for now - ideas welcome or I may come back to it at some point. As per a comment in the code, we could choose different forms that may reduce the number of debug_addr entries we emit, but that will require further study. llvm-svn: 204697
*	DebugInfo: Remove unnecessary zero-size check	David Blaikie	2014-03-25	1	-3/+0
\| \| \| \| \| \| \| \| \|	This seems excessive - switching section isn't expensive (or if it is we're already being wasteful, since we emitted the debug_loc section symbol earlier anyway) and otherwise there's no work that happens in this function when the list is empty. llvm-svn: 204696
*	Support: Functions for consuming endian specific data from a buffer.	Justin Bogner	2014-03-25	1	-0/+9
\| \| \| \| \| \| \| \|	This adds a function to Endian.h that reads from and updates a pointer into a buffer with endian specific data. This is more convenient for stream-like reading of data than endian::read. llvm-svn: 204693
*	Register Allocator: check other options before using a CSR for the first time.	Manman Ren	2014-03-25	2	-6/+354
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	When register allocator's stage is RS_Spill, we choose spill over using the CSR for the first time, if the spill cost is lower than CSRCost. When register allocator's stage is < RS_Split, we choose pre-splitting over using the CSR for the first time, if the cost of splitting is lower than CSRCost. CSRCost is set with command-line option "regalloc-csr-first-time-cost". The default value is 0 to generate the same codes as before this commit. With a value of 15 (1 << 14 is the entry frequency), I measured performance gain of 3% on 253.perlbmk and 1.7% on 197.parser, with instrumented PGO, on an arm device. rdar://16162005 llvm-svn: 204690
*	Fix crashes when assembler directives are used that are not	Kevin Enderby	2014-03-25	2	-1/+88
\| \| \| \| \| \| \| \|	for Mach-O object files by generating an error instead. rdar://16335232 llvm-svn: 204687
*	Register Allocator: refactoring (no functionality change).	Manman Ren	2014-03-24	1	-6/+30
\| \| \| \| \| \| \| \| \|	Factor out two functions calculateRegionSplitCost and doRegionSplit from tryRegionSplit. These two functions will be used in coming patches. rdar://16162005 llvm-svn: 204684
*	DebugInfo: Simplify debug loc list handling by keeping separate lists	David Blaikie	2014-03-24	3	-33/+17
\| \| \| \| \| \| \|	Rather than using a flat list with "empty" entries (ala the actual on-disk format), keep separate lists for each variable. llvm-svn: 204680
*	DwarfDebug: Simplify debug_loc merging	David Blaikie	2014-03-24	3	-29/+13
\| \| \| \| \| \| \| \| \| \|	No functional change intended. Merging up-front rather than delaying this task until later. This just seems simpler and more efficient (avoiding growing the debug loc list only to have to skip over those post-merged entries, etc). llvm-svn: 204679
*	Get rid of an unnecessary use of the * and & operators.	Adrian Prantl	2014-03-24	1	-1/+1
\| \| \| \|	llvm-svn: 204673
*	DebugInfo: Add DW_AT_GNU_ranges_base to skeleton CUs	David Blaikie	2014-03-24	2	-9/+16
\| \| \| \| \| \| \| \| \|	This is used to avoid relocations in the dwo file by allowing DW_AT_ranges specified in debug_info.dwo to be relative to this base address. (r204667 implements the base-relative DW_AT_ranges side of this) llvm-svn: 204672
*	Support: Document Endian.h functions	Justin Bogner	2014-03-24	1	-0/+3
\| \| \| \|	llvm-svn: 204671
*	DebugInfo: Implement relative addressing for DW_AT_ranges under fission	David Blaikie	2014-03-24	2	-5/+19
\| \| \| \| \| \| \| \|	This removes the debug_ranges relocations from debug_info.dwo (but doesn't implement the DW_AT_GNU_ranges_base which is also necessary for correct functioning) llvm-svn: 204668
*	DebugInfo: Don't emit relocations to abbreviations in debug_info.dwo	David Blaikie	2014-03-24	3	-2/+7
\| \| \| \|	llvm-svn: 204667
*	DwarfDebug: Remove an unused parameter	David Blaikie	2014-03-24	2	-9/+4
\| \| \| \|	llvm-svn: 204665
*	R600: Don't viewCFG() under DEBUG() except on failure.	Matt Arsenault	2014-03-24	1	-9/+6
\| \| \| \| \| \| \|	Having these popping up every time you use -debug is really irritating. llvm-svn: 204664
*	Remove unused parameter	David Blaikie	2014-03-24	3	-10/+6
\| \| \| \|	llvm-svn: 204663
*	R600/SI: Fix extra mov from legalizing 64-bit SALU ops.	Matt Arsenault	2014-03-24	2	-19/+31
\| \| \| \| \| \| \|	Check the register class of each operand individually to avoid an extra copy to a vgpr. llvm-svn: 204662
*	R600/SI: Sub-optimial fix for 64-bit immediates with SALU ops.	Matt Arsenault	2014-03-24	3	-50/+104
\| \| \| \| \| \| \|	No longer asserts, but now you get moves loading legal immediates into the split 32-bit operations. llvm-svn: 204661
*	R600/SI: Fix 64-bit bit ops that require the VALU.	Matt Arsenault	2014-03-24	4	-8/+108
\| \| \| \| \| \| \| \|	Try to match scalar and first like the other instructions. Expand 64-bit ands to a pair of 32-bit ands since that is not available on the VALU. llvm-svn: 204660
*	In Release modes, Visual Studio complains that the Operator destructor in ↵	Yaron Keren	2014-03-24	1	-0/+20
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	User.cpp never returns, which is true by design. Initially assumed that the reason is llvm_unreachable being dependent on NDEBUG. However, even if llvm_unreachable is replaced by __assume(false), VC still warns in Release modes but not in Debug modes... The real reason turned out to be optimization flags. With /Od in Debug modes the warning is not issued whereas with /O1 it is. I could not find any documentation to this effect, but it is reproducable: Try compiling http://msdn.microsoft.com/en-us/library/khwfyc5d(v=vs.90).aspx with /O1 and then with /Od. llvm-svn: 204659
*	R600: Implement isNarrowingProfitable.	Matt Arsenault	2014-03-24	3	-5/+30
\| \| \| \|	llvm-svn: 204658
*	R600/SI: Move splitting 64-bit immediates to separate function.	Matt Arsenault	2014-03-24	2	-39/+59
\| \| \| \|	llvm-svn: 204651
*	Adding some very nascent information about the clang tablegen backends, with ↵	Aaron Ballman	2014-03-24	1	-17/+49
\| \| \| \| \| \|	a promise to add more information later. llvm-svn: 204635
*	[PowerPC] Generate little-endian object files	Ulrich Weigand	2014-03-24	24	-4301/+6629
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	As a first step towards real little-endian code generation, this patch changes the PowerPC MC layer to actually generate little-endian object files. This involves passing the little-endian flag through the various layers, including down to createELFObjectWriter so we actually get basic little-endian ELF objects, emitting instructions in little-endian order, and handling fixups and relocations as appropriate for little-endian. The bulk of the patch is to update most test cases in test/MC/PowerPC to verify both big- and little-endian encodings. (The only test cases not updated are those that create actual big-endian ABI code, like the TLS tests.) Note that while the object files are now little-endian, the generated code itself is not yet updated, in particular, it still does not adhere to the ELFv2 ABI. llvm-svn: 204634
*	[X86][ISelDAG] Add missing fallback patterns for avx2 broadcast instructions.	Quentin Colombet	2014-03-24	2	-0/+183
\| \| \| \| \| \| \| \| \| \| \| \| \|	Those patterns are used when the load cannot be folded into the related broadcast during the select phase. This happens when the load gets additional uses that were not anticipated during the previous lowering phases (constant vector to constant load, then constant load reused) or when selection DAG is not able to prove that folding the load will not create a cycle in the DAG. <rdar://problem/16074331> llvm-svn: 204631
*	R600/SI: Fix 64-bit private loads.	Matt Arsenault	2014-03-24	2	-9/+69
\| \| \| \|	llvm-svn: 204630
*	VS integration installer: set SUCCESS=1 if we find VS 2013	Hans Wennborg	2014-03-24	1	-2/+3
\| \| \| \| \| \| \| \| \| \|	Previously we would print an error message on machines where the only VS version we find is 2013, even though we successfully install the integration files for it. Also, we shouldn't have two END labels. llvm-svn: 204629
*	Add test to test/CodeGen/NVPTX for "alloca buffer" arguments.	Eli Bendersky	2014-03-24	1	-0/+66
\| \| \| \| \| \|	Make sure such IR gets properly lowered to PTX. llvm-svn: 204624
*	[X86] Fix non-determinism in LowerVectorAllZeroTest	Adam Nemet	2014-03-24	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This can be observed with the old testcase of CodeGen/X86/pr12312.ll: 47c47 < vorps %ymm0, %ymm1, %ymm0 --- > vorps %ymm1, %ymm0, %ymm0 97c97 < vorps %ymm1, %ymm0, %ymm0 --- > vorps %ymm0, %ymm1, %ymm0 The vector VecIns is populated with all the values from VecInMap. This is done while iterating VecInMap. VecInMap uses a hash of pointer values so the resulting order can vary depending on the memory layout. The fix is to populate the vector VecIns earlier as VecInMap is populated. This is done in DAG traversal order. Fixes <rdar://problem/16398806> llvm-svn: 204623
*	[mips] Add error message when trying to use $at in '.set noat' mode.	Daniel Sanders	2014-03-24	2	-1/+33
\| \| \| \| \| \| \| \| \| \|	Summary: Patch by David Chisnall His work was sponsored by: DARPA, AFRL Differential Revision: http://llvm-reviews.chandlerc.com/D3158 llvm-svn: 204621
*	Removes the NVPTXSplitBBatBar pass.	Eli Bendersky	2014-03-24	4	-118/+0
\| \| \| \| \| \| \|	This pass is a historic remnant and actually causes less efficient code to be generated in some cases. llvm-svn: 204620
*	R600/SI: Fix warning with gcc 4.8.2	Tom Stellard	2014-03-24	1	-1/+1
\| \| \| \|	llvm-svn: 204618
*	R600/SI: Promote fp64 SELECT to i64	Tom Stellard	2014-03-24	2	-12/+2
\| \| \| \| \| \| \|	This type promotion is replacing a Tablegen pattern and it is already covered by existing tests. llvm-svn: 204617
*	SelectionDAG: Allow promotion of SELECT nodes from float to int types	Tom Stellard	2014-03-24	1	-1/+2
\| \| \| \| \| \| \| \|	And vice-versa, as long as the types are the same width. There are a few R600 tests that will cover this. llvm-svn: 204616
*	R600: Reorganize tablegen instruction definitions	Tom Stellard	2014-03-24	5	-781/+826
\| \| \| \| \| \|	Each GPU family now has its own file. llvm-svn: 204615
*	[PPC64LE] ELFv2 ABI updates for the .opd section	Will Schmidt	2014-03-24	1	-0/+5
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	[PPC64LE] ELFv2 ABI updates for the .opd section The PPC64 Little Endian (PPC64LE) target supports the ELFv2 ABI, and as such, does not have a ".opd" section. This is keyed off a _CALL_ELF=2 macro check. The CALL_ELF check is not clearly documented at this time. The basis for usage in this patch is from the gcc thread here: http://gcc.gnu.org/ml/gcc-patches/2013-11/msg01144.html > Adding comment from Uli: Looks good to me. I think the old-style JIT doesn't really work anyway for 64-bit, but at least with this patch LLVM will compile and link again on a ppc64le host ... llvm-svn: 204614
*	[mips] Add regression tests for parenthetic expressions in MIPS assembly.	Daniel Sanders	2014-03-24	1	-0/+12
\| \| \| \| \| \| \| \| \| \| \| \|	Summary: These expressions already worked but weren't tested. Patch by Robert N. M. Watson and David Chisnall (it was originally two patches) Their work was sponsored by: DARPA, AFRL Differential Revision: http://llvm-reviews.chandlerc.com/D3156 llvm-svn: 204612
*	[mips] Allow dsubu to take an immediate as an alias for dsubiu.	Daniel Sanders	2014-03-24	2	-0/+5
\| \| \| \| \| \| \| \| \| \|	Summary: Patch by David Chisnall His work was sponsored by: DARPA, AFRL Differential Revision: http://llvm-reviews.chandlerc.com/D3155 llvm-svn: 204611
*	[PowerPC] Mark many instructions as commutative	Hal Finkel	2014-03-24	4	-4/+63
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	I'm under the impression that we used to infer the isCommutable flag from the instruction-associated pattern. Regardless, we don't seem to do this (at least by default) any more. I've gone through all of our instruction definitions, and marked as commutative all of those that should be trivial to commute (by exchanging the first two operands). There has been special code for the RL* instructions, and that's not changed. Before this change, we had the following commutative instructions: RLDIMI RLDIMIo RLWIMI RLWIMI8 RLWIMI8o RLWIMIo XSADDDP XSMULDP XVADDDP XVADDSP XVMULDP XVMULSP After: ADD4 ADD4o ADD8 ADD8o ADDC ADDC8 ADDC8o ADDCo ADDE ADDE8 ADDE8o ADDEo AND AND8 AND8o ANDo CRAND CREQV CRNAND CRNOR CROR CRXOR EQV EQV8 EQV8o EQVo FADD FADDS FADDSo FADDo FMADD FMADDS FMADDSo FMADDo FMSUB FMSUBS FMSUBSo FMSUBo FMUL FMULS FMULSo FMULo FNMADD FNMADDS FNMADDSo FNMADDo FNMSUB FNMSUBS FNMSUBSo FNMSUBo MULHD MULHDU MULHDUo MULHDo MULHW MULHWU MULHWUo MULHWo MULLD MULLDo MULLW MULLWo NAND NAND8 NAND8o NANDo NOR NOR8 NOR8o NORo OR OR8 OR8o ORo RLDIMI RLDIMIo RLWIMI RLWIMI8 RLWIMI8o RLWIMIo VADDCUW VADDFP VADDSBS VADDSHS VADDSWS VADDUBM VADDUBS VADDUHM VADDUHS VADDUWM VADDUWS VAND VAVGSB VAVGSH VAVGSW VAVGUB VAVGUH VAVGUW VMADDFP VMAXFP VMAXSB VMAXSH VMAXSW VMAXUB VMAXUH VMAXUW VMHADDSHS VMHRADDSHS VMINFP VMINSB VMINSH VMINSW VMINUB VMINUH VMINUW VMLADDUHM VMULESB VMULESH VMULEUB VMULEUH VMULOSB VMULOSH VMULOUB VMULOUH VNMSUBFP VOR VXOR XOR XOR8 XOR8o XORo XSADDDP XSMADDADP XSMAXDP XSMINDP XSMSUBADP XSMULDP XSNMADDADP XSNMSUBADP XVADDDP XVADDSP XVMADDADP XVMADDASP XVMAXDP XVMAXSP XVMINDP XVMINSP XVMSUBADP XVMSUBASP XVMULDP XVMULSP XVNMADDADP XVNMADDASP XVNMSUBADP XVNMSUBASP XXLAND XXLNOR XXLOR XXLXOR This is a by-inspection change, and I'm not sure how to write a reliable test case. I would like advice on this, however. llvm-svn: 204609
*	[mips] Implement shorthand add / sub forms for MIPS.	Daniel Sanders	2014-03-24	5	-1/+138
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Summary: - If only two registers are passed to a three-register operation, then the first argument is both source and destination register. - If a non-register is passed as the last argument, generate the immediate version of the instruction. Also mark DADD commutative and add scheduling information (to the generic scheduler), and implement DSUB. Patch by David Chisnall His work was sponsored by: DARPA, AFRL CC: theraven Differential Revision: http://llvm-reviews.chandlerc.com/D3148 llvm-svn: 204605
*	[NVPTX] Add isel patterns for addrspacecast	Justin Holewinski	2014-03-24	3	-0/+163
\| \| \| \|	llvm-svn: 204600
*	Update release notes with EHABI current behaviour	Renato Golin	2014-03-24	1	-2/+1
\| \| \| \|	llvm-svn: 204598
*	[PowerPC] Don't schedule VSX copy legalization unless VSX is enabled	Hal Finkel	2014-03-24	1	-1/+2
\| \| \| \| \| \|	There is no need to schedule this extra pass if it will have nothing to do. llvm-svn: 204594
*	[PowerPC] Update comment re: VSX copy-instruction selection	Hal Finkel	2014-03-24	1	-2/+4
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	I've done some experimentation with this, and it looks like using the lower-latency (but lower throughput) copy instruction is essentially always the right thing to do. My assumption is that, in order to be relatively sure that the higher-latency copy will increase throughput, we'd want to have it unlikely to be in-flight with its use. On the P7, the global completion table (GCT) can hold a maximum of 120 instructions, shared among all active threads (up to 4), giving 30 instructions per thread. So specifically, I'd require at least that many instructions between the copy and the use before the high-latency variant is used. Trying this, however, over the entire test suite resulted in zero cases where the high-latency form would be preferable. This may be a consequence of the fact that the scheduler views copies as free, and so they tend to end up close to their uses. For this experiment I created a function: unsigned chooseVSXCopy(MachineBasicBlock &MBB, MachineBasicBlock::iterator I, unsigned DestReg, unsigned SrcReg, unsigned StartDist = 1, unsigned Depth = 3) const; with an implementation like: if (!Depth) return PPC::XXLOR; const unsigned MaxDist = 30; unsigned Dist = StartDist; for (auto J = I, JE = MBB.end(); J != JE && Dist <= MaxDist; ++J) { if (J->isTransient() && !J->isCopy()) continue; if (J->isCall() \|\| J->isReturn() \|\| J->readsRegister(DestReg, TRI)) return PPC::XXLOR; ++Dist; } // We've exceeded the required distance for the high-latency form, use it. if (Dist > MaxDist) return PPC::XVCPSGNDP; // If this is only an exit block, use the low-latency form. if (MBB.succ_empty()) return PPC::XXLOR; // We've reached the end of the block, check the successor blocks (up to some // depth), and use the high-latency form if that is okay with all successors. for (auto J = MBB.succ_begin(), JE = MBB.succ_end(); J != JE; ++J) { if (chooseVSXCopy(*J, (J)->begin(), DestReg, SrcReg, Dist, --Depth) == PPC::XXLOR) return PPC::XXLOR; } // All of our successor blocks seem okay with the high-latency variant, so // we'll use it. return PPC::XVCPSGNDP; and then changed the copy opcode selection from: Opc = PPC::XXLOR; to: Opc = chooseVSXCopy(MBB, std::next(I), DestReg, SrcReg); In conclusion, I'm removing the FIXME from the comment, because I believe that there is, at least absent other examples, nothing to fix. llvm-svn: 204591
*	Teach llvm-readobj to print human friendly description of reserved sections.	Rafael Espindola	2014-03-24	24	-68/+88
\| \| \| \|	llvm-svn: 204584
*	Allow constant folding of ceil function whenever feasible	Karthik Bhat	2014-03-24	2	-0/+59
\| \| \| \|	llvm-svn: 204583
*	Add back tests that were reverted in r204203.	Rafael Espindola	2014-03-24	1	-11/+56
\| \| \| \| \| \|	They pass again with the fix in r204581. llvm-svn: 204582
*	Propagate section from base to derived symbol.	Rafael Espindola	2014-03-24	3	-27/+22
\| \| \| \| \| \| \| \| \| \| \| \|	We were already propagating the section in a = b With this patch we also propagate it for a = b + 1 llvm-svn: 204581
*	InstrProf: Silence spurious warnings in GCC 4.8	Duncan P. N. Exon Smith	2014-03-24	1	-9/+13
\| \| \| \| \| \|	No functionality change. llvm-svn: 204580
*	SupportTests.LockFileManagerTest: Add assertions for Win32.	NAKAMURA Takumi	2014-03-23	1	-2/+16
\| \| \| \| \| \| \|	- create_link doesn't work for nonexistent file. - remove cannot remove working directory. llvm-svn: 204579