external_llvm.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	Implement NEON domain switching for scalar <-> S-register vmovs on ARM	Tim Northover	2012-08-17	1	-32/+32
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@162094 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add ADD and SUB to the predicable ARM instructions.	Jakob Stoklund Olesen	2012-08-16	3	-21/+35
\| \| \| \| \| \| \| \| \| \|	It is not my plan to duplicate the entire ARM instruction set with predicated versions. We need a way of representing predicated instructions in SSA form without requiring a separate opcode. Then the pseudo-instructions can go away. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@162061 91177308-0d34-0410-b5e6-96231b3b80d8
*	[arm-fast-isel] Add support for fastcc.	Jush Lu	2012-08-16	1	-0/+66
\| \| \| \| \| \| \| \| \|	Without fastcc support, the caller just falls through to CallingConv::C for fastcc, but callee still uses fastcc, this inconsistency of calling convention is a problem, and fastcc support can fix it. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@162013 91177308-0d34-0410-b5e6-96231b3b80d8
*	Test case for r162008.	Akira Hatanaka	2012-08-16	1	-0/+12
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@162009 91177308-0d34-0410-b5e6-96231b3b80d8
*	Fold predicable instructions into MOVCC / t2MOVCC.	Jakob Stoklund Olesen	2012-08-15	2	-1/+61
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	The ARM select instructions are just predicated moves. If the select is the only use of an operand, the instruction defining the operand can be predicated instead, saving one instruction and decreasing register pressure. This implementation can turn AND/ORR/EOR instructions into their corresponding ANDCC/ORRCC/EORCC variants. Ideally, we should be able to predicate any instruction, but we don't yet support predicated instructions in SSA form. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161994 91177308-0d34-0410-b5e6-96231b3b80d8
*	Rework test so that it reproduces the error without the horrible flag.	Bill Wendling	2012-08-15	1	-8/+2
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161989 91177308-0d34-0410-b5e6-96231b3b80d8
*	Remove invalid test. This test requires that dead basic blocks be kept	Bill Wendling	2012-08-15	1	-19/+0
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	around. That's not how we do things. Besides, the commit message tells us that it is covered by the GCC test suite. ------------------------------------------------------------------------ r127497 \| zwarich \| 2011-03-11 13:51:56 -0800 (Fri, 11 Mar 2011) \| 3 lines Fix the GCC test suite issue exposed by r127477, which was caused by stack protector insertion not working correctly with unreachable code. Since that revision was rolled out, this test doesn't actual fail before this fix. ------------------------------------------------------------------------ git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161985 91177308-0d34-0410-b5e6-96231b3b80d8
*	Use vld1/vst1 to load/store f64 if alignment is < 4 and the target allows ↵	Evan Cheng	2012-08-15	1	-17/+49
\| \| \| \| \| \|	unaligned access. rdar://12091029 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161962 91177308-0d34-0410-b5e6-96231b3b80d8
*	The names of VFP variants of half-to-float conversion instructions were	Anton Korobeynikov	2012-08-14	1	-3/+3
\| \| \| \| \| \| \| \| \|	reversed. This leads to wrong codegen for float-to-half conversion intrinsics which are used to support storage-only fp16 type. NEON variants of same instructions are fine. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161907 91177308-0d34-0410-b5e6-96231b3b80d8
*	fix PR11334	Michael Liao	2012-08-14	1	-0/+56
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	- FP_EXTEND only support extending from vectors with matching elements. This results in the scalarization of extending to v2f64 from v2f32, which will be legalized to v4f32 not matching with v2f64. - add X86-specific VFPEXT supproting extending from v4f32 to v2f64. - add BUILD_VECTOR lowering helper to recover back the original extending from v4f32 to v2f64. - test case is enhanced to include different vector width. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161894 91177308-0d34-0410-b5e6-96231b3b80d8
*	During the CodeGenPrepare we often lower intrinsics (such as objsize)	Nadav Rotem	2012-08-14	5	-15/+16
\| \| \| \| \| \| \| \| \| \| \| \| \|	and allow some optimizations to turn conditional branches into unconditional. This commit adds a simple control-flow optimization which merges two consecutive basic blocks which are connected by a single edge. This allows the codegen to operate on larger basic blocks. rdar://11973998 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161852 91177308-0d34-0410-b5e6-96231b3b80d8
*	llvm/test/CodeGen/ARM/floorf.ll: Add explicit -mtriple=arm-unknown-unknown, ↵	NAKAMURA Takumi	2012-08-14	1	-1/+1
\| \| \| \| \| \|	or it fails on msvc. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161825 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add a roundToIntegral method to APFloat, which can be parameterized over ↵	Owen Anderson	2012-08-13	1	-0/+29
\| \| \| \| \| \|	various rounding modes. Use this to implement SelectionDAG constant folding of FFLOOR, FCEIL, and FTRUNC. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161807 91177308-0d34-0410-b5e6-96231b3b80d8
*	Rename test since it's not linux-specific.	Bill Wendling	2012-08-13	1	-0/+0
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161792 91177308-0d34-0410-b5e6-96231b3b80d8
*	Handle extra Tail predecessors in if-conversion.	Jakob Stoklund Olesen	2012-08-13	1	-0/+30
\| \| \| \| \| \| \| \|	It is still possible to if-convert if the tail block has extra predecessors, but the tail phis must be rewritten instead of being removed. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161781 91177308-0d34-0410-b5e6-96231b3b80d8
*	[Hexagon] Don't mark callee saved registers as clobbered by a tail call	Arnold Schwaighofer	2012-08-13	1	-0/+14
\| \| \| \| \| \| \| \| \| \| \| \|	This was causing unnecessary spills/restores of callee saved registers. Fixes PR13572. Patch by Pranav Bhandarkar! git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161778 91177308-0d34-0410-b5e6-96231b3b80d8
*	Fix failure on Atom bot due to r161769	Manman Ren	2012-08-13	1	-1/+1
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161777 91177308-0d34-0410-b5e6-96231b3b80d8
*	Do not optimize (or (and X,Y), Z) into BFI and other sequences if the AND ↵	Nadav Rotem	2012-08-13	1	-0/+17
\| \| \| \| \| \| \| \| \| \|	ISDNode has more than one user. rdar://11876519 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161775 91177308-0d34-0410-b5e6-96231b3b80d8
*	X86: move Int_CVTSD2SSrr, Int_CVTSI2SSrr, Int_CVTSI2SDrr, Int_CVTSS2SDrr from	Manman Ren	2012-08-13	1	-0/+14
\| \| \| \| \| \| \| \| \| \|	OpTbl1 to OpTbl2 since they have 3 operands and the last operand can be changed to a memory operand. PR13576 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161769 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add support for the %H output modifier.	Eric Christopher	2012-08-13	1	-0/+9
\| \| \| \| \| \|	Patch by Weiming Zhao. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161768 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add test for previous commit correcting NEON load patterns.	Tim Northover	2012-08-13	1	-0/+102
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161750 91177308-0d34-0410-b5e6-96231b3b80d8
*	Revert 161581: Patch to implement UMLAL/SMLAL instructions for the ARM	Arnold Schwaighofer	2012-08-12	2	-88/+0
\| \| \| \| \| \| \| \| \| \|	architecture It broke MultiSource/Applications/JM/ldecod/ldecod on armv7 thumb O0 g and armv7 thumb O3. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161736 91177308-0d34-0410-b5e6-96231b3b80d8
*	fix PR13577, an issue introduced by r161687	Michael Liao	2012-08-11	1	-0/+8
\| \| \| \| \| \| \| \| \| \|	- FCMOV only supports a subset of X86 conditions. Skip boolean simplification if X86 condition is not valid for FCMOV. - add a minimal test case for PR13577. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161732 91177308-0d34-0410-b5e6-96231b3b80d8
*	PR13578: Teach MachineCSE that instructions that use a constant register can ↵	Benjamin Kramer	2012-08-11	2	-2/+24
\| \| \| \| \| \| \| \|	be CSE'd safely. This is common e.g. when doing rip-relative addressing on x86_64. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161728 91177308-0d34-0410-b5e6-96231b3b80d8
*	X86: when we are auto-detecting the subtarget features, make sure we turn on	Manman Ren	2012-08-10	1	-1/+3
\| \| \| \| \| \| \| \| \| \| \| \|	FeatureFastUAMem for Nehalem, Westmere and Sandy Bridge. FeatureFastUAMem is already on if we pass in nehalem or westmere as a command argument. rdar: 7252306 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161717 91177308-0d34-0410-b5e6-96231b3b80d8
*	The normal edge of an invoke is not allowed to branch to a block with a	Eli Friedman	2012-08-10	1	-19/+0
\| \| \| \| \| \| \| \|	landingpad. Enforce it in the verifier, and fix the regression tests to match. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161697 91177308-0d34-0410-b5e6-96231b3b80d8
*	add X86-specific DAG optimization to simplify boolean test	Michael Liao	2012-08-10	1	-0/+42
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	- if a boolean test (X86ISD::CMP or X86ISD:SUB) checks a boolean value generated from X86ISD::SETCC, try to simplify the boolean value generation and checking by reusing the original EFLAGS with proper condition code - add hooks to X86 specific SETCC/BRCOND/CMOV, the major 3 places consuming EFLAGS part of patches fixing PR12312 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161687 91177308-0d34-0410-b5e6-96231b3b80d8
*	Update edge weights correctly in replaceSuccessor().	Jakob Stoklund Olesen	2012-08-10	1	-1/+1
\| \| \| \| \| \| \| \|	When replacing Old with New, it can happen that New is already a successor. Add the old and new edge weights instead of creating a duplicate edge. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161653 91177308-0d34-0410-b5e6-96231b3b80d8
*	Reapply r161633-161634 "Partition use lists so defs always come before uses.""	Jakob Stoklund Olesen	2012-08-10	2	-2/+2
\| \| \| \| \| \| \|	No changes to these patches, MRI needed to be notified when changing uses into defs and vice versa. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161644 91177308-0d34-0410-b5e6-96231b3b80d8
*	Revert r161633-161634 "Partition use lists so defs always come before uses."	Jakob Stoklund Olesen	2012-08-09	2	-2/+2
\| \| \| \| \| \|	These commits broke a number of buildbots. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161640 91177308-0d34-0410-b5e6-96231b3b80d8
*	Partition use lists so defs always come before uses.	Jakob Stoklund Olesen	2012-08-09	2	-3/+3
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This makes it possible to speed up def_iterator by stopping at the first use. This makes def_empty() and getUniqueVRegDef() much faster when there are many uses. In a +Asserts build, LiveVariables is 100x faster in one case because getVRegDef() has an assertion that would scan to the end of a def_iterator chain. Spill weight calculation is significantly faster (300x in one case) because isTriviallyReMaterializable() calls MRI->isConstantPhysReg(%RIP) which calls def_empty(%RIP). git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161634 91177308-0d34-0410-b5e6-96231b3b80d8
*	Don't use pointer-pointers for the register use lists.	Jakob Stoklund Olesen	2012-08-09	2	-3/+3
\| \| \| \| \| \| \| \| \| \| \|	Use a more conventional doubly linked list where the Prev pointers form a cycle. This means it is no longer necessary to adjust the Prev pointers when reallocating the VRegInfo array. The test changes are required because the register allocation hint is using the use-list order to break ties. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161633 91177308-0d34-0410-b5e6-96231b3b80d8
*	Don't modify MO while use_iterator is still pointing to it.	Jakob Stoklund Olesen	2012-08-09	1	-0/+1
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161626 91177308-0d34-0410-b5e6-96231b3b80d8
*	Patch to implement UMLAL/SMLAL instructions for the ARM architecture	Arnold Schwaighofer	2012-08-09	2	-0/+88
\| \| \| \| \| \| \| \| \| \| \| \|	This patch corrects the definition of umlal/smlal instructions and adds support for matching them to the ARM dag combiner. Bug 12213 Patch by Yin Ma! git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161581 91177308-0d34-0410-b5e6-96231b3b80d8
*	Fix the legalization of ExtLoad on ARM. ExpandUnalignedLoad did not properly	Nadav Rotem	2012-08-09	1	-0/+12
\| \| \| \| \| \| \| \| \|	handle the cases where the memory value type was illegal. PR 13111. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161565 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add test triples to fix win32 failures. Revert workaround from r161292.	Bob Wilson	2012-08-08	10	-22/+22
\| \| \| \| \| \|	I don't have a win32 system to test, so hopefully I got them all fixed here. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161519 91177308-0d34-0410-b5e6-96231b3b80d8
*	X86: enable CSE between CMP and SUB	Manman Ren	2012-08-08	1	-0/+23
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	We perform the following: 1> Use SUB instead of CMP for i8,i16,i32 and i64 in ISel lowering. 2> Modify MachineCSE to correctly handle implicit defs. 3> Convert SUB back to CMP if possible at peephole. Removed pattern matching of (a>b) ? (a-b):0 and like, since they are handled by peephole now. rdar://11873276 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161462 91177308-0d34-0410-b5e6-96231b3b80d8
*	X86 cmp lowering is looking past truncate on the condition node. It should only	Evan Cheng	2012-08-07	1	-0/+36
\| \| \| \| \| \| \| \| \|	do so when the high bits are known zero. This caused a subtle miscompilation. rdar://12027825 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161451 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add a much more conservative strategy for aligning branch targets.	Chandler Carruth	2012-08-07	3	-5/+10
\| \| \| \| \| \| \| \| \| \| \| \| \|	Previously, MBP essentially aligned every branch target it could. This bloats code quite a bit, especially non-looping code which has no real reason to prefer aligned branch targets so heavily. As Andy said in review, it's still a bit odd to do this without a real cost model, but this at least has much more plausible heuristics. Fixes PR13265. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161409 91177308-0d34-0410-b5e6-96231b3b80d8
*	MachineCSE: Update the heuristics for isProfitableToCSE.	Manman Ren	2012-08-07	2	-1/+36
\| \| \| \| \| \| \| \| \| \|	If the result of a common subexpression is used at all uses of the candidate expression, CSE should not increase the live range of the common subexpression. rdar://11393714 and rdar://11819721 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161396 91177308-0d34-0410-b5e6-96231b3b80d8
*	MFTB on PPC64 should really be encoded using MFSPR.	Hal Finkel	2012-08-06	1	-1/+1
\| \| \| \| \| \| \| \| \| \| \|	The MFTB instruction itself is being phased out, and its functionality is provided by MFSPR. According to the ISA docs, using MFSPR works on all known chips except for the 601 (which did not have a timebase register anyway) and the POWER3. Thanks to Adhemerval Zanella for pointing this out! git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161346 91177308-0d34-0410-b5e6-96231b3b80d8
*	Implement proper handling for pcmpistri/pcmpestri intrinsics. Requires ↵	Craig Topper	2012-08-06	1	-10/+10
\| \| \| \| \| \|	custom handling in DAGISelToDAG due to limitations in TableGen's implicit def handling. Fixes PR11305. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161318 91177308-0d34-0410-b5e6-96231b3b80d8
*	Update test to check for r161305	Craig Topper	2012-08-05	1	-0/+2
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161307 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add readcyclecounter lowering on PPC64.	Hal Finkel	2012-08-04	1	-0/+15
\| \| \| \| \| \| \| \|	On PPC64, this can be done with a simple TableGen pattern. To enable this, I've added the (otherwise missing) readcyclecounter SDNode definition to TargetSelectionDAG.td. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161302 91177308-0d34-0410-b5e6-96231b3b80d8
*	Add stack spill / reload instructions for DTriple and DQuad register ↵	Anton Korobeynikov	2012-08-04	1	-0/+174
\| \| \| \| \| \| \| \| \|	classes, which were missed for no reason. This fixes PR13377 git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161299 91177308-0d34-0410-b5e6-96231b3b80d8
*	Refactor and check "onlyReadsMemory" before optimizing builtins.	Bob Wilson	2012-08-03	11	-19/+19
\| \| \| \| \| \| \| \| \|	This patch is mostly just refactoring a bunch of copy-and-pasted code, but it also adds a check that the call instructions are readnone or readonly. That check was already present for sin, cos, sqrt, log2, and exp2 calls, but it was missing for the rest of the builtins being handled in this code. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161282 91177308-0d34-0410-b5e6-96231b3b80d8
*	1. Redo mips16 instructions to avoid multiple opcodes for same instruction.	Akira Hatanaka	2012-08-03	19	-0/+336
\| \| \| \| \| \| \| \| \| \| \|	Change these to patterns. 2. Add another 16 instructions. Patch by Reed Kotler. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161272 91177308-0d34-0410-b5e6-96231b3b80d8
*	Fix memcmp code-gen to honor -fno-builtin.	Bob Wilson	2012-08-03	1	-0/+3
\| \| \| \| \| \| \| \| \|	I noticed that SelectionDAGBuilder::visitCall was missing a check for memcmp in TargetLibraryInfo, so that it would use custom code for memcmp calls even with -fno-builtin. I also had to add a new -disable-simplify-libcalls option to llc so that I could write a test for this. git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161262 91177308-0d34-0410-b5e6-96231b3b80d8
*	Fall back to selection DAG isel for calls to builtin functions.	Bob Wilson	2012-08-03	1	-0/+8
\| \| \| \| \| \| \| \| \| \|	Fast isel doesn't currently have support for translating builtin function calls to target instructions. For embedded environments where the library functions are not available, this is a matter of correctness and not just optimization. Most of this patch is just arranging to make the TargetLibraryInfo available in fast isel. <rdar://problem/12008746> git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161232 91177308-0d34-0410-b5e6-96231b3b80d8
*	[arm-fast-isel] Add support for shl, lshr, and ashr.	Jush Lu	2012-08-03	1	-0/+50
\| \| \| \|	git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@161230 91177308-0d34-0410-b5e6-96231b3b80d8