llvm-project

Commit Graph

Author	SHA1	Message	Date
Chris Lattner	8c5cdddfb9	simplify some code. If we can infer alignment for source and dest that are greater than memcpy alignment, and if we lower to load/store, use the best alignment info we have. llvm-svn: 45943	2008-01-13 22:30:28 +00:00
Chris Lattner	5a86612d3f	simplify some code by adding a InsertBitCastBefore method, make memmove->memcpy conversion a bit simpler. llvm-svn: 45942	2008-01-13 22:23:22 +00:00
Chris Lattner	5bc253c8f2	Fix PR1907, a nasty miscompilation because instcombine didn't realize that ne & sgt was a signed comparison (it was only looking at whether the left compare was signed). llvm-svn: 45937	2008-01-13 20:59:02 +00:00
Duncan Sands	781f6549db	When turning a call to a bitcast function into a direct call, if this becomes a varargs call then deal correctly with any parameter attributes on the newly vararg call arguments. llvm-svn: 45931	2008-01-13 08:02:44 +00:00
Chris Lattner	2940c5c56d	Implement PR1795, an instcombine hack for forming GEPs with integer pointer arithmetic. llvm-svn: 45745	2008-01-08 07:23:51 +00:00
Duncan Sands	b18c30acec	Small cleanup for handling of type/parameter attribute incompatibility. llvm-svn: 45704	2008-01-07 17:16:06 +00:00
Duncan Sands	404eb05247	The transform that tries to turn calls to bitcast functions into direct calls bails out unless caller and callee have essentially equivalent parameter attributes. This is illogical - the callee's attributes should be of no relevance here. Rework the logic, which incidentally fixes a crash when removed arguments have attributes. llvm-svn: 45658	2008-01-06 18:27:01 +00:00
Duncan Sands	55e5090fe8	When transforming a call to a bitcast function into a direct call with cast parameters and cast return value (if any), instcombine was prepared to cast any non-void return value into any other, whether castable or not. Add a new predicate for testing whether casting is valid, and check it both for the return value and (as a cleanup) for the parameters. llvm-svn: 45657	2008-01-06 10:12:28 +00:00
Chris Lattner	e666bc272d	remove a couple more unsafe xforms in the face of overflow. llvm-svn: 45613	2008-01-05 01:22:42 +00:00
Chris Lattner	db026d703b	remove the (x-y) < 0 comparison xform, it miscompiles things that are not equality comparisons, for example: (2147479553+4096)-2147479553 < 0 != (2147479553+4096) < 2147479553 llvm-svn: 45612	2008-01-05 01:18:20 +00:00
Chris Lattner	f3ebc3f3d2	Remove attribution from file headers, per discussion on llvmdev. llvm-svn: 45418	2007-12-29 20:36:04 +00:00
Christopher Lamb	b053b80b79	Disable null pointer folding transforms for non-generic address spaces. This should probably be a target-specific predicate based on address space. That way for targets where this isn't applicable the predicate can be optimized away. llvm-svn: 45403	2007-12-29 07:56:53 +00:00
Owen Anderson	7363914ef7	Repair a transform that Chris noticed a bug in. Thanks to Nicholas for pointing out my stupid mistakes when writing this patch. :-) llvm-svn: 45384	2007-12-28 07:42:12 +00:00
Chris Lattner	5179819beb	disable this instcombine xform, it miscompiles: define i32 @main() { entry: %z = alloca i32 ; <i32> [#uses=2] store i32 0, i32 %z %tmp = load i32* %z ; <i32> [#uses=1] %sub = sub i32 %tmp, 1 ; <i32> [#uses=1] %cmp = icmp ult i32 %sub, 0 ; <i1> [#uses=1] %retval = select i1 %cmp, i32 1, i32 0 ; <i32> [#uses=1] ret i32 %retval } into ret 1, instead of ret 0. Christopher, please investigate. llvm-svn: 45383	2007-12-28 06:24:31 +00:00
Chris Lattner	74b2ab59fd	implement InstCombine/shift-trunc-shift.ll. This allows us to compile: #include <math.h> int t1(double d) { return signbit(d); } into: _t1: movd %xmm0, %rax shrq $63, %rax ret instead of: _t1: movd %xmm0, %rax shrq $32, %rax shrl $31, %eax ret on x86-64. llvm-svn: 45311	2007-12-22 09:07:47 +00:00
Christopher Lamb	7d82bc46b8	Implement review feedback, including additional transforms (icmp slt (sub A B) 1) -> (icmp sle A B) icmp sgt (sub A B) -1) -> (icmp sge A B) and add testcase. llvm-svn: 45256	2007-12-20 07:21:11 +00:00
Chris Lattner	16a51da0e2	simplify this code with the new m_Zero() pattern. Make sure the select only has a single use, and generalize it to not require N to be a constant. llvm-svn: 45250	2007-12-20 01:56:58 +00:00
Duncan Sands	aa31b92508	When inlining through an 'nounwind' call, mark inlined calls 'nounwind'. It is important for correct C++ exception handling that nounwind markings do not get lost, so this transformation is actually needed for correctness. llvm-svn: 45218	2007-12-19 21:13:37 +00:00
Christopher Lamb	f00ac6dd93	Fold subtracts into integer compares vs. zero. This improves generate code for this case on X86 from _foo: movl $99, %ecx movl 4(%esp), %eax subl %eax, %ecx xorl %edx, %edx testl %ecx, %ecx cmovs %edx, %eax ret to _foo: xorl %ecx, %ecx movl 4(%esp), %eax cmpl $99, %eax cmovg %ecx, %eax ret llvm-svn: 45173	2007-12-18 21:32:20 +00:00
Christopher Lamb	b7016c53d1	Fix comments llvm-svn: 45170	2007-12-18 20:33:11 +00:00
Christopher Lamb	74dbad9216	Remove an orthogonal transformation of the selection condition from my most recent submission. llvm-svn: 45169	2007-12-18 20:30:28 +00:00
Duncan Sands	3353ed09ac	Rename isNoReturn to doesNotReturn, and isNoUnwind to doesNotThrow. llvm-svn: 45160	2007-12-18 09:59:50 +00:00
Christopher Lamb	30291f4a30	Fix typos. llvm-svn: 45159	2007-12-18 09:45:40 +00:00
Christopher Lamb	8b09a464b4	Fold certain additions through selects (and their compares) so as to eliminate subtractions. This code is often produced by the SMAX expansion in SCEV. This implements test/Transforms/InstCombine/2007-12-18-AddSelCmpSub.ll llvm-svn: 45158	2007-12-18 09:34:41 +00:00
Christopher Lamb	edf0788758	Change the PointerType api for creating pointer types. The old functionality of PointerType::get() has become PointerType::getUnqual(), which returns a pointer in the generic address space. The new prototype of PointerType::get() requires both a type and an address space. llvm-svn: 45082	2007-12-17 01:12:55 +00:00
Duncan Sands	8e4847ee95	Make instcombine promote inline asm calls to 'nounwind' calls. Remove special casing of inline asm from the inliner. There is a potential problem: the verifier rejects invokes of inline asm (not sure why). If an asm call is not marked "nounwind" in some .ll, and instcombine is not run, but the inliner is run, then an illegal module will be created. This is bad but I'm not sure what the best approach is. I'm tempted to remove the check in the verifier... llvm-svn: 45073	2007-12-16 15:51:49 +00:00
Wojciech Matyjewicz	309e5a723b	1. "Upgrage" comments. 2. Using zero-extended value of Scale and unsigned division is safe provided that Scale doesn't have the sign bit set. Previously these 2 instructions: %p = bitcast [100 x {i8,i8,i8}]* %x to i8* %q = getelementptr i8* %p, i32 -4 were combined into: %q = getelementptr [100 x { i8, i8, i8 }]* %x, i32 0, i32 1431655764, i32 0 what was incorrect. llvm-svn: 44936	2007-12-12 15:21:32 +00:00
Chris Lattner	d2bbbabbfb	simplify some code. llvm-svn: 44655	2007-12-06 06:25:04 +00:00
Chris Lattner	0ccb663cca	move some ashr-specific code out of commonShiftTransforms into visitAShr. llvm-svn: 44650	2007-12-06 01:59:46 +00:00
Duncan Sands	ad0ea2d430	Fix PR1146: parameter attributes are longer part of the function type, instead they belong to functions and function calls. This is an updated and slightly corrected version of Reid Spencer's original patch. The only known problem is that auto-upgrading of bitcode files doesn't seem to work properly (see test/Bitcode/AutoUpgradeIntrinsics.ll). Hopefully a bitcode guru (who might that be? :) ) will fix it. llvm-svn: 44359	2007-11-27 13:23:08 +00:00
Chris Lattner	c00e8adfe0	Implement PR1822 llvm-svn: 44318	2007-11-25 21:27:53 +00:00
Duncan Sands	185eeac0f8	Fix PR1816. If a bitcast of a function only exists because of a trivial difference in function attributes, allow calls to it to be converted to direct calls. Based on a patch by Török Edwin. While there, move the various lists of mutually incompatible parameters etc out of the verifier and into ParameterAttributes.h. llvm-svn: 44315	2007-11-25 14:10:56 +00:00
Chris Lattner	0cf083815a	add a comment. llvm-svn: 44293	2007-11-23 22:35:18 +00:00
Chris Lattner	1985d96dc9	Fix PR1817. llvm-svn: 44284	2007-11-22 23:47:13 +00:00
Chris Lattner	c53b18362a	Fix PR1800 by correcting mistaken logic. llvm-svn: 44188	2007-11-16 06:04:17 +00:00
Andrew Lenharth	19ca5c7021	Better check llvm-svn: 43897	2007-11-08 18:45:15 +00:00
Andrew Lenharth	8cf11aa330	Fix PR1780 llvm-svn: 43893	2007-11-08 17:39:28 +00:00
Chris Lattner	d8515f8e80	Implement PR1777 by detecting dependent phis that all compute the same value. llvm-svn: 43777	2007-11-06 21:52:06 +00:00
Chris Lattner	362709dff1	wrap long lines llvm-svn: 43745	2007-11-06 01:15:27 +00:00
Dan Gohman	4decbc5002	Fix an abort in instcombine when folding creates a vector rem instruction. llvm-svn: 43743	2007-11-05 23:16:33 +00:00
Duncan Sands	44b8721de8	Executive summary: getTypeSize -> getTypeStoreSize / getABITypeSize. The meaning of getTypeSize was not clear - clarifying it is important now that we have x86 long double and arbitrary precision integers. The issue with long double is that it requires 80 bits, and this is not a multiple of its alignment. This gives a primitive type for which getTypeSize differed from getABITypeSize. For arbitrary precision integers it is even worse: there is the minimum number of bits needed to hold the type (eg: 36 for an i36), the maximum number of bits that will be overwriten when storing the type (40 bits for i36) and the ABI size (i.e. the storage size rounded up to a multiple of the alignment; 64 bits for i36). This patch removes getTypeSize (not really - it is still there but deprecated to allow for a gradual transition). Instead there is: (1) getTypeSizeInBits - a number of bits that suffices to hold all values of the type. For a primitive type, this is the minimum number of bits. For an i36 this is 36 bits. For x86 long double it is 80. This corresponds to gcc's TYPE_PRECISION. (2) getTypeStoreSizeInBits - the maximum number of bits that is written when storing the type (or read when reading it). For an i36 this is 40 bits, for an x86 long double it is 80 bits. This is the size alias analysis is interested in (getTypeStoreSize returns the number of bytes). There doesn't seem to be anything corresponding to this in gcc. (3) getABITypeSizeInBits - this is getTypeStoreSizeInBits rounded up to a multiple of the alignment. For an i36 this is 64, for an x86 long double this is 96 or 128 depending on the OS. This is the spacing between consecutive elements when you form an array out of this type (getABITypeSize returns the number of bytes). This is TYPE_SIZE in gcc. Since successive elements in a SequentialType (arrays, pointers and vectors) need to be aligned, the spacing between them will be given by getABITypeSize. This means that the size of an array is the length times the getABITypeSize. It also means that GEP computations need to use getABITypeSize when computing offsets. Furthermore, if an alloca allocates several elements at once then these too need to be aligned, so the size of the alloca has to be the number of elements multiplied by getABITypeSize. Logically speaking this doesn't have to be the case when allocating just one element, but it is simpler to also use getABITypeSize in this case. So alloca's and mallocs should use getABITypeSize. Finally, since gcc's only notion of size is that given by getABITypeSize, if you want to output assembler etc the same as gcc then getABITypeSize is the size you want. Since a store will overwrite no more than getTypeStoreSize bytes, and a read will read no more than that many bytes, this is the notion of size appropriate for alias analysis calculations. In this patch I have corrected all type size uses except some of those in ScalarReplAggregates, lib/Codegen, lib/Target (the hard cases). I will get around to auditing these too at some point, but I could do with some help. Finally, I made one change which I think wise but others might consider pointless and suboptimal: in an unpacked struct the amount of space allocated for a field is now given by the ABI size rather than getTypeStoreSize. I did this because every other place that reserves memory for a type (eg: alloca) now uses getABITypeSize, and I didn't want to make an exception for unpacked structs, i.e. I did it to make things more uniform. This only effects structs containing long doubles and arbitrary precision integers. If someone wants to pack these types more tightly they can always use a packed struct. llvm-svn: 43620	2007-11-01 20:53:16 +00:00
Chris Lattner	74709473ed	Fix InstCombine/2007-10-31-RangeCrash.ll llvm-svn: 43596	2007-11-01 02:18:41 +00:00
Chris Lattner	55b8302dfe	simplify some code by using the new isNaN predicate llvm-svn: 43305	2007-10-24 18:54:45 +00:00
Chris Lattner	c62877e9da	Implement a couple of foldings for ordered and unordered comparisons, implementing cases related to PR1738. llvm-svn: 43289	2007-10-24 05:38:08 +00:00
Devang Patel	df49cf52e2	Try again. Instead of loading small global string from memory, use integer constant. llvm-svn: 43148	2007-10-18 19:52:32 +00:00
Evan Cheng	cdcc1d0444	Reverting r43070 for now. It's causing llc test failures. llvm-svn: 43103	2007-10-17 23:51:13 +00:00
Devang Patel	91ff13edcc	Apply "Instead of loading small c string constant, use integer constant directly" transformation while processing load instruction. llvm-svn: 43070	2007-10-17 07:24:40 +00:00
Devang Patel	8d818f5e80	Use immediate stores. llvm-svn: 43055	2007-10-16 23:44:18 +00:00
Devang Patel	bff4aea328	Achieve same result but use fewer lines of code. llvm-svn: 42985	2007-10-15 15:31:35 +00:00
Devang Patel	371e6ca690	Dest type is always i8 *. This allows some simplification. Do not filter memmove. llvm-svn: 42930	2007-10-12 20:10:21 +00:00
Chris Lattner	ad618f66e6	Fix a bug in my patch last night that broke InstCombine/2007-10-12-Crash.ll llvm-svn: 42920	2007-10-12 18:05:47 +00:00
Gabor Greif	5d8f7e0cc7	eliminate warning llvm-svn: 42892	2007-10-12 07:44:54 +00:00
Chris Lattner	d8675e4915	Fix some 80 column violations. Fix DecomposeSimpleLinearExpr to handle simple constants better. Don't nuke gep(bitcast(allocation)) if the bitcast(allocation) will fold the allocation. This fixes PR1728 and Instcombine/malloc3.ll llvm-svn: 42891	2007-10-12 05:30:59 +00:00
Devang Patel	899cc56612	Lower memcpy if it makes sense. llvm-svn: 42864	2007-10-11 17:21:57 +00:00
Dale Johannesen	9d559cfff5	Tone down an overzealous optimization. llvm-svn: 42582	2007-10-03 17:45:27 +00:00
Duncan Sands	d31649bc59	Improve comment. llvm-svn: 42132	2007-09-19 10:25:38 +00:00
Duncan Sands	56df7dec2b	A global variable with external weak linkage can be null, while an alias could alias such a global variable. llvm-svn: 42130	2007-09-19 10:10:31 +00:00
Dan Gohman	2ac2652779	Instcombine x-((x/y)*y) into a remainder operator. llvm-svn: 42035	2007-09-17 17:31:57 +00:00
Duncan Sands	6d5da71288	Factor the trampoline transformation into a subroutine. llvm-svn: 42021	2007-09-17 10:26:40 +00:00
Dale Johannesen	98d3a08d8f	Remove the assumption that FP's are either float or double from some of the many places in the optimizers it appears, and do something reasonable with x86 long double. Make APInt::dump() public, remove newline, use it to dump ConstantSDNode's. Allow APFloats in FoldingSet. Expand X86 backend handling of long doubles (conversions to/from int, mostly). llvm-svn: 41967	2007-09-14 22:26:36 +00:00
Chris Lattner	d9111b88d1	silence a bogus gcc warning. llvm-svn: 41949	2007-09-14 03:07:24 +00:00
Duncan Sands	9204663bcb	Turn calls to trampolines into calls to the underlying nested function. llvm-svn: 41844	2007-09-11 14:35:41 +00:00
Chris Lattner	e804567cd8	remove some dead code, this is handled by constant folding. llvm-svn: 41819	2007-09-10 23:46:29 +00:00
Chris Lattner	85a51e0060	Don't zap back to back volatile load/stores llvm-svn: 41759	2007-09-07 05:33:03 +00:00
Dale Johannesen	bed9dc423c	Next round of APFloat changes. Use APFloat in UpgradeParser and AsmParser. Change all references to ConstantFP to use the APFloat interface rather than double. Remove the ConstantFP double interfaces. Use APFloat functions for constant folding arithmetic and comparisons. (There are still way too many places APFloat is just a wrapper around host float/double, but we're getting there.) llvm-svn: 41747	2007-09-06 18:13:44 +00:00
Nick Lewycky	0c5c47944a	Use isTrueWhenEqual. Thanks Chris! llvm-svn: 41741	2007-09-06 02:40:25 +00:00
Nick Lewycky	b0b066eaaa	When the two operands of an icmp are equal, there are five possible predicates that would make the icmp true. Fixes PR1637. llvm-svn: 41740	2007-09-06 01:10:22 +00:00
Chuck Rose III	2320323647	Forgot to obey 80 column rule. Fixing that. llvm-svn: 41725	2007-09-05 20:36:41 +00:00
Chuck Rose III	e58572233d	Added default parameters to GetElementPtrInstr constructor call. Visual Studio 2k5 was getting confused and was unable to compile it. Suspected compiler error. llvm-svn: 41721	2007-09-05 16:54:38 +00:00
David Greene	c656cbb8c2	Update GEP constructors to use an iterator interface to fix GLIBCXX_DEBUG issues. llvm-svn: 41697	2007-09-04 15:46:09 +00:00
Chris Lattner	0e258b8518	Cut off crazy computation. This helps PR1622 slightly. llvm-svn: 41522	2007-08-28 04:23:55 +00:00
David Greene	703623d571	Update InvokeInst to work like CallInst llvm-svn: 41506	2007-08-27 19:04:21 +00:00
Chris Lattner	99c8ee2977	Transform a load from an undef/zero global into an undef/global even if we have complex pointer manipulation going on. This allows us to compile stuff like this: __m128i foo(__m128i x){ static const unsigned int c_0[4] = { 0, 0, 0, 0 }; __m128i v_Zero = _mm_loadu_si128((__m128i*)c_0); x = _mm_unpacklo_epi8(x, v_Zero); return x; } into: _foo: xorps %xmm1, %xmm1 punpcklbw %xmm1, %xmm0 ret llvm-svn: 41022	2007-08-11 18:48:48 +00:00
Chris Lattner	a8e4b4bc7b	when we see a unaligned load from an insufficiently aligned global or alloca, increase the alignment of the load, turning it into an aligned load. This allows us to compile: #include <xmmintrin.h> __m128i foo(__m128i x){ static const unsigned int c_0[4] = { 0, 0, 0, 0 }; __m128i v_Zero = _mm_loadu_si128((__m128i*)c_0); x = _mm_unpacklo_epi8(x, v_Zero); return x; } into: _foo: punpcklbw _c_0.5944, %xmm0 ret .data .lcomm _c_0.5944,16,4 # c_0.5944 instead of: _foo: movdqu _c_0.5944, %xmm1 punpcklbw %xmm1, %xmm0 ret .data .lcomm _c_0.5944,16,2 # c_0.5944 llvm-svn: 40971	2007-08-09 19:05:49 +00:00
Nick Lewycky	8052019a20	It's safe to fold not of fcmp. llvm-svn: 40870	2007-08-06 20:04:16 +00:00
Chris Lattner	f0da7975ea	at the end of instcombine, explicitly clear WorklistMap. This shrinks it down to something small. On the testcase from PR1432, this speeds up instcombine from 0.7959s to 0.5000s, (59%) llvm-svn: 40840	2007-08-05 08:47:58 +00:00
Chandler Carruth	7132e00de7	This is the patch to provide clean intrinsic function overloading support in LLVM. It cleans up the intrinsic definitions and generally smooths the process for more complicated intrinsic writing. It will be used by the upcoming atomic intrinsics as well as vector and float intrinsics in the future. This also changes the syntax for llvm.bswap, llvm.part.set, llvm.part.select, and llvm.ct* intrinsics. They are automatically upgraded by both the LLVM ASM reader and the bitcode reader. The test cases have been updated, with special tests added to ensure the automatic upgrading is supported. llvm-svn: 40807	2007-08-04 01:51:18 +00:00
Chris Lattner	dc2cf228ce	Replacing a cast with another one does not reduce the number of casts in the input. llvm-svn: 40741	2007-08-02 17:23:38 +00:00
Chris Lattner	222b214be7	Disable an xform that causes an infinite loop. This fixes PR1594 llvm-svn: 40739	2007-08-02 16:56:32 +00:00
Chris Lattner	2740694450	wrap some long lines. Major offenders that are left include gvn, gvnpre, dse, and predsimplify. To see these, use: make check-line-length llvm-svn: 40738	2007-08-02 16:53:43 +00:00
Chris Lattner	b0418fc607	Enhance instcombine to be more aggressive about folding casts of operations of casts. This implements InstCombine/zext-fold.ll llvm-svn: 40726	2007-08-02 06:11:14 +00:00
David Greene	17a5dfe6f7	New CallInst interface to address GLIBCXX_DEBUG errors caused by indexing an empty std::vector. Updates to all clients. llvm-svn: 40660	2007-08-01 03:43:44 +00:00
Lauro Ramos Venancio	549e775e67	Fix a bug in GetKnownAlignment of packed structs. llvm-svn: 40649	2007-07-31 20:13:21 +00:00
Reid Spencer	dff9d69cfb	Fix a typo/thinko. llvm-svn: 40599	2007-07-30 19:53:57 +00:00
Chris Lattner	4512cd2cab	completely remove a transformation that is unsafe in the face of undefs. llvm-svn: 40439	2007-07-23 17:10:17 +00:00
Devang Patel	5e39293e62	Apply temporary work around to fix llvm mis-compilation reported in PR 1556. llvm-svn: 40133	2007-07-21 00:34:29 +00:00
Chris Lattner	d82e4a19cc	this xform is already done by the constant folder. llvm-svn: 40124	2007-07-20 22:06:41 +00:00
Dan Gohman	e31a61eeca	Optimize alignment of loads and stores. llvm-svn: 40102	2007-07-20 16:34:21 +00:00
Dan Gohman	06c60b6032	Fix comments about vectors to use the current wording. llvm-svn: 39921	2007-07-16 14:29:03 +00:00
Chris Lattner	640fd5124d	Repair a regression in Transforms/InstCombine/mul.ll that Reid noticed. llvm-svn: 39896	2007-07-16 04:15:34 +00:00
Chris Lattner	d4fef8dbca	Implement shift-simplify.ll:test[45]. First teach instcombine that sign bit checks only demand the sign bit, this allows simplify demanded bits to hack on expressions better. Second, teach instcombine that ashr is useless if only the sign bit is demanded. llvm-svn: 39880	2007-07-15 20:54:51 +00:00
Chris Lattner	06205d5567	Implement shift-simplify.ll:test3, turning: (X << 31) <s 0 --> (X&1) != 0 This happens dozens of times in the CFE. llvm-svn: 39879	2007-07-15 20:42:37 +00:00
Chris Lattner	fb032b176b	Significantly improve the documentation of the instcombine divide/compare transformation. Also, keep track of which end of the integer interval overflows occur on. This fixes Transforms/InstCombine/2007-06-21-DivCompareMiscomp.ll and rdar://5278853, a miscompilation of perl. llvm-svn: 37692	2007-06-21 18:11:19 +00:00
Chris Lattner	3bbec59e8b	refactor a bunch of code out of visitICmpInstWithInstAndIntCst into its own routine. llvm-svn: 37679	2007-06-20 23:46:26 +00:00
Chris Lattner	09a33a4f64	silence a bogus warning Duraid ran into. llvm-svn: 37649	2007-06-19 05:43:49 +00:00
Chris Lattner	373389260f	Generalize many transforms to work on ~ of vectors in addition to ~ of integer ops. This implements Transforms/InstCombine/and-or-not.ll test3/test4, and finishes off PR1510 llvm-svn: 37589	2007-06-15 06:23:19 +00:00
Chris Lattner	481e28b1f5	Implement two xforms: 1. ~(~X \| Y) === (X & ~Y) 2. (A\|B) & ~(A&B) -> A^B This allows us to transform ~(~(a\|b) \| (a&b)) -> a^b. This implements PR1510 for scalar values. llvm-svn: 37584	2007-06-15 05:58:24 +00:00
Chris Lattner	f14e5175ed	delete some obviously dead vector operations, which deletes a few thousand operations from Duraids example. llvm-svn: 37582	2007-06-15 05:26:55 +00:00
Lauro Ramos Venancio	368e8872db	Fix PR1499. llvm-svn: 37472	2007-06-06 17:08:48 +00:00
Chris Lattner	f79577d314	fix a miscompilation when passing a float through varargs llvm-svn: 37297	2007-05-23 01:17:04 +00:00
Chris Lattner	a655a157a0	Fix Transforms/InstCombine/2007-05-18-CastFoldBug.ll, a bug that devastates objc code due to the way the FE lowers objc message sends. llvm-svn: 37256	2007-05-19 06:51:32 +00:00
Chris Lattner	234f96daa8	Fix Transforms/InstCombine/2007-05-14-Crash.ll llvm-svn: 37057	2007-05-15 00:16:00 +00:00
Dan Gohman	b5650ebd6a	Fix typos. llvm-svn: 36994	2007-05-11 21:10:54 +00:00
Chris Lattner	600db3eb96	fix regressions from my previous checking, including Transforms/InstCombine/2006-12-08-ICmp-Combining.ll llvm-svn: 36989	2007-05-11 16:58:45 +00:00
Chris Lattner	fe2b44de9f	fix Transforms/InstCombine/2007-05-10-icmp-or.ll llvm-svn: 36984	2007-05-11 05:55:56 +00:00
Nick Lewycky	e7da2d6ac3	Fix typo in comment. llvm-svn: 36873	2007-05-06 13:37:16 +00:00
Chris Lattner	9b35b3e863	Fix a bug in my previous patch llvm-svn: 36857	2007-05-06 07:24:03 +00:00
Chris Lattner	5aa73fe34c	Implement Transforms/InstCombine/cast_ptr.ll llvm-svn: 36809	2007-05-05 22:41:33 +00:00
Chris Lattner	361e981415	wrap long lines llvm-svn: 36807	2007-05-05 22:32:24 +00:00
Chris Lattner	5c827bda0d	Fix InstCombine/2007-05-04-Crash.ll and PR1384 llvm-svn: 36775	2007-05-05 01:59:31 +00:00
Devang Patel	8c78a0bff0	Drop 'const' llvm-svn: 36662	2007-05-03 01:11:54 +00:00
Devang Patel	e95c6ad802	Use 'static const char' instead of 'static const int'. Due to darwin gcc bug, one version of darwin linker coalesces static const int, which defauts PassID based pass identification. llvm-svn: 36652	2007-05-02 21:39:20 +00:00
Devang Patel	09f162ca6a	Do not use typeinfo to identify pass in pass manager. llvm-svn: 36632	2007-05-01 21:15:47 +00:00
Chris Lattner	089e35cc57	fix a bug triggered by 403.gcc llvm-svn: 36527	2007-04-28 05:27:36 +00:00
Chris Lattner	6e880871e9	Fix several latent bugs in EmitGEPOffset that didn't manifest with its previous clients. This fixes MallocBench/gs llvm-svn: 36525	2007-04-28 04:52:43 +00:00
Chris Lattner	c753800800	uhn zap cvs llvm-svn: 36523	2007-04-28 03:50:56 +00:00
Chris Lattner	acbf6a401d	Implement PR1345 and Transforms/InstCombine/bitcast-gep.ll llvm-svn: 36521	2007-04-28 00:57:34 +00:00
Chris Lattner	1db224db92	refactor some code relating to pointer cast xforms, pulling it out of the codepath for unrelated casts. llvm-svn: 36511	2007-04-27 17:44:50 +00:00
Zhou Sheng	aafe4e216e	Make use of ConstantInt::isZero instead of ConstantInt::isNullValue. llvm-svn: 36261	2007-04-19 05:39:12 +00:00
Chris Lattner	4a6e0cbd41	Extend store merging to support the 'if/then' version in addition to if/then/else. This sinks the two stores in this example into a single store in cond_next. In this case, it allows elimination of the load as well: store double 0.000000e+00, double* @s.3060 %tmp3 = fcmp ogt double %tmp1, 5.000000e-01 ; <i1> [#uses=1] br i1 %tmp3, label %cond_true, label %cond_next cond_true: ; preds = %entry store double 1.000000e+00, double* @s.3060 br label %cond_next cond_next: ; preds = %entry, %cond_true %tmp6 = load double* @s.3060 ; <double> [#uses=1] This implements Transforms/InstCombine/store-merge.ll:test2 llvm-svn: 36040	2007-04-15 01:02:18 +00:00
Chris Lattner	14a251b937	refactor some code, no functionality change. llvm-svn: 36037	2007-04-15 00:07:55 +00:00
Chris Lattner	28d921d04f	fix long lines llvm-svn: 36031	2007-04-14 23:32:02 +00:00
Chris Lattner	7bfdd0abe1	Implement Transforms/InstCombine/vec_extract_elt.ll, transforming: define i32 @test(float %f) { %tmp7 = insertelement <4 x float> undef, float %f, i32 0 %tmp17 = bitcast <4 x float> %tmp7 to <4 x i32> %tmp19 = extractelement <4 x i32> %tmp17, i32 0 ret i32 %tmp19 } into: define i32 @test(float %f) { %tmp19 = bitcast float %f to i32 ; <i32> [#uses=1] ret i32 %tmp19 } On PPC, this is the difference between: _test: mfspr r2, 256 oris r3, r2, 8192 mtspr 256, r3 stfs f1, -16(r1) addi r3, r1, -16 addi r4, r1, -32 lvx v2, 0, r3 stvx v2, 0, r4 lwz r3, -32(r1) mtspr 256, r2 blr and: _test: stfs f1, -4(r1) nop nop nop lwz r3, -4(r1) blr llvm-svn: 36025	2007-04-14 23:02:14 +00:00
Chris Lattner	b37fb6a0da	Implement InstCombine/vec_demanded_elts.ll:test2. This allows us to turn unsigned test(float f) { return _mm_cvtsi128_si32( (__m128i) _mm_set_ss( f*f )); } into: _test: movss 4(%esp), %xmm0 mulss %xmm0, %xmm0 movd %xmm0, %eax ret instead of: _test: movss 4(%esp), %xmm0 mulss %xmm0, %xmm0 xorps %xmm1, %xmm1 movss %xmm0, %xmm1 movd %xmm1, %eax ret GCC gets: _test: subl $28, %esp movss 32(%esp), %xmm0 mulss %xmm0, %xmm0 xorps %xmm1, %xmm1 movss %xmm0, %xmm1 movaps %xmm1, %xmm0 movd %xmm0, 12(%esp) movl 12(%esp), %eax addl $28, %esp ret llvm-svn: 36020	2007-04-14 22:29:23 +00:00
Chris Lattner	efb33d28c6	Implement PR1201 and test/Transforms/InstCombine/malloc-free-delete.ll llvm-svn: 35981	2007-04-14 00:20:02 +00:00
Chris Lattner	74ff60ff84	Turn stuff like: icmp slt i32 %X, 0 ; <i1>:0 [#uses=1] sext i1 %0 to i32 ; <i32>:1 [#uses=1] into: %X.lobit = ashr i32 %X, 31 ; <i32> [#uses=1] This implements InstCombine/icmp.ll:test[34] llvm-svn: 35891	2007-04-11 06:57:46 +00:00
Chris Lattner	d0f7942e23	Simplify some comparisons to arithmetic, this implements: Transforms/InstCombine/icmp.ll llvm-svn: 35890	2007-04-11 06:53:04 +00:00
Chris Lattner	20f2372a7c	canonicalize (x <u 2147483648) -> (x >s -1) and (x >u 2147483647) -> (x <s 0) llvm-svn: 35886	2007-04-11 06:12:58 +00:00
Chris Lattner	7ddbff090a	fix a miscompilation of: define i32 @test(i32 %X) { entry: %Y = and i32 %X, 4 ; <i32> [#uses=1] icmp eq i32 %Y, 0 ; <i1>:0 [#uses=1] sext i1 %0 to i32 ; <i32>:1 [#uses=1] ret i32 %1 } by moving code out of commonIntCastTransforms into visitZExt. Simplify the APInt gymnastics in it etc. llvm-svn: 35885	2007-04-11 05:45:39 +00:00
Chris Lattner	467b69cabb	Strengthen the boundary conditions of this fold, implementing InstCombine/set.ll:test25 llvm-svn: 35852	2007-04-09 23:52:13 +00:00
Chris Lattner	a87c9f6114	Fix PR1304 and Transforms/InstCombine/2007-04-08-SingleEltVectorCrash.ll llvm-svn: 35792	2007-04-09 01:37:55 +00:00
Chris Lattner	4ca9cbb170	Eliminate useless insertelement instructions. This implements Transforms/InstCombine/vec_insertelt.ll and fixes PR1286. We now compile the code from that bug into: _foo: movl 4(%esp), %eax movdqa (%eax), %xmm0 movl 8(%esp), %ecx psllw (%ecx), %xmm0 movdqa %xmm0, (%eax) ret instead of: _foo: subl $4, %esp movl %ebp, (%esp) movl %esp, %ebp movl 12(%ebp), %eax movdqa (%eax), %xmm0 #IMPLICIT_DEF %eax pinsrw $2, %eax, %xmm0 xorl %ecx, %ecx pinsrw $3, %ecx, %xmm0 pinsrw $4, %eax, %xmm0 pinsrw $5, %ecx, %xmm0 pinsrw $6, %eax, %xmm0 pinsrw $7, %ecx, %xmm0 movl 8(%ebp), %eax movdqa (%eax), %xmm1 psllw %xmm0, %xmm1 movdqa %xmm1, (%eax) movl %ebp, %esp popl %ebp ret woo :) llvm-svn: 35788	2007-04-09 01:11:16 +00:00
Chris Lattner	c8d3788f71	reenable this xform, whoops :) llvm-svn: 35765	2007-04-08 08:01:49 +00:00
Chris Lattner	7621a031d8	Fix regression on Instcombine/apint-or2.ll llvm-svn: 35763	2007-04-08 07:55:22 +00:00
Chris Lattner	1150df9cc4	Generalize the code that handles (A&B)\|(A&C) to work where B/C are not constants. Add a new xform to simplify (A&B)\|(~A&C). THis implements InstCombine/or2.ll:test1 llvm-svn: 35760	2007-04-08 07:47:01 +00:00
Chris Lattner	3dbe65f80a	implement Transforms/InstCombine/malloc2.ll and PR1313 llvm-svn: 35700	2007-04-06 18:57:34 +00:00
Dale Johannesen	7c2001d014	Prevent transformConstExprCastCall from generating conversions that assert elsewhere. llvm-svn: 35668	2007-04-04 19:16:42 +00:00
Jeff Cohen	5a1c750f31	Fix 2007-04-04-BadFoldBitcastIntoMalloc.ll llvm-svn: 35665	2007-04-04 16:58:57 +00:00
Duncan Sands	f01a47c93c	Fix comment. llvm-svn: 35655	2007-04-04 06:42:45 +00:00
Chris Lattner	e5bbb3cb1a	Fix a bug I introduced with my patch yesterday which broke Qt (I converted some constant exprs to apints). Thanks to Anton for tracking down a small testcase that triggered this! llvm-svn: 35633	2007-04-03 23:29:39 +00:00
Chris Lattner	a74deafb13	reinstate the previous two patches, with a bugfix :) ldecod now passes. llvm-svn: 35626	2007-04-03 17:43:25 +00:00
Evan Cheng	7511fa280d	Reverting back to 1.723. The last two commits broke JM (and possibily others) on ARM. llvm-svn: 35620	2007-04-03 08:11:50 +00:00
Chris Lattner	64c764cebc	Split a whole ton of code out of visitICmpInst into visitICmpInstWithInstAndIntCst. llvm-svn: 35614	2007-04-03 04:46:52 +00:00
Chris Lattner	8b2ec5f506	Fix PR1253 and xor2.ll:test[01] llvm-svn: 35612	2007-04-03 01:47:41 +00:00
Zhou Sheng	9bc8ab100d	1. Make use of APInt operation instead of using ConstantExpr::getXXX. 2. Use cheaper APInt methods. llvm-svn: 35594	2007-04-02 13:45:30 +00:00
Zhou Sheng	56cda95658	Use uint32_t for bitwidth instead of unsigned. llvm-svn: 35593	2007-04-02 08:20:41 +00:00
Chris Lattner	9d5aacee92	Wrap long line llvm-svn: 35588	2007-04-02 05:48:58 +00:00
Chris Lattner	50490d54f2	use more obvious function name. llvm-svn: 35587	2007-04-02 05:42:22 +00:00
Chris Lattner	b24acc7bee	simplify (x+c)^signbit as (x+c+signbit), pointed out by PR1288. This implements test/Transforms/InstCombine/xor.ll:test28 llvm-svn: 35584	2007-04-02 05:36:22 +00:00
Chris Lattner	c3eeb42809	simplify this code, make it work for ap ints llvm-svn: 35561	2007-04-01 20:57:36 +00:00
Zhou Sheng	150f3bbab2	Avoid unnecessary APInt construction. llvm-svn: 35555	2007-04-01 17:13:37 +00:00
Reid Spencer	6bba6c8143	For PR1297: Support overloaded intrinsics bswap, ctpop, cttz, ctlz. llvm-svn: 35547	2007-04-01 07:35:23 +00:00
Chris Lattner	0427799531	Fix InstCombine/2007-03-31-InfiniteLoop.ll llvm-svn: 35536	2007-04-01 05:36:37 +00:00
Zhou Sheng	82c42284f4	Delete dead code. llvm-svn: 35525	2007-03-31 02:50:26 +00:00
Zhou Sheng	4f16402e0d	Use APInt operators to calculate the carry bits, remove this loop. llvm-svn: 35524	2007-03-31 02:38:39 +00:00
Zhou Sheng	fd28a33031	Make sure the use of ConstantInt::getZExtValue() for shift amount safe. llvm-svn: 35510	2007-03-30 17:20:39 +00:00
Zhou Sheng	b25806fa5f	1. Make sure the use of ConstantInt::getZExtValue() for getting shift amount is safe. 2. Use new method on ConstantInt instead of (? :) operator. 3. Use new method uge() on ConstantInt to simplify codes. llvm-svn: 35505	2007-03-30 09:29:48 +00:00
Zhou Sheng	5e60a4a6b0	Use APInt operation instead of ConstantExpr::getXX. llvm-svn: 35503	2007-03-30 05:45:18 +00:00
Zhou Sheng	b3a80b1d70	1. Make more use of APInt::getHighBitsSet/getLowBitsSet. 2. Let APInt variable do the binary operation stuff instead of using ConstantExpr::getXXX. llvm-svn: 35450	2007-03-29 08:15:12 +00:00
Zhou Sheng	444af49cc0	Clean up some codes in InstCombiner::SimplifyDemandedBits(). llvm-svn: 35446	2007-03-29 04:45:55 +00:00
Zhou Sheng	a4475575c0	Clean up codes in InstCombiner::SimplifyDemandedBits(): 1. Line out nested call of APInt::zext/trunc. 2. Make more use of APInt::getHighBitsSet/getLowBitsSet. 3. Use APInt[] operator instead of expression like "APIntVal & SignBit". llvm-svn: 35444	2007-03-29 02:26:30 +00:00
Zhou Sheng	4961cf1c06	1. Make the APInt variable do the binary operation stuff if possible instead of using ConstantExpr::getXX. 2. Use constant reference to APInt if possible instead of expensive APInt copy. llvm-svn: 35443	2007-03-29 01:57:21 +00:00
Zhou Sheng	117477e28b	Avoid unnecessary APInt construction. llvm-svn: 35431	2007-03-28 17:38:21 +00:00
Zhou Sheng	23f7a1c947	1. Make more use of getLowBitsSet/getHighBitsSet. 2. Use APInt[] instead of "X & SignBit". 3. Clean up some codes. 4. Make the expression like "ShiftAmt = ShiftAmtC->getZExtValue()" safe. llvm-svn: 35424	2007-03-28 15:02:20 +00:00
Zhou Sheng	2777a31850	1. Make more use of getLowBitsSet/getHighBitsSet. 2. Make the APInt value do the zext/trunc stuff instead of using ConstantExpr::getZExt(). llvm-svn: 35422	2007-03-28 09:19:01 +00:00
Zhou Sheng	c2d3309b99	Use UnknownBIts[BitWidth-1] instead of UnknownBIts & SignBits. llvm-svn: 35418	2007-03-28 05:15:57 +00:00
Zhou Sheng	18570b1f14	Remove unused APInt variable. llvm-svn: 35414	2007-03-28 03:02:21 +00:00
Zhou Sheng	57e3f7324b	Clean up codes in ComputeMaskedBits(): 1. Line out nested use of zext/trunc. 2. Make more use of getHighBitsSet/getLowBitsSet. 3. Use APInt[] != 0 instead of "(APInt & SignBit) != 0". llvm-svn: 35408	2007-03-28 02:19:03 +00:00
Reid Spencer	a5c18bf798	For PR1280: When converting an add/xor/and triplet into a trunc/sext, only do so if the intermediate integer type is a bitwidth that the targets can handle. llvm-svn: 35400	2007-03-28 01:36:16 +00:00
Evan Cheng	a4ed8a512a	Unbreaks non-debug builds. llvm-svn: 35383	2007-03-27 16:44:48 +00:00
Reid Spencer	54d5b1b8f8	Implement some minor review feedback. llvm-svn: 35373	2007-03-26 23:58:26 +00:00
Reid Spencer	441486c172	For PR1271: Fix another incorrectly converted shift mask. llvm-svn: 35371	2007-03-26 23:45:51 +00:00
Chris Lattner	d2602d5054	eliminate use of std::set llvm-svn: 35361	2007-03-26 20:40:50 +00:00
Reid Spencer	755d0e7ffc	Get better debug output by having modified instructions print both the original and new instruction. A slight performance hit with ostringstream but it is only for debug. Also, clean up an uninitialized variable warning noticed in a release build. llvm-svn: 35358	2007-03-26 17:44:01 +00:00
Reid Spencer	769a5a8e0b	Get the number of bits to set in a mask correct for a shl/lshr transform. llvm-svn: 35357	2007-03-26 17:18:58 +00:00
Reid Spencer	50898607a9	For PR1271: Fix SingleSource/Regression/C/2003-05-21-UnionBitFields.c by changing a getHighBitsSet call to getLowBitsSet call that was incorrectly converted from the original lshr constant expression. llvm-svn: 35348	2007-03-26 05:25:00 +00:00
Reid Spencer	52830327e9	For PR1271: Remove a use of getLowBitsSet that caused the mask used for replacement of shl/lshr pairs with an AND instruction to be computed incorrectly. Its not clear exactly why this is the case. This solves the disappearing shifts problem, but it doesn't fix Regression/C/2003-05-21-UnionBitFields. It seems there is more going on. llvm-svn: 35342	2007-03-25 21:11:44 +00:00
Chris Lattner	9bf53ffaa2	implement Transforms/InstCombine/cast2.ll:test3 and PR1263 llvm-svn: 35341	2007-03-25 20:43:09 +00:00
Reid Spencer	624766f8a2	Some cleanup from review: * Don't assume shift amounts are <= 64 bits * Avoid creating an extra APInt in SubOne and AddOne by using -- and ++ * Add another use of getLowBitsSet * Convert a series of if statements to a switch llvm-svn: 35339	2007-03-25 19:55:33 +00:00
Reid Spencer	80263aadf3	Refactor several ConstantExpr::getXXX calls with ConstantInt arguments using the facilities of APInt. While this duplicates a tiny fraction of the constant folding code, it also makes the code easier to read and avoids large ConstantExpr overhead for simple, known computations. llvm-svn: 35335	2007-03-25 05:33:51 +00:00
Zhou Sheng	222d5ebfd2	1. Avoid unnecessary APInt construction if possible. 2. Use isStrictlyPositive() instead of isPositive() in two places where they need APInt value > 0 not only >=0. llvm-svn: 35333	2007-03-25 05:01:29 +00:00
Reid Spencer	cd99fbdf3b	Make more uses of getHighBitsSet and get rid of some pointless & of an APInt with its type mask. llvm-svn: 35325	2007-03-25 04:26:16 +00:00
Reid Spencer	d8aad61d4d	More APIntification: * Convert the last use of a uint64_t that should have been an APInt. * Change ComputeMaskedBits to have a const reference argument for the Mask so that recursions don't cause unneeded temporaries. This causes temps to be needed in other places (where the mask has to change) but this change optimizes for the recursion which is more frequent. * Remove two instances of &ing a Mask with getAllOnesValue. Its not needed any more because APInt is accurate in its bit computations. * Start using the getLowBitsSet and getHighBits set methods on APInt instead of shifting. This makes it more clear in the code what is going on. llvm-svn: 35321	2007-03-25 02:03:12 +00:00
Chris Lattner	3a8248f79d	fix a regression on vector or instructions. llvm-svn: 35314	2007-03-24 23:56:43 +00:00
Zhou Sheng	e9ebd3f6ba	Make some codes more efficient. llvm-svn: 35297	2007-03-24 15:34:37 +00:00
Reid Spencer	a962d18774	For PR1205: Convert some calls to ConstantInt::getZExtValue() into getValue() and use APInt facilities in the subsequent computations. llvm-svn: 35294	2007-03-24 00:42:08 +00:00
Reid Spencer	959a21d3dc	For PR1205: * APIntify visitAdd and visitSelectInst * Remove unused uint64_t versions of utility functions that have been replaced with APInt versions. This completes most of the changes for APIntification of InstCombine. This passes llvm-test and llvm/test/Transforms/InstCombine/APInt. Patch by Zhou Sheng. llvm-svn: 35287	2007-03-23 21:24:59 +00:00
Reid Spencer	6d39206bc2	For PR1205: APIntify visitDiv, visitMul and visitRem. Patch by Zhou Sheng. llvm-svn: 35283	2007-03-23 20:05:17 +00:00
Chris Lattner	12b89cc148	switch AddReachableCodeToWorklist from being recursive to being iterative. llvm-svn: 35282	2007-03-23 19:17:18 +00:00
Reid Spencer	6274c72ee1	For PR1205: APIntify several utility functions supporting logical operators and shift operators. Patch by Zhou Sheng. llvm-svn: 35281	2007-03-23 18:46:34 +00:00
Zhou Sheng	0900993ebc	Make the "KnownZero ^ TypeMask" computation just once. llvm-svn: 35276	2007-03-23 03:13:21 +00:00
Zhou Sheng	755f04b5d7	Simplify the code. llvm-svn: 35275	2007-03-23 02:39:25 +00:00
Reid Spencer	b722f2b110	For PR1205: APInt support for logical operators in visitAnd, visitOr, and visitXor. Patch by Zhou Sheng. llvm-svn: 35273	2007-03-22 22:19:58 +00:00
Reid Spencer	4154e732e6	For PR1205: * APIntify commonIntCastTransforms * APIntify visitTrunc * APIntify visitZExt Patch by Zhou Sheng. llvm-svn: 35271	2007-03-22 20:56:53 +00:00
Reid Spencer	c3e3b8a32f	For PR1205: * Re-enable the APInt version of MaskedValueIsZero. * APIntify the Comput{Un}SignedMinMaxValuesFromKnownBits functions * APIntify visitICmpInst. llvm-svn: 35270	2007-03-22 20:36:03 +00:00
Dan Gohman	dcb291faa4	Change uses of Function::front to Function::getEntryBlock for readability. llvm-svn: 35265	2007-03-22 16:38:57 +00:00
Reid Spencer	f40711637f	For PR1248: * Fix some indentation and comments in InsertRangeTest * Add an "IsSigned" parameter to AddWithOverflow and make it handle signed additions. Also, APIntify this function so it works with any bitwidth. * For the icmp pred ([us]div %X, C1), C2 transforms, exit early if the div instruction's RHS is zero. * Finally, for icmp pred (sdiv %X, C1), -C2, fix an off-by-one error. The HiBound needs to be incremented in order to get the range test correct. llvm-svn: 35247	2007-03-21 23:19:50 +00:00
Zhou Sheng	b3949340c8	Simplify isHighOnes(). llvm-svn: 35211	2007-03-20 12:49:06 +00:00
Reid Spencer	6682721316	Make isOneBitSet faster by using APInt::isPowerOf2. Thanks Chris. llvm-svn: 35194	2007-03-20 00:16:52 +00:00
Reid Spencer	cc031a43aa	APIntify the isHighOnes utility function. llvm-svn: 35190	2007-03-19 21:29:50 +00:00
Reid Spencer	ef599b0786	Implement isMaxValueMinusOne in terms of APInt instead of uint64_t. Patch by Sheng Zhou. llvm-svn: 35188	2007-03-19 21:10:28 +00:00
Reid Spencer	3b93db72b4	Implement isMinValuePlusOne using facilities of APInt instead of uint64_t Patch by Zhou Sheng. llvm-svn: 35187	2007-03-19 21:08:07 +00:00
Reid Spencer	129a86792d	Implement isOneBitSet in terms of APInt::countPopulation. llvm-svn: 35186	2007-03-19 21:04:43 +00:00
Reid Spencer	450434ed65	1. Use APInt::getSignBit to reduce clutter (patch by Sheng Zhou) 2. Replace uses of the "isPositive" utility function with APInt::isPositive llvm-svn: 35185	2007-03-19 20:58:18 +00:00
Reid Spencer	03c31d5bb0	Remove a redundant clause in an if statement. Patch by Sheng Zhou. llvm-svn: 35184	2007-03-19 20:47:50 +00:00
Chris Lattner	0741842b3b	Implement InstCombine/and-xor-merge.ll:test[12]. Rearrange some code to simplify it now that shifts are binops llvm-svn: 35145	2007-03-18 22:51:34 +00:00
Zhou Sheng	d8c645b0ba	ShiftAmt might equal to zero. Handle this situation. llvm-svn: 35094	2007-03-14 09:07:33 +00:00
Zhou Sheng	b912844554	Enable KnownZero/One.clear(). llvm-svn: 35093	2007-03-14 03:21:24 +00:00
Chris Lattner	d1bce956b4	ifdef out some dead code. Fix PR1244 and Transforms/InstCombine/2007-03-13-CompareMerge.ll llvm-svn: 35082	2007-03-13 14:27:42 +00:00
Zhou Sheng	ebe634e662	For expression like "APInt::getAllOnesValue(ShiftAmt).zextOrCopy(BitWidth)", to handle ShiftAmt == BitWidth situation, use zextOrCopy() instead of zext(). llvm-svn: 35080	2007-03-13 06:40:59 +00:00
Zhou Sheng	af4341d441	In APInt version ComputeMaskedBits(): 1. Ensure VTy, KnownOne and KnownZero have same bitwidth. 2. Make code more efficient. llvm-svn: 35078	2007-03-13 02:23:10 +00:00
Reid Spencer	1791f23803	Add an APInt version of SimplifyDemandedBits. Patch by Zhou Sheng. llvm-svn: 35064	2007-03-12 17:25:59 +00:00
Reid Spencer	d9281784be	Add an APInt version of ShrinkDemandedConstant. Patch by Zhou Sheng. llvm-svn: 35063	2007-03-12 17:15:10 +00:00
Zhou Sheng	be171ee5cd	Avoid to assert on "(KnownZero & KnownOne) == 0". llvm-svn: 35062	2007-03-12 16:54:56 +00:00
Zhou Sheng	b3e00c4656	In function ComputeMaskedBits(): 1. Replace getSignedMinValue() with getSignBit() for better code readability. 2. Replace APIntOps::shl() with operator<<= for convenience. 3. Make APInt construction more effective. llvm-svn: 35060	2007-03-12 05:44:52 +00:00
Zhou Sheng	d1eb3d593e	Fix a bug in function ComputeMaskedBits(). llvm-svn: 35027	2007-03-08 15:15:18 +00:00
Zhou Sheng	387d7b1a35	Fix a bug in APIntified ComputeMaskedBits(). llvm-svn: 35022	2007-03-08 05:42:00 +00:00
Reid Spencer	bb5741fb02	For PR1205: Provide an APIntified version of MaskedValueIsZero. This will (temporarily) cause a "defined but not used" message from the compiler. It will be used in the next patch in this series. Patch by Sheng Zhou. llvm-svn: 35019	2007-03-08 01:52:58 +00:00
Reid Spencer	aa69640b10	For PR1205: Add a new ComputeMaskedBits function that is APIntified. We'll slowly convert things over to use this version. When its all done, we'll remove the existing version. llvm-svn: 35018	2007-03-08 01:46:38 +00:00
Reid Spencer	3939b1a274	Remove an unnecessary if statement and adjust indentation. llvm-svn: 34939	2007-03-05 23:36:13 +00:00
Chris Lattner	fe53cf2459	fix a subtle bug that caused an MSVC warning. Thanks to Jeffc for pointing this out. llvm-svn: 34920	2007-03-05 00:11:19 +00:00
Chris Lattner	5fdded1d2f	Add some simplifications for demanded bits, this allows instcombine to turn: define i64 @test(i64 %A, i32 %B) { %tmp12 = zext i32 %B to i64 ; <i64> [#uses=1] %tmp3 = shl i64 %tmp12, 32 ; <i64> [#uses=1] %tmp5 = add i64 %tmp3, %A ; <i64> [#uses=1] %tmp6 = and i64 %tmp5, 123 ; <i64> [#uses=1] ret i64 %tmp6 } into: define i64 @test(i64 %A, i32 %B) { %tmp6 = and i64 %A, 123 ; <i64> [#uses=1] ret i64 %tmp6 } This implements Transforms/InstCombine/add2.ll:test1 llvm-svn: 34919	2007-03-05 00:02:29 +00:00
Jeff Cohen	b622c11f77	Unbreak VC++ build. llvm-svn: 34917	2007-03-05 00:00:42 +00:00
Chris Lattner	ab2f913b68	simplify some code llvm-svn: 34914	2007-03-04 23:16:36 +00:00
Chris Lattner	8258b44b22	Speed up -instcombine by 20% by avoiding a particularly expensive passmgr call. llvm-svn: 34902	2007-03-04 04:27:24 +00:00
Chris Lattner	da1d04a057	my recent change caused a failure in a bswap testcase, because it changed the order that instcombine processed instructions in the testcase. The end result is that instcombine finished with: define i16 @test1(i16 %a) { %tmp = zext i16 %a to i32 ; <i32> [#uses=2] %tmp21 = lshr i32 %tmp, 8 ; <i32> [#uses=1] %tmp5 = shl i32 %tmp, 8 ; <i32> [#uses=1] %tmp.upgrd.32 = or i32 %tmp21, %tmp5 ; <i32> [#uses=1] %tmp.upgrd.3 = trunc i32 %tmp.upgrd.32 to i16 ; <i16> [#uses=1] ret i16 %tmp.upgrd.3 } which can't get matched as a bswap. This patch makes instcombine more sophisticated about removing truncating casts, allowing it to turn this into: define i16 @test2(i16 %a) { %tmp211 = lshr i16 %a, 8 %tmp52 = shl i16 %a, 8 %tmp.upgrd.323 = or i16 %tmp211, %tmp52 ret i16 %tmp.upgrd.323 } which then matches as bswap. This fixes bswap.ll and implements InstCombine/cast2.ll:test[12]. This also implements cast elimination of add/sub. llvm-svn: 34870	2007-03-03 05:27:34 +00:00
Chris Lattner	960a543037	add a top-level iteration loop to instcombine. This means that it will never finish without combining something it is capable of. llvm-svn: 34865	2007-03-03 02:04:50 +00:00
Chris Lattner	b15e2b182f	Fix a significant algorithm problem with the instcombine worklist. removing a value from the worklist required scanning the entire worklist to remove all entries. We now use a combination map+vector to prevent duplicates from happening and prevent the scan. This speeds up instcombine on a large file from the llvm-gcc bootstrap from 189.7s to 4.84s in a debug build and from 5.04s to 1.37s in a release build. llvm-svn: 34848	2007-03-02 21:28:56 +00:00
Chris Lattner	51f5457ad4	minor cleanup llvm-svn: 34846	2007-03-02 19:59:19 +00:00
Reid Spencer	24f1a0e78f	The 64-bit constructor for ConstantInt changes from int64_t to uint64_t. This caused a warning for construction with -1. Avoid the warning by using -1ULL instead. llvm-svn: 34796	2007-03-01 19:33:52 +00:00
Chris Lattner	c4d8e7e614	Fix InstCombine/2007-02-23-PhiFoldInfLoop.ll and PR1217 llvm-svn: 34546	2007-02-24 01:03:45 +00:00
Chris Lattner	99c6cf60f1	convert more vectors to smallvectors, 2.8% speedup llvm-svn: 34333	2007-02-15 22:52:10 +00:00
Chris Lattner	af6094fe3f	change some vectors to smallvectors. This speeds up instcombine on 447.dealII by 5%. llvm-svn: 34332	2007-02-15 22:48:32 +00:00
Chris Lattner	7907e5fe07	switch an std::set to a SmallPtr set, this speeds up instcombine by 9.5% on 447.dealII llvm-svn: 34323	2007-02-15 19:41:52 +00:00
Reid Spencer	d84d35ba70	For PR1195: Rename PackedType -> VectorType, ConstantPacked -> ConstantVector, and PackedTyID -> VectorTyID. No functional changes. llvm-svn: 34293	2007-02-15 02:26:10 +00:00
Chris Lattner	945e437c65	Generalize TargetData strings, to support more interesting forms of data. Patch by Scott Michel. llvm-svn: 34266	2007-02-14 05:52:17 +00:00
Chris Lattner	a06a8fd2d7	Eliminate use of ctors that take vectors. llvm-svn: 34219	2007-02-13 02:10:56 +00:00
Chris Lattner	a731513406	stop using methods that take vectors. llvm-svn: 34205	2007-02-12 22:56:41 +00:00
Chris Lattner	6e0123b17f	Simplify code by using value::takename llvm-svn: 34176	2007-02-11 01:23:03 +00:00
Chris Lattner	83ac5ae9f3	Fix miscompilations of consumer-typeset, telecomm-gsm, and 176.gcc. llvm-svn: 33902	2007-02-05 05:57:49 +00:00
Chris Lattner	0a28e90f2c	fix a miscompilation of 176.gcc llvm-svn: 33900	2007-02-05 04:09:35 +00:00
Chris Lattner	3e009e8b8f	rewrite shift/shift folding, now that types are not signed. llvm-svn: 33892	2007-02-05 00:57:54 +00:00
Reid Spencer	3f4e6e84dc	For PR1163: Make the Module's dependent library use a std::vector instead of SetVector adjust #includes in .cpp files because SetVector.h is no longer included. llvm-svn: 33855	2007-02-04 00:40:42 +00:00
Chris Lattner	6c344e56b1	remove some dead code llvm-svn: 33845	2007-02-03 23:28:07 +00:00
Reid Spencer	2f34b98cbf	Remove dead code and fix indentation per Chris' review comments. llvm-svn: 33785	2007-02-02 14:41:37 +00:00
Reid Spencer	0d5f9237b6	Use short form of binary operator create functions. llvm-svn: 33783	2007-02-02 14:08:20 +00:00
Chris Lattner	d5fea61d98	bugfix for reid's shift patch. llvm-svn: 33779	2007-02-02 05:29:55 +00:00
Reid Spencer	2341c22ec7	Changes to support making the shift instructions be true BinaryOperators. This feature is needed in order to support shifts of more than 255 bits on large integer types. This changes the syntax for llvm assembly to make shl, ashr and lshr instructions look like a binary operator: shl i32 %X, 1 instead of shl i32 %X, i8 1 Additionally, this should help a few passes perform additional optimizations. llvm-svn: 33776	2007-02-02 02:16:23 +00:00
Chris Lattner	c904205d28	Fix Transforms/InstCombine/2007-02-01-LoadSinkAlloca.ll, a serious code pessimization where instcombine can sink a load (good for code size) that prevents an alloca from being promoted by mem2reg (bad for everything). llvm-svn: 33771	2007-02-01 22:30:07 +00:00
Chris Lattner	416a8939c3	remove temporary vectors. llvm-svn: 33715	2007-01-31 20:08:52 +00:00
Chris Lattner	4fc18a4cb8	Revert another incorrectly applied chunk, which fixes InstCombine/vec_insert_to_shuffle.ll llvm-svn: 33705	2007-01-31 18:09:17 +00:00
Chris Lattner	f96f4a874c	eliminate temporary vectors llvm-svn: 33693	2007-01-31 04:40:53 +00:00
Chris Lattner	aa17576933	Move symbolic constant folding code to libanalysis. llvm-svn: 33688	2007-01-31 00:53:10 +00:00
Chris Lattner	024f4ab383	Adjust #includes to match movement of constant folding code from transformutils to libanalysis. llvm-svn: 33680	2007-01-30 23:46:24 +00:00
Chris Lattner	e3eda25641	pass TD to constant folding apis llvm-svn: 33674	2007-01-30 23:16:15 +00:00
Chris Lattner	2b15f2ba9d	remove some bits that are not yet meant to land. llvm-svn: 33666	2007-01-30 22:50:32 +00:00
Chris Lattner	4284f6463a	Symbolically evaluate constant expressions like &A[123] - &A[4].f. This occurs in C++ code like: #include <iostream> #include <iterator> int a[] = { 1, 2, 3, 4, 5 }; int main() { using namespace std; copy(a, a + sizeof(a)/sizeof(a[0]), ostream_iterator<int>(cout, "\n")); return 0; } Before we would decide the loop trip count is: sdiv (i32 sub (i32 ptrtoint (i32* getelementptr ([5 x i32]* @a, i32 0, i32 5) to i32), i32 ptrtoint ([5 x i32]* @a to i32)), i32 4) Now we decide it is "5". Amazing. This code will need to be refactored, but I'm doing that as a separate commit. llvm-svn: 33665	2007-01-30 22:32:46 +00:00
Reid Spencer	5301e7c605	For PR1136: Rename GlobalVariable::isExternal as isDeclaration to avoid confusion with external linkage types. llvm-svn: 33663	2007-01-30 20:08:39 +00:00
Chris Lattner	c8fb6de78c	Fix test/Transforms/InstCombine/2007-01-27-AndICmp.ll, a miscompilation of Mozilla that Anton tracked down. llvm-svn: 33591	2007-01-27 23:08:34 +00:00
Reid Spencer	31a4ef4dc1	Cleanup checks in the load and store of casted pointer transforms. Two changes: (1) don't special case for i1 any more, (2) use the new TargetData::getTypeSizeInBits method to ensure source and dest are the same bit width. llvm-svn: 33427	2007-01-22 05:51:25 +00:00
Reid Spencer	9a4bed06dd	Revise the store V, (cast P) -> store (cast V) -> P transform. We only want to do this if the src and destination types have the same bit width. This patch uses TargetData::getTypeSizeInBits() instead of making a special case for integer types and avoiding the transform if they don't match. llvm-svn: 33414	2007-01-20 23:35:48 +00:00
Chris Lattner	50ee0e40e5	Teach TargetData to handle 'preferred' alignment for each target, and use these alignment amounts to align scalars when we can. Patch by Scott Michel! llvm-svn: 33409	2007-01-20 22:35:55 +00:00
Reid Spencer	e928a15c9e	For this transform: store V, (cast P) -> store (cast V), P don't allow the transform if V and the pointer's element type are different width integer types. llvm-svn: 33371	2007-01-19 21:20:31 +00:00
Reid Spencer	a94d394ad2	For PR1043: This is the final patch for this PR. It implements some minor cleanup in the use of IntegerType, to wit: 1. Type::getIntegerTypeMask -> IntegerType::getBitMask 2. Type::IntTy changed to IntegerType from Type* 3. ConstantInt::getType() returns IntegerType* now, not Type* This also fixes PR1120. Patch by Sheng Zhou. llvm-svn: 33370	2007-01-19 21:13:56 +00:00
Chris Lattner	120ab038eb	Fix InstCombine/2007-01-18-VectorInfLoop.ll, a case where instcombine infinitely loops. llvm-svn: 33343	2007-01-18 22:16:33 +00:00
Reid Spencer	c050af9126	Clean up some code around the store V, (cast P) -> store (cast V), P transform. Change some variable names so it is clear what is source and what is dest of the cast. Also, add an assert to ensure that the integer to integer case is asserting if the bitwidths are different. This prevents illegal casts from being formed and catches bitwidth bugs sooner. llvm-svn: 33337	2007-01-18 18:54:33 +00:00
Chris Lattner	479a9fc492	Fix a regression in my isIntegral patch that broke 471.omnetpp. This is because TargetData::getTypeSize() returns the same for i1 and i8. This fix is not right for the full generality of bitwise types, but it fixes the regression. llvm-svn: 33237	2007-01-15 17:55:20 +00:00
Chris Lattner	c8dcede292	Implement InstCombine/phi.ll:test7, deletion of trivial value loops for induction variables. llvm-svn: 33234	2007-01-15 07:30:06 +00:00
Chris Lattner	27df1db485	simplify some code now that types are signless llvm-svn: 33232	2007-01-15 07:02:54 +00:00
Chris Lattner	a4beeef76c	delete stores to allocas with one use. This is a trivial form of DSE which often kicks in for ?: expressions. llvm-svn: 33231	2007-01-15 06:51:56 +00:00
Chris Lattner	03c4953cdd	rename Type::isIntegral to Type::isInteger, eliminating the old Type::isInteger. rename Type::getIntegralTypeMask to Type::getIntegerTypeMask. This makes naming much more consistent. For example, there are now no longer any instances of IntegerType that are not considered isInteger! :) llvm-svn: 33225	2007-01-15 02:27:26 +00:00
Chris Lattner	1942249c5b	Eliminate calls to isInteger, generalizing code and tightening checks as needed. llvm-svn: 33218	2007-01-15 01:55:30 +00:00
Chris Lattner	6ee923f3bb	instcombine has always been miscompiling fcmp x, x, disregarding possible NANs. This fixes PR1111 and Transforms/InstCombine/2007-01-14-FcmpSelf.ll llvm-svn: 33208	2007-01-14 19:42:17 +00:00
Chris Lattner	387bf3f700	Fix Transforms/InstCombine/2007-01-13-ExtCompareMiscompile.ll, which is part of PR1107 llvm-svn: 33185	2007-01-13 23:11:38 +00:00
Reid Spencer	7a9c62baa6	For PR1064: Implement the arbitrary bit-width integer feature. The feature allows integers of any bitwidth (up to 64) to be defined instead of just 1, 8, 16, 32, and 64 bit integers. This change does several things: 1. Introduces a new Derived Type, IntegerType, to represent the number of bits in an integer. The Type classes SubclassData field is used to store the number of bits. This allows 2^23 bits in an integer type. 2. Removes the five integer Type::TypeID values for the 1, 8, 16, 32 and 64-bit integers. These are replaced with just IntegerType which is not a primitive any more. 3. Adjust the rest of LLVM to account for this change. Note that while this incremental change lays the foundation for arbitrary bit-width integers, LLVM has not yet been converted to actually deal with them in any significant way. Most optimization passes, for example, will still only deal with the byte-width integer types. Future increments will rectify this situation. llvm-svn: 33113	2007-01-12 07:05:14 +00:00
Reid Spencer	cddc9dfe97	Implement review feedback for the ConstantBool->ConstantInt merge. Chris recommended that getBoolValue be replaced with getZExtValue and that get(bool) be replaced by get(const Type*, uint64_t). This implements those changes. llvm-svn: 33110	2007-01-12 04:24:46 +00:00
Reid Spencer	542964f55b	Rename BoolTy as Int1Ty. Patch by Sheng Zhou. llvm-svn: 33076	2007-01-11 18:21:29 +00:00
Zhou Sheng	bd23db9968	Remove unnecessary boolean type check. llvm-svn: 33075	2007-01-11 14:38:17 +00:00
Zhou Sheng	75b871fb1e	For PR1043: Merge ConstantIntegral and ConstantBool into ConstantInt. Remove ConstantIntegral and ConstantBool from LLVM. llvm-svn: 33073	2007-01-11 12:24:14 +00:00
Jeff Cohen	223004cd12	Unbreak VC++ build. llvm-svn: 33021	2007-01-08 20:17:17 +00:00
Reid Spencer	8f166b0ef3	Comparison of primitive type sizes should now be done in bits, not bytes. This patch converts getPrimitiveSize to getPrimitiveSizeInBits where it is appropriate to do so (comparison of integer primitive types). llvm-svn: 33012	2007-01-08 16:32:00 +00:00
Chris Lattner	fbc524fe87	relax some types llvm-svn: 32980	2007-01-07 06:58:05 +00:00
Chris Lattner	7051d758de	Fix regressions in InstCombine/call-cast-target.ll and InstCombine/2003-11-13-ConstExprCastCall.ll llvm-svn: 32959	2007-01-06 19:53:32 +00:00
Chris Lattner	c343a99786	this final call to canLosslesslyBitCastTo is dead, because ValueRequiresCast is only called on integers. llvm-svn: 32949	2007-01-06 02:11:56 +00:00
Chris Lattner	400f959a0c	simplify some more code now that there are not multiple different integer types of the same size llvm-svn: 32948	2007-01-06 02:09:32 +00:00
Chris Lattner	64d87b0215	eliminate some uses of canLosslesslyBitCastTo, this actually makes the code stronger, by nuking relational pointer comparisons with casts. llvm-svn: 32947	2007-01-06 01:45:59 +00:00
Chris Lattner	d7b6ea166d	Implement InstCombine/vec_shuffle.ll:%test7, simplifying shuffles with undef operands. llvm-svn: 32899	2007-01-05 07:36:08 +00:00
Chris Lattner	17c7c030c2	fold things like a^b != c^a -> b != c. This implements InstCombine/xor.ll:test27 llvm-svn: 32893	2007-01-05 03:04:57 +00:00
Chris Lattner	23eb8ec78b	Compile X + ~X to -1. This implements Instcombine/add.ll:test34 llvm-svn: 32890	2007-01-05 02:17:46 +00:00
Reid Spencer	6ff3e73db6	Death to useless bitcast instructions! llvm-svn: 32866	2007-01-04 05:23:51 +00:00
Reid Spencer	c635f47d9a	For PR950: This patch replaces signed integer types with signless ones: 1. [US]Byte -> Int8 2. [U]Short -> Int16 3. [U]Int -> Int32 4. [U]Long -> Int64. 5. Removal of isSigned, isUnsigned, getSignedVersion, getUnsignedVersion and other methods related to signedness. In a few places this warranted identifying the signedness information from other sources. llvm-svn: 32785	2006-12-31 05:48:39 +00:00
Reid Spencer	193df25eb9	For PR1066: Fix this by ensuring that a bitcast is inserted to do sign switching. This is only temporarily needed as the merging of signed and unsigned is next on the SignlessTypes plate. llvm-svn: 32757	2006-12-24 00:40:59 +00:00
Reid Spencer	910f23f7d7	Shut up some compilers that can't accurately analyze variable usage correctly and emit "may be used uninitialized" warnings. llvm-svn: 32756	2006-12-23 19:17:57 +00:00
Reid Spencer	43c77d53ff	For PR1065: Don't allow CmpInst instances to be processed in FoldSelectOpOp because you can't easily swap their operands. llvm-svn: 32753	2006-12-23 18:58:04 +00:00
Reid Spencer	266e42b312	For PR950: This patch removes the SetCC instructions and replaces them with the ICmp and FCmp instructions. The SetCondInst instruction has been removed and been replaced with ICmpInst and FCmpInst. llvm-svn: 32751	2006-12-23 06:05:41 +00:00
Chris Lattner	79a42ac941	Switch over Transforms/Scalar to use the STATISTIC macro. For each statistic converted, we lose a static initializer. This also allows GCC to emit warnings about unused statistics. llvm-svn: 32690	2006-12-19 21:40:18 +00:00
Reid Spencer	668d90f289	Convert the last uses of CastInst::createInferredCast to a normal cast creation. These changes are still temporary but at least this pushes knowledge of signedness out closer to where it can be determined properly and allows signedness to be removed from VMCore. llvm-svn: 32654	2006-12-18 08:47:13 +00:00
Reid Spencer	74a528b427	Fix a bug in EvaluateInDifferentType. The type of operand should not be used to determine whether a ZExt or SExt cast is performed. Instead, pass an "isSigned" bool to the function and determine its value from the opcode of the cast involved. Also, clean up some cruft from previous patches. llvm-svn: 32548	2006-12-13 18:21:21 +00:00
Reid Spencer	2a499b0b6c	Implement review feedback. Most of this has to do with removing unnecessary cast instructions. A few are bug fixes. llvm-svn: 32544	2006-12-13 17:19:09 +00:00
Reid Spencer	612683b0d7	For mul transforms, when checking for a cast from bool as either operand, make sure to also check that it is a zext from bool, not any other cast operation type. llvm-svn: 32539	2006-12-13 08:33:33 +00:00

... 4 5 6 7 8 ...

1117 Commits