llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-11-25 20:23:11 +01:00

Author	SHA1	Message	Date
Owen Anderson	84faaa8d14	Break critical edges coming into blocks with PHI nodes. llvm-svn: 44019	2007-11-12 17:27:27 +00:00
Evan Cheng	d887bfa88e	Refactor some code. llvm-svn: 44010	2007-11-12 06:35:08 +00:00
Owen Anderson	40abf86e03	As Chris and Evan pointed out, BreakCriticalMachineEdges doesn't really need to be a pass of its own. Instead, move it out into a helper method. llvm-svn: 44002	2007-11-12 01:05:09 +00:00
Hartmut Kaiser	72dc0f6e89	Fixed a strange construct. Please review. llvm-svn: 43960	2007-11-09 19:59:00 +00:00
Duncan Sands	edf7e3b5f4	Move MinAlign to MathExtras.h. llvm-svn: 43944	2007-11-09 13:41:39 +00:00
Duncan Sands	7df7c7aed1	Fix some load/store logic that would be wrong for apints on big-endian machines if the bitwidth is not a multiple of 8. Introduce a new helper, MVT::getStoreSizeInBits, and use it. llvm-svn: 43934	2007-11-09 08:57:19 +00:00
Duncan Sands	3a9cf68f35	Add terminating newline. llvm-svn: 43933	2007-11-09 08:30:21 +00:00
Evan Cheng	7d8deec92f	Much improved pic jumptable codegen: Then: call "L1$pb" "L1$pb": popl %eax ... LBB1_1: # entry imull $4, %ecx, %ecx leal LJTI1_0-"L1$pb"(%eax), %edx addl LJTI1_0-"L1$pb"(%ecx,%eax), %edx jmpl %edx .align 2 .set L1_0_set_3,LBB1_3-LJTI1_0 .set L1_0_set_2,LBB1_2-LJTI1_0 .set L1_0_set_5,LBB1_5-LJTI1_0 .set L1_0_set_4,LBB1_4-LJTI1_0 LJTI1_0: .long L1_0_set_3 .long L1_0_set_2 Now: call "L1$pb" "L1$pb": popl %eax ... LBB1_1: # entry addl LJTI1_0-"L1$pb"(%eax,%ecx,4), %eax jmpl %eax .align 2 .set L1_0_set_3,LBB1_3-"L1$pb" .set L1_0_set_2,LBB1_2-"L1$pb" .set L1_0_set_5,LBB1_5-"L1$pb" .set L1_0_set_4,LBB1_4-"L1$pb" LJTI1_0: .long L1_0_set_3 .long L1_0_set_2 llvm-svn: 43924	2007-11-09 01:32:10 +00:00
Evan Cheng	781ed42681	Didn't mean to check these in. llvm-svn: 43923	2007-11-09 01:28:33 +00:00
Evan Cheng	7e1d8b99ab	Bug fix. Passive nodes are not in SUnitMap. llvm-svn: 43922	2007-11-09 01:27:11 +00:00
Owen Anderson	17a70ad8c6	This preserves critical edge breaking. llvm-svn: 43911	2007-11-08 22:23:57 +00:00
Owen Anderson	8f46b986d9	Make BreakCriticalMachineEdges available as a pass that can be depended on. llvm-svn: 43910	2007-11-08 22:20:23 +00:00
Evan Cheng	d9bab93a44	If both parts of smul_lohi, etc. are used, don't simplify. If only one part is used, try simplify it. llvm-svn: 43888	2007-11-08 09:25:29 +00:00
Owen Anderson	1032243608	Add the majority of machine-level critical edge breaking pass. Most of this was written by Fernando, cleanup and updating to TOT by me. This still needs a bit of work, particularly to handle jump tables properly. llvm-svn: 43885	2007-11-08 07:55:43 +00:00
Owen Anderson	b735086b08	Take another stab at getting isLiveIn() and isLiveOut() right. llvm-svn: 43869	2007-11-08 01:32:45 +00:00
Owen Anderson	ba84ab5b21	Bring UsedBlocks back. StrongPHIElimination needs this information. llvm-svn: 43866	2007-11-08 01:20:48 +00:00
Evan Cheng	5b53732be2	Simplify my (il)logic. llvm-svn: 43819	2007-11-07 08:08:25 +00:00
Owen Anderson	7021ab1c78	Add some more of StrongPHIElim. llvm-svn: 43805	2007-11-07 05:17:15 +00:00
Dan Gohman	ff12f4602f	Remainder operations must be either integer or floating-point. llvm-svn: 43781	2007-11-06 22:11:54 +00:00
Evan Cheng	c401482711	When the allocator rewrite a spill register with new virtual register, it replaces other operands of the same register. Watch out for situations where only some of the operands are sub-register uses. llvm-svn: 43776	2007-11-06 21:12:10 +00:00
Evan Cheng	bd9a038bd7	First step towards moving the coalescer to priority_queue based machinery. llvm-svn: 43764	2007-11-06 08:52:21 +00:00
Evan Cheng	f3e53ebd0e	Fix a bug where a def use operand isn't being detected as a sub-register use. llvm-svn: 43763	2007-11-06 08:50:44 +00:00
Evan Cheng	3764ad2bac	Add pseudo dependency to force two-address instruction to be scheduled after other uses. There was a overly restricted check that prevented some obvious cases. llvm-svn: 43762	2007-11-06 08:44:59 +00:00
Owen Anderson	c3baea32f3	Add a few comments. llvm-svn: 43755	2007-11-06 05:26:02 +00:00
Owen Anderson	03decb2fca	DomForest is a forest of registers, not instructions. llvm-svn: 43754	2007-11-06 05:22:43 +00:00
Owen Anderson	d0fb7600f9	StrongPHIElimination requires LiveVariables. llvm-svn: 43751	2007-11-06 04:49:43 +00:00
Dan Gohman	6255ce9f5d	Add support for vector remainder operations. llvm-svn: 43744	2007-11-05 23:35:22 +00:00
Rafael Espindola	ec025c3042	Move the LowerMEMCPY and LowerMEMCPYCall to a common place. Thanks for the suggestions Bill :-) llvm-svn: 43742	2007-11-05 23:12:20 +00:00
Dale Johannesen	1f70f86c7a	Make labels work in asm blocks; allow labels as parameters. Rename ValueRefList to ParamList in AsmParser, since its only use is for parameters. llvm-svn: 43734	2007-11-05 21:20:28 +00:00
Duncan Sands	c338eafe35	Don't output ABI size padding twice. By using the store size for the field we get ABI padding automatically, so no need to put it in again when we emit the field. llvm-svn: 43720	2007-11-05 18:03:02 +00:00
Evan Cheng	ec35b58b0a	Move SimpleRegisterCoalescing.h to lib/CodeGen since there is now a common register coalescer interface: RegisterCoalescing. llvm-svn: 43714	2007-11-05 17:41:38 +00:00
Evan Cheng	28c61e33a4	Skip over deleted val#'s. llvm-svn: 43700	2007-11-05 06:46:45 +00:00
Evan Cheng	e5eac2c5ac	Handle cases where a register and one of its super-register are both marked as defined on the same instruction. This fixes PR1767. llvm-svn: 43699	2007-11-05 03:11:55 +00:00
Evan Cheng	13d79ab67a	Fix PR1187. llvm-svn: 43692	2007-11-05 00:59:10 +00:00
Duncan Sands	d1bdbd010b	Eliminate the remaining uses of getTypeSize. This should only effect x86 when using long double. Now 12/16 bytes are output for long double globals (the exact amount depends on the alignment). This brings globals in line with the rest of LLVM: the space reserved for an object is now always the ABI size. One tricky point is that only 10 bytes should be output for long double if it is a field in a packed struct, which is the reason for the additional argument to EmitGlobalConstant. llvm-svn: 43688	2007-11-05 00:04:43 +00:00
Owen Anderson	5fff5dcf65	Another step of stronger PHI elimination down. llvm-svn: 43684	2007-11-04 22:33:26 +00:00
Evan Cheng	9a5c2f1169	If an interval is being undone clear its preference as well since the source interval may have been undone as well. llvm-svn: 43670	2007-11-04 08:32:21 +00:00
Evan Cheng	1771f6da9c	There are times when the coalescer would not coalesce away a copy but the copy can be eliminated by the allocator is the destination and source targets the same register. The most common case is when the source and destination registers are in different class. For example, on x86 mov32to32_ targets GR32_ which contains a subset of the registers in GR32. The allocator can do 2 things: 1. Set the preferred allocation for the destination of a copy to that of its source. 2. After allocation is done, change the allocation of a copy destination (if legal) so the copy can be eliminated. This eliminates 443 extra moves from 403.gcc. llvm-svn: 43662	2007-11-03 07:20:12 +00:00
Dan Gohman	19d88d511b	Add std:: to sort calls. llvm-svn: 43652	2007-11-02 22:24:01 +00:00
Dan Gohman	26c8800fbd	Change illegal uses of ++ to uses of STLExtra.h's next function. llvm-svn: 43651	2007-11-02 22:22:02 +00:00
Evan Cheng	65a07e73e2	One more extract_subreg coalescing bug. llvm-svn: 43644	2007-11-02 17:35:08 +00:00
Duncan Sands	281da5e25f	Fix a thinko. llvm-svn: 43639	2007-11-02 15:18:06 +00:00
Duncan Sands	eb464e976f	Executive summary: getTypeSize -> getTypeStoreSize / getABITypeSize. The meaning of getTypeSize was not clear - clarifying it is important now that we have x86 long double and arbitrary precision integers. The issue with long double is that it requires 80 bits, and this is not a multiple of its alignment. This gives a primitive type for which getTypeSize differed from getABITypeSize. For arbitrary precision integers it is even worse: there is the minimum number of bits needed to hold the type (eg: 36 for an i36), the maximum number of bits that will be overwriten when storing the type (40 bits for i36) and the ABI size (i.e. the storage size rounded up to a multiple of the alignment; 64 bits for i36). This patch removes getTypeSize (not really - it is still there but deprecated to allow for a gradual transition). Instead there is: (1) getTypeSizeInBits - a number of bits that suffices to hold all values of the type. For a primitive type, this is the minimum number of bits. For an i36 this is 36 bits. For x86 long double it is 80. This corresponds to gcc's TYPE_PRECISION. (2) getTypeStoreSizeInBits - the maximum number of bits that is written when storing the type (or read when reading it). For an i36 this is 40 bits, for an x86 long double it is 80 bits. This is the size alias analysis is interested in (getTypeStoreSize returns the number of bytes). There doesn't seem to be anything corresponding to this in gcc. (3) getABITypeSizeInBits - this is getTypeStoreSizeInBits rounded up to a multiple of the alignment. For an i36 this is 64, for an x86 long double this is 96 or 128 depending on the OS. This is the spacing between consecutive elements when you form an array out of this type (getABITypeSize returns the number of bytes). This is TYPE_SIZE in gcc. Since successive elements in a SequentialType (arrays, pointers and vectors) need to be aligned, the spacing between them will be given by getABITypeSize. This means that the size of an array is the length times the getABITypeSize. It also means that GEP computations need to use getABITypeSize when computing offsets. Furthermore, if an alloca allocates several elements at once then these too need to be aligned, so the size of the alloca has to be the number of elements multiplied by getABITypeSize. Logically speaking this doesn't have to be the case when allocating just one element, but it is simpler to also use getABITypeSize in this case. So alloca's and mallocs should use getABITypeSize. Finally, since gcc's only notion of size is that given by getABITypeSize, if you want to output assembler etc the same as gcc then getABITypeSize is the size you want. Since a store will overwrite no more than getTypeStoreSize bytes, and a read will read no more than that many bytes, this is the notion of size appropriate for alias analysis calculations. In this patch I have corrected all type size uses except some of those in ScalarReplAggregates, lib/Codegen, lib/Target (the hard cases). I will get around to auditing these too at some point, but I could do with some help. Finally, I made one change which I think wise but others might consider pointless and suboptimal: in an unpacked struct the amount of space allocated for a field is now given by the ABI size rather than getTypeStoreSize. I did this because every other place that reserves memory for a type (eg: alloca) now uses getABITypeSize, and I didn't want to make an exception for unpacked structs, i.e. I did it to make things more uniform. This only effects structs containing long doubles and arbitrary precision integers. If someone wants to pack these types more tightly they can always use a packed struct. llvm-svn: 43620	2007-11-01 20:53:16 +00:00
Evan Cheng	7c2d41dcc1	- Coalesce extract_subreg when both intervals are relatively small. - Some code clean up. llvm-svn: 43606	2007-11-01 06:22:48 +00:00
Duncan Sands	b86535ad9a	Promotion of sdiv/srem/udiv/urem. llvm-svn: 43551	2007-10-31 08:57:43 +00:00
Duncan Sands	c945dfa9af	Add a newline at the end of the file. llvm-svn: 43550	2007-10-31 08:49:24 +00:00
Owen Anderson	ae0530aaf4	Add the skeleton of a better PHI elimination pass. llvm-svn: 43542	2007-10-31 03:37:57 +00:00
Owen Anderson	4314cf9d58	Some fixes to get MachineDomTree working better. llvm-svn: 43541	2007-10-31 03:30:14 +00:00
Dale Johannesen	9bc04ae496	Make i64=expand_vector_elt(v2i64) work in 32-bit mode. llvm-svn: 43535	2007-10-31 00:32:36 +00:00
Evan Cheng	1fef4e369a	Typo. llvm-svn: 43511	2007-10-30 20:11:21 +00:00

1 2 3 4 5 ...

4221 Commits