llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-10-26 14:33:02 +02:00

Author	SHA1	Message	Date
Evan Cheng	36b3babfde	Added support for new condition code modeling scheme (i.e. physical register dependency). These are a bunch of instructions that are duplicated so the x86 backend can support both the old and new schemes at the same time. They will be deleted after all the kinks are worked out. llvm-svn: 42285	2007-09-25 01:57:46 +00:00
Dan Gohman	96d5f979bc	Add support on x86 for having Legalize lower ISD::LOCATION to ISD::DEBUG_LOC instead of ISD::LABEL with a manual .debug_line entry when the assembler supports .file and .loc directives. llvm-svn: 42278	2007-09-24 21:54:14 +00:00
Chris Lattner	594d3aa066	claim that "st" is from the 80-bit register file. This causes x87-using inline asm to die with: ScheduleDAG.cpp:269: failed assertion `false && "Couldn't find the register class"' instead of: failed assertion `RegMap->getRegClass(VReg) == RC && "Register class of operand and regclass of use don't agree!"' yay. llvm-svn: 42259	2007-09-24 05:27:37 +00:00
Dale Johannesen	ea6ffa0b36	Fix PR 1681. When X86 target uses +sse -sse2, keep f32 in SSE registers and f64 in x87. This is effectively a new codegen mode. Change addLegalFPImmediate to permit float and double variants to do different things. Adjust callers. llvm-svn: 42246	2007-09-23 14:52:20 +00:00
Rafael Espindola	11ee0898b9	Don't add a default STACK_ALIGN (use the generic ABI alignment) Implement calls to functions with byval arguments on X86 llvm-svn: 42192	2007-09-21 15:50:22 +00:00
Rafael Espindola	b0b536b597	small cleanup: use LowerMemArgument in LowerFastCCArguments also llvm-svn: 42189	2007-09-21 14:55:38 +00:00
Dale Johannesen	04682bdc81	More long double fixes. x86_64 should build now. llvm-svn: 42155	2007-09-19 23:55:34 +00:00
Dan Gohman	1aeaeec570	Emit integer x<1 as x<=0, as comparisons with zero (now includeing 64-bit) can use test instead of cmp with an immediate. llvm-svn: 42026	2007-09-17 14:49:27 +00:00
Dale Johannesen	575bd6070a	Remove the assumption that FP's are either float or double from some of the many places in the optimizers it appears, and do something reasonable with x86 long double. Make APInt::dump() public, remove newline, use it to dump ConstantSDNode's. Allow APFloats in FoldingSet. Expand X86 backend handling of long doubles (conversions to/from int, mostly). llvm-svn: 41967	2007-09-14 22:26:36 +00:00
Rafael Espindola	5d8b225881	Add support for functions with byval arguments on x86 llvm-svn: 41953	2007-09-14 15:48:13 +00:00
Dale Johannesen	7bc3969cea	Add APInt interfaces to APFloat (allows directly access to bits). Use them in place of float and double interfaces where appropriate. First bits of x86 long double constants handling (untested, probably does not work). llvm-svn: 41858	2007-09-11 18:32:33 +00:00
Duncan Sands	c358890f73	Fold the adjust_trampoline intrinsic into init_trampoline. There is now only one trampoline intrinsic. llvm-svn: 41841	2007-09-11 14:10:23 +00:00
Dale Johannesen	86f367a6b7	Next round of APFloat changes. Use APFloat in UpgradeParser and AsmParser. Change all references to ConstantFP to use the APFloat interface rather than double. Remove the ConstantFP double interfaces. Use APFloat functions for constant folding arithmetic and comparisons. (There are still way too many places APFloat is just a wrapper around host float/double, but we're getting there.) llvm-svn: 41747	2007-09-06 18:13:44 +00:00
Anton Korobeynikov	cf91be2c79	Reapply r41578 with proper fix llvm-svn: 41680	2007-09-03 00:36:06 +00:00
Rafael Espindola	4ddaad4de0	Initial support for calling functions with byval arguments on x86-64 llvm-svn: 41643	2007-08-31 15:06:30 +00:00
Dale Johannesen	81d6ecb886	Enhance APFloat to retain bits of NaNs (fixes oggenc). Use APFloat interfaces for more references, mostly of ConstantFPSDNode. llvm-svn: 41632	2007-08-31 04:03:46 +00:00
Dale Johannesen	e91a908971	Change LegalFPImmediates to use APFloat. Add APFloat interfaces to ConstantFP, SelectionDAG. Fix integer bit in double->APFloat conversion. Convert LegalizeDAG to use APFloat interface in ConstantFPSDNode uses. llvm-svn: 41587	2007-08-30 00:23:21 +00:00
Duncan Sands	26ef2a1767	Move getX86RegNum into X86RegisterInfo and use it in the trampoline lowering. Lookup the jump and mov opcodes for the trampoline rather than hard coding them. llvm-svn: 41577	2007-08-29 19:01:20 +00:00
Rafael Espindola	dc5450f7fb	Add a comment about using libc memset/memcpy or generating inline code. llvm-svn: 41502	2007-08-27 17:48:26 +00:00
Rafael Espindola	3d52fe3ef3	call libc memcpy/memset if array size is bigger then threshold. Coping 100MB array (after a warmup) shows that glibc 2.6.1 implementation on x86-64 (core 2) is 30% faster (from 0.270917s to 0.188079s) llvm-svn: 41479	2007-08-27 10:18:20 +00:00
Chris Lattner	1e089aac3a	rename isOperandValidForConstraint to LowerAsmOperandForConstraint, changing the interface to allow for future changes. llvm-svn: 41384	2007-08-25 00:47:38 +00:00
Rafael Espindola	68d95ff2b1	Partial implementation of calling functions with byval arguments: ) The needed information is propagated to the DAG ) The X86-64 backend detects it and aborts llvm-svn: 41179	2007-08-20 15:18:24 +00:00
Anton Korobeynikov	3094846993	Move ReturnAddrIndex variable to X86MachineFunctionInfo structure. This fixed hard to catch bugs with retaddr lowering llvm-svn: 41104	2007-08-15 17:12:32 +00:00
Evan Cheng	eef13203e7	Fix a typo pointd out by Maarten ter Huurne. llvm-svn: 41059	2007-08-13 23:27:11 +00:00
Christopher Lamb	450f6815b9	Increase efficiency of sign_extend_inreg by using subregisters for truncation. As the README suggests sign_extend_subreg is selected to (sext(trunc)). llvm-svn: 41010	2007-08-10 21:48:46 +00:00
Rafael Espindola	b20b9e985a	propagate struct size and alignment of byval arguments to the DAG llvm-svn: 40986	2007-08-10 14:44:42 +00:00
Dale Johannesen	79551baaad	long double 9 of N. This finishes up the X86-32 bits (constants are still not handled). Adds ConvertActions to control fp-to-fp conversions (these are currently defaulted for all other targets, so no changes there). llvm-svn: 40958	2007-08-09 01:04:01 +00:00
Dale Johannesen	2c35d56edd	Long double patch 7 of N, unless I lost count:). Last x87 bits for full functionality (not thoroughly tested, and long doubles do not work in SSE modes at all - use -mcpu=i486 for now) llvm-svn: 40886	2007-08-07 01:17:37 +00:00
Dale Johannesen	a85f11d870	Long double patch 4 of N: initial x87 implementation. Lots of problems yet but some simple things work. llvm-svn: 40847	2007-08-05 18:49:15 +00:00
Dan Gohman	1afde4166e	Fix the alignment requirements of several unpck and shuf instructions. Generalize isPSHUFDMask and add a unary SHUFPD pattern so that SHUFPD's memory operand alignment can be tested as well, with a fix to avoid breaking MMX's use of isPSHUFDMask. llvm-svn: 40756	2007-08-02 21:17:01 +00:00
Evan Cheng	019ecf3b91	Can't handle offset and scale if rip-relative addressing is to be used. llvm-svn: 40703	2007-08-01 23:46:47 +00:00
Evan Cheng	e90ad40aa1	This isn't safe when there are uses of load's chain result. llvm-svn: 40617	2007-07-31 06:21:44 +00:00
Duncan Sands	35a77d857b	Trampoline codegen support for X86-32. llvm-svn: 40566	2007-07-27 20:02:49 +00:00
Dan Gohman	0252aa07ee	Re-apply 40504, but with a fix for the segfault it caused in oggenc: Make the alignedload and alignedstore patterns always require 16-byte alignment. This way when they are used in the "Fs" instructions, in which a vector instruction is used for a scalar purpose, they can still require the full vector alignment. And add a regression test for this. llvm-svn: 40555	2007-07-27 17:16:43 +00:00
Evan Cheng	cb8f08ebca	Reverting 40504 for now. It's breaking oggenc. llvm-svn: 40547	2007-07-27 01:37:47 +00:00
Dan Gohman	513dcba4f8	Remove X86ISD::LOAD_PACK and X86ISD::LOAD_UA and associated code from the x86 target, replacing them with the new alignment attributes on memory references. llvm-svn: 40504	2007-07-26 00:31:09 +00:00
Dan Gohman	a2e07a38bc	Use movaps to load a v4f32 build_vector of all-constant values into a register instead of loading each element individually. llvm-svn: 40478	2007-07-24 22:55:08 +00:00
Dan Gohman	54b8032d64	Fix some uses of dyn_cast to be uses of cast. llvm-svn: 40443	2007-07-23 20:24:29 +00:00
Evan Cheng	ba990bbc3f	Fix custom lowering of SSE FXOR. llvm-svn: 40071	2007-07-19 23:36:01 +00:00
Anton Korobeynikov	5635277c36	Long live the exception handling! This patch fills the last necessary bits to enable exceptions handling in LLVM. Currently only on x86-32/linux. In fact, this patch adds necessary intrinsics (and their lowering) which represent really weird target-specific gcc builtins used inside unwinder. After corresponding llvm-gcc patch will land (easy) exceptions should be more or less workable. However, exceptions handling support should not be thought as 'finished': I expect many small and not so small glitches everywhere. llvm-svn: 39855	2007-07-14 14:06:15 +00:00
Dan Gohman	928144b051	Define non-intrinsic instructions for vector min, max, sqrt, rsqrt, and rcp, in addition to the intrinsic forms. Add spill-folding entries for these new instructions, and for the scalar min and max instrinsic instructions which were missing. And add some preliminary ISelLowering code for using the new non-intrinsic vector sqrt instruction, and fneg and fabs. llvm-svn: 38478	2007-07-10 00:05:58 +00:00
Anton Korobeynikov	e8215d1780	Proper flag __alloca call llvm-svn: 37923	2007-07-05 20:36:08 +00:00
Dale Johannesen	9072b65b0b	Refactor X87 instructions. As a side effect, all their names are changed. llvm-svn: 37876	2007-07-04 21:07:47 +00:00
Dale Johannesen	7af19491d3	Fix for PR 1505 (and 1489). Rewrite X87 register model to include f32 variants. Some factoring improvments forthcoming. llvm-svn: 37847	2007-07-03 00:53:03 +00:00
Evan Cheng	992f296a71	No vector fneg. llvm-svn: 37786	2007-06-29 00:18:15 +00:00
Evan Cheng	e233ec5e46	Type of vector extract / insert index operand should be iPTR. llvm-svn: 37784	2007-06-29 00:01:20 +00:00
Dan Gohman	354f02e03d	Generalize MVT::ValueType and associated functions to be able to represent extended vector types. Remove the special SDNode opcodes used for pre-legalize vector operations, and the special MVT::Vector type used with them. Adjust lowering and legalize to work with the normal SDNode kinds instead, and to use the normal MVT functions to work with vector types instead of using the two special operands that the pre-legalize nodes held. This allows pre-legalize and post-legalize DAGs, and the code that operates on them, to be more consistent. Pre-legalize vector operators can be handled more consistently with scalar operators. And, -view-dag-combine1-dags and -view-legalize-dags now look prettier for vector code. llvm-svn: 37719	2007-06-25 16:23:39 +00:00
Dan Gohman	a62327ea40	Move ComputeMaskedBits, MaskedValueIsZero, and ComputeNumSignBits from TargetLowering to SelectionDAG so that they have more convenient access to the current DAG, in preparation for the ValueType routines being changed from standalone functions to members of SelectionDAG for the pre-legalize vector type changes. llvm-svn: 37704	2007-06-22 14:59:07 +00:00
Chris Lattner	e13fac05d7	If a function is vararg, never pass inreg arguments in registers. Thanks to Anton for half of this patch. llvm-svn: 37641	2007-06-19 00:13:10 +00:00
Evan Cheng	80f0d5ae45	Look for VECTOR_SHUFFLE that's identity operation on either LHS or RHS. This can happen before DAGCombiner catches it. llvm-svn: 37636	2007-06-19 00:02:56 +00:00
Bill Wendling	94f3474832	Revert patch. It regresses: define double @test2(i64 %A) { %B = bitcast i64 %A to double ret double %B } $ llvm-as < t.ll \| llc -march=x86-64 before: .align 4 .globl _test2 _test2: movd %rdi, %xmm0 ret after: _test2: subq $8, %rsp movq %rdi, (%rsp) movsd (%rsp), %xmm0 addq $8, %rsp ret llvm-svn: 37617	2007-06-16 23:57:15 +00:00
Bill Wendling	a1f8f0aa97	Fix a failure to bit_convert from integer GPR to MMX register. llvm-svn: 37611	2007-06-16 06:17:31 +00:00
Dan Gohman	2fd7d26df8	Rename MVT::getVectorBaseType to MVT::getVectorElementType. llvm-svn: 37579	2007-06-14 22:58:02 +00:00
Chris Lattner	5f85da00bb	fix x86-64 mmx calling convention for real, which passes in integer gprs. llvm-svn: 37534	2007-06-09 05:08:10 +00:00
Chris Lattner	e965432273	fix mmx handling bug llvm-svn: 37533	2007-06-09 05:01:50 +00:00
Dan Gohman	1b1932dda5	Add explicit qualification for namespace MVT members. llvm-svn: 37320	2007-05-24 14:33:05 +00:00
Dan Gohman	ec87afe526	Use MVT::FIRST_VECTOR_VALUETYPE and MVT::LAST_VECTOR_VALUETYPE. llvm-svn: 37234	2007-05-18 18:44:07 +00:00
Evan Cheng	1b4af5f975	Fix a bogus check that prevented folding VECTOR_SHUFFLE to UNDEF; add an optimization to fold VECTOR_SHUFFLE to a zero vector. llvm-svn: 37173	2007-05-17 18:45:50 +00:00
Chris Lattner	9a53871650	This is the correct fix for PR1427. This fixes mmx-shuffle.ll and doesn't cause other regressions. llvm-svn: 37160	2007-05-17 17:13:13 +00:00
Anton Korobeynikov	375cafc275	Revert patch for PR1427. It breaks almost all vector tests. llvm-svn: 37159	2007-05-17 07:50:14 +00:00
Chris Lattner	f65fe1d931	Fix PR1427 and test/CodeGen/X86/mmx-shuffle.ll llvm-svn: 37141	2007-05-17 03:29:42 +00:00
Chris Lattner	ce20a357f1	fix subtle bugs in inline asm operand selection llvm-svn: 37065	2007-05-15 01:28:08 +00:00
Chris Lattner	60cd08c23e	Fix two classes of bugs: 1. x86 backend rejected (&gv+c) for the 'i' constraint when in static mode. 2. the matcher didn't correctly reject and accept some global addresses. the right predicate is GVRequiresExtraLoad, not "relomodel = pic". llvm-svn: 36670	2007-05-03 16:52:29 +00:00
Anton Korobeynikov	44aa4c588b	Emit correct register move information in eh frames for X86. This allows Shootout-C++/except to pass on x86/linux with non-llvm-compiled (e.g. "native") unwind runtime. llvm-svn: 36647	2007-05-02 19:53:33 +00:00
Bill Wendling	6856e741fa	Support for the special case of a vector with the canonical form: vector_shuffle v1, v2, <2, 6, 3, 7> I.e. vector_shuffle v, undef, <2, 2, 3, 3> MMX only has a shuffle for v4i16 vectors. It needs to use the unpackh for this type of operation. llvm-svn: 36403	2007-04-24 21:16:55 +00:00
Lauro Ramos Venancio	b1a101f0e7	X86 TLS: fix and optimize the implementation of "initial exec" model. llvm-svn: 36355	2007-04-22 22:50:52 +00:00
Lauro Ramos Venancio	3b60b9546e	X86 TLS: Implement review feedback. llvm-svn: 36318	2007-04-21 20:56:26 +00:00
Lauro Ramos Venancio	bc32d90b46	Implement "general dynamic", "initial exec" and "local exec" TLS models for X86 32 bits. llvm-svn: 36283	2007-04-20 21:38:10 +00:00
Anton Korobeynikov	60de2ce283	Add comment llvm-svn: 36213	2007-04-17 19:34:00 +00:00
Chris Lattner	c7109ece27	rename X86FunctionInfo to X86MachineFunctionInfo to match the header file it is defined in. llvm-svn: 36196	2007-04-17 17:21:52 +00:00
Anton Korobeynikov	9bc4b792bf	Implemented correct stack probing on mingw/cygwin for dynamic alloca's. Also, fixed static case in presence of eax livin. This fixes PR331 PS: Why don't we still have push/pop instructions? :) llvm-svn: 36195	2007-04-17 09:20:00 +00:00
Anton Korobeynikov	f3e62a428a	Removed tabs everywhere except autogenerated & external files. Add make target for tabs checking. llvm-svn: 36146	2007-04-16 18:10:23 +00:00
Chris Lattner	2b6b79b896	Fix mmx paddq, add support for the 'y' register class, though it isn't tested. llvm-svn: 35940	2007-04-12 04:14:49 +00:00
Chris Lattner	3f9ff05309	remove some dead hooks llvm-svn: 35845	2007-04-09 23:31:19 +00:00
Chris Lattner	ae6e2c0ee5	remove some dead target hooks, subsumed by isLegalAddressingMode llvm-svn: 35840	2007-04-09 22:27:04 +00:00
Chris Lattner	de148c7887	move a bunch of register constraints from being handled by getRegClassForInlineAsmConstraint to being handled by getRegForInlineAsmConstraint. This allows us to let the llvm register allocator allocate, which gives us better code. For example, X86/2007-01-29-InlineAsm-ir.ll used to compile to: _run_init_process: subl $4, %esp movl %ebx, (%esp) xorl %ebx, %ebx movl $11, %eax movl %ebx, %ecx movl %ebx, %edx # InlineAsm Start push %ebx ; movl %ebx,%ebx ; int $0x80 ; pop %ebx # InlineAsm End Now we get: _run_init_process: xorl %ecx, %ecx movl $11, %eax movl %ecx, %edx # InlineAsm Start push %ebx ; movl %ecx,%ebx ; int $0x80 ; pop %ebx # InlineAsm End llvm-svn: 35804	2007-04-09 05:49:22 +00:00
Chris Lattner	b940a717ac	implement support for CodeGen/X86/inline-asm-x-scalar.ll:test3 - i32/i64 values used with x constraints. llvm-svn: 35803	2007-04-09 05:31:48 +00:00
Chris Lattner	e2d3bf8ecf	implement CodeGen/X86/inline-asm-x-scalar.ll llvm-svn: 35799	2007-04-09 05:11:28 +00:00
Chris Lattner	c0405a348d	implement the new addressing mode description hook. llvm-svn: 35521	2007-03-30 23:15:24 +00:00
Bill Wendling	e8eccb1684	Remove cruft I put in there... llvm-svn: 35394	2007-03-28 01:02:54 +00:00
Bill Wendling	1087888176	Unbreak mmx arithmetic. It was barfing trying to do v8i8 arithmetic. llvm-svn: 35392	2007-03-28 00:57:11 +00:00
Bill Wendling	d43819da2f	Fix so that pandn is emitted instead of an xor/and combo. Add integer comparison operators. llvm-svn: 35385	2007-03-27 20:22:40 +00:00
Bill Wendling	8065cc3173	Promote to v1i64 type... llvm-svn: 35353	2007-03-26 08:03:33 +00:00
Bill Wendling	a42484728c	Add support for the v1i64 type. This makes better code for this: #include <mmintrin.h> extern __m64 C; void baz(__v2si A, __v2si B) { *A = C; _mm_empty(); } We get this: _baz: call "L1$pb" "L1$pb": popl %eax movl L_C$non_lazy_ptr-"L1$pb"(%eax), %eax movq (%eax), %mm0 movl 4(%esp), %eax movq %mm0, (%eax) emms ret GCC gives us this: _baz: pushl %ebx call L3 "L00000000001$pb": L3: popl %ebx subl $8, %esp movl L_C$non_lazy_ptr-"L00000000001$pb"(%ebx), %eax movl (%eax), %edx movl 4(%eax), %ecx movl 16(%esp), %eax movl %edx, (%eax) movl %ecx, 4(%eax) emms addl $8, %esp popl %ebx ret llvm-svn: 35351	2007-03-26 07:53:08 +00:00
Chris Lattner	b19069959d	switch TargetLowering::getConstraintType to take the entire constraint, not just the first letter. No functionality change. llvm-svn: 35322	2007-03-25 02:14:49 +00:00
Chris Lattner	104e73382c	enforce the proper range for the i386 N constraint. llvm-svn: 35319	2007-03-25 01:57:35 +00:00
Bill Wendling	1bcad4c1cd	Support added for shifts and unpacking MMX instructions. llvm-svn: 35266	2007-03-22 18:42:45 +00:00
Dale Johannesen	44c0a5d545	repair x86 performance, dejagnu problems from previous change llvm-svn: 35245	2007-03-21 21:51:52 +00:00
Chris Lattner	59fe2be1c4	fix a warning llvm-svn: 35152	2007-03-19 00:39:32 +00:00
Devang Patel	2dabb16eac	Support 'I' inline asm constraint. llvm-svn: 35129	2007-03-17 00:13:28 +00:00
Bill Wendling	8ced23ee5a	And now support for MMX logical operations. llvm-svn: 35125	2007-03-16 09:44:46 +00:00
Bill Wendling	feaff80149	Multiplication support for MMX. llvm-svn: 35118	2007-03-15 21:24:36 +00:00
Evan Cheng	00edaa08b5	Under X86-64 large code model, do not emit 32-bit pc relative calls. llvm-svn: 35108	2007-03-14 22:11:11 +00:00
Evan Cheng	0eeb8b59eb	More flexible TargetLowering LSR hooks for testing whether an immediate is a legal target address immediate or scale. llvm-svn: 35073	2007-03-12 23:28:50 +00:00
Evan Cheng	4224fa3617	Stupid bug: SSE2 supports v2i64 add / sub. llvm-svn: 35070	2007-03-12 22:58:52 +00:00
Bill Wendling	236cfc4344	Adding more arithmetic operators to MMX. This is an almost exact copy of the addition. Please let me know if you have suggestions. llvm-svn: 35055	2007-03-10 09:57:05 +00:00
Bill Wendling	5fef3fd7e7	Added "padd*" support for MMX. Added MMX move stuff to X86InstrInfo so that moves, loads, etc. are recognized. llvm-svn: 35031	2007-03-08 22:09:11 +00:00
Anton Korobeynikov	85d6c1ebad	Refactoring of formal parameter flags. Enable properly use of zext/sext/aext stuff. llvm-svn: 35008	2007-03-07 16:25:09 +00:00
Bill Wendling	3c201ddd02	Properly support v8i8 and v4i16 types. It now converts them to v2i32 for load and stores. llvm-svn: 35002	2007-03-07 05:43:18 +00:00
Bill Wendling	a02d43fbbd	Add LOAD/STORE support for MMX. llvm-svn: 34978	2007-03-06 18:53:42 +00:00
Anton Korobeynikov	6da6c8c48b	Use new SDIselParamAttr enumeration. This removes "magick" constants from formal attributes' flags processing. llvm-svn: 34963	2007-03-06 08:12:33 +00:00
Evan Cheng	2fb461c1b5	X86-64 VACOPY needs custom expansion. va_list is a struct { i32, i32, i8, i8 }. llvm-svn: 34857	2007-03-02 23:16:35 +00:00
Anton Korobeynikov	7cec92bcd2	Simplify things llvm-svn: 34849	2007-03-02 21:50:27 +00:00
Chris Lattner	55dcf58453	argument lowering should copy from the vreg shadows of live-in arguments passed in registers, not directly from the pregs themselves. llvm-svn: 34838	2007-03-02 05:12:29 +00:00
Anton Korobeynikov	eaf27d276a	Ensure that fastcall'ed function is correctly mangled & stack is properly aligned llvm-svn: 34788	2007-03-01 16:29:22 +00:00
Chris Lattner	bcc44762bc	remove dead option llvm-svn: 34754	2007-02-28 18:39:53 +00:00
Chris Lattner	a66d550298	use high-level functions in CCState llvm-svn: 34739	2007-02-28 07:09:55 +00:00
Chris Lattner	3663b6e73a	make use of helper functions in CCState for analyzing formals and calls. llvm-svn: 34737	2007-02-28 07:00:42 +00:00
Chris Lattner	3762b44a0c	switch LowerFastCCCallTo over to using the new fastcall description. llvm-svn: 34734	2007-02-28 06:26:33 +00:00
Chris Lattner	a8dd712470	switch LowerFastCCArguments over to using the autogenerated Fastcall description. llvm-svn: 34733	2007-02-28 06:21:19 +00:00
Chris Lattner	3b16744840	rearrange code llvm-svn: 34731	2007-02-28 06:10:12 +00:00
Chris Lattner	023751c20b	remove fastcc (not fastcall) support llvm-svn: 34730	2007-02-28 06:05:16 +00:00
Chris Lattner	012066f78b	switch LowerCCCArguments over to using autogenerated CC. llvm-svn: 34729	2007-02-28 05:46:49 +00:00
Chris Lattner	6424f8e245	simplify sret handling llvm-svn: 34728	2007-02-28 05:39:26 +00:00
Chris Lattner	76147834d6	switch LowerCCCCallTo over to using an autogenerated callingconv llvm-svn: 34727	2007-02-28 05:31:48 +00:00
Chris Lattner	eef57fed6e	switch return value passing and the x86-64 calling convention information over to being autogenerated from the X86CallingConv.td file. llvm-svn: 34722	2007-02-28 04:55:35 +00:00
Chris Lattner	9117648533	switch x86-64 return value lowering over to using same mechanism as argument lowering uses. llvm-svn: 34657	2007-02-27 05:28:59 +00:00
Chris Lattner	11a1c2113c	Minor refactoring of CC Lowering interfaces llvm-svn: 34656	2007-02-27 05:13:54 +00:00
Chris Lattner	e34136f6d5	move CC Lowering stuff to its own public interface llvm-svn: 34655	2007-02-27 04:43:02 +00:00
Chris Lattner	cac44e283d	refactor x86-64 argument lowering yet again, this time eliminating templates, 'clients', etc, and adding CCValAssign instead. llvm-svn: 34654	2007-02-27 04:18:15 +00:00
Chris Lattner	7165ee9b6b	switch to smallvector llvm-svn: 34633	2007-02-26 07:59:53 +00:00
Chris Lattner	3fe1132dcd	initial hack at splitting the x86-64 calling convention info out from the mechanics that process it. I'm still not happy with this, but it's a step in the right direction. llvm-svn: 34631	2007-02-26 07:50:02 +00:00
Chris Lattner	d0c941c89e	the truncate must always be done, it's only the assert that is conditional. llvm-svn: 34628	2007-02-26 05:21:05 +00:00
Chris Lattner	2e7125dc74	in X86-64 CCC, i8/i16 arguments are already properly zext/sext'd on input. Capture this so that downstream zext/sext's are optimized out. This compiles: int test(short X) { return (int)X; } to: _test: movl %edi, %eax ret instead of: _test: movswl %di, %eax ret GCC produces this bizarre code: _test: movw %di, -12(%rsp) movswl -12(%rsp),%eax ret llvm-svn: 34623	2007-02-26 03:18:56 +00:00
Chris Lattner	ad14e21b97	Fix an X86-64 abi bug. We now compile: void foo(short); void bar(unsigned short A) { foo(A); } into: _bar: subq $8, %rsp movswl %di, %edi call _foo addq $8, %rsp ret instead of: _bar: subq $8, %rsp call _foo addq $8, %rsp ret Testcase here: test/CodeGen/X86/x86-64-shortint.ll llvm-svn: 34615	2007-02-25 23:10:46 +00:00
Chris Lattner	15c167cc61	fix CodeGen/X86/2007-02-25-FastCCStack.ll, a regression from my patch last night: fastcc returns should only go in XMM0 if we have SSE2 or above. llvm-svn: 34613	2007-02-25 22:23:46 +00:00
Chris Lattner	65ba08d627	fastcc functions that return double values now return them in xmm0 on x86-32. This implements CodeGen/X86/fp-stack-ret.ll:test[23] llvm-svn: 34592	2007-02-25 09:31:16 +00:00
Chris Lattner	e4ba88824d	allow vectors to be passed to stdcall/fastcall functions llvm-svn: 34590	2007-02-25 09:14:25 +00:00
Chris Lattner	fac0b30da0	move LowerRET into the 'Return Value Calling Convention Implementation' section of the file. llvm-svn: 34589	2007-02-25 09:12:39 +00:00
Chris Lattner	65d915a3b6	make all Lower*CallTo implementations use LowerCallResult to handle their result value stuff. This eliminates a bunch of duplicated code and now GetRetValueLocs is the sole place that decides where a value is returned. llvm-svn: 34588	2007-02-25 09:10:05 +00:00
Chris Lattner	423224a7b4	pass the calling convention into Lower*CallTo, instead of using ad-hoc flags. llvm-svn: 34587	2007-02-25 09:06:15 +00:00
Chris Lattner	8fa75c3ae8	factor a bunch of code out of LowerCCCCallTo into a new LowerCallResult function. This function now uses GetRetValueLocs to determine where the result values are located and concerns itself with how to pull the values out. llvm-svn: 34586	2007-02-25 08:59:22 +00:00
Chris Lattner	3bfbc23ccd	move some code around, pass in calling conv, even though it is unused llvm-svn: 34585	2007-02-25 08:29:00 +00:00
Chris Lattner	f119813ff4	simplify result value lowering by splitting the selection of where to return registers out from the logic of how to return them. This changes X86-64 to mark EAX live out when returning a 32-bit value, where before it marked RAX liveout. llvm-svn: 34582	2007-02-25 08:15:11 +00:00
Chris Lattner	bcce79717b	make void-return not a special case llvm-svn: 34579	2007-02-25 07:18:38 +00:00
Chris Lattner	d00fcb3277	eliminate a bunch more temporary vectors from X86 lowering. llvm-svn: 34578	2007-02-25 07:10:00 +00:00
Chris Lattner	f7eeef816d	eliminate temporary vectors created during X86 lowering. llvm-svn: 34577	2007-02-25 06:40:16 +00:00
Chris Lattner	6f25082e67	remove std::vector's in RET lowering. llvm-svn: 34576	2007-02-25 06:21:57 +00:00
Jim Laskey	b57ee1fc37	Simplify lowering and selection of exception ops. llvm-svn: 34488	2007-02-22 14:56:36 +00:00
Jim Laskey	6a937ad320	Support to provide exception and selector registers. llvm-svn: 34482	2007-02-21 22:54:50 +00:00
Evan Cheng	0e7be3c4e0	ELF / PIC requires GOT be in the EBX register during calls via PLT GOT pointer. Add implicit uses of EBX to calls to ensure liveintervalanalysis does not treat the GOT in EBX move as dead upon definition. This should fix PR1207. llvm-svn: 34470	2007-02-21 21:18:14 +00:00
Anton Korobeynikov	c469cbc2e7	Fixed uninitialized stuff inside LegalizeDAG. Fortunately, the only affected part is codegen of "memove" inside x86 backend. This fixes PR1144 llvm-svn: 33752	2007-02-01 08:39:52 +00:00
Nate Begeman	dc46021355	Finish off bug 680, allowing targets to custom lower frame and return address nodes. llvm-svn: 33636	2007-01-29 22:58:52 +00:00
Nick Lewycky	e788dc93d5	Fix compile error "jump to case label crosses initialization". What compiler are people using that accepts this code? llvm-svn: 33603	2007-01-28 15:39:16 +00:00
Anton Korobeynikov	611d5e2eda	Propagate changes from my local tree. This patch includes: 1. New parameter attribute called 'inreg'. It has meaning "place this parameter in registers, if possible". This is some generalization of gcc's regparm(n) attribute. It's currently used only in X86-32 backend. 2. Completely rewritten CC handling/lowering code inside X86 backend. Merged stdcall + c CCs and fastcall + fast CC. 3. Dropped CSRET CC. We cannot add struct return variant for each target-specific CC (e.g. stdcall + csretcc and so on). 4. Instead of CSRET CC introduced 'sret' parameter attribute. Setting in on first attribute has meaning 'This is hidden pointer to structure return. Handle it gently'. 5. Fixed small bug in llvm-extract + add new feature to FunctionExtraction pass, which relinks all internal-linkaged callees from deleted function to external linkage. This will allow further linking everything together. NOTEs: 1. Documentation will be updated soon. 2. llvm-upgrade should be improved to translate csret => sret. Before this, there will be some unexpected test fails. llvm-svn: 33597	2007-01-28 13:31:35 +00:00
Jim Laskey	23ed7d2625	Make LABEL a builtin opcode. llvm-svn: 33537	2007-01-26 14:34:52 +00:00
Evan Cheng	818c6bdfa2	Linux GOT indirect reference is only necessary in PIC mode. llvm-svn: 33441	2007-01-22 21:34:25 +00:00
Anton Korobeynikov	548b9af9c2	* PIC codegen for X86/Linux has been implemented * PIC-aware internal structures in X86 Codegen have been refactored * Visibility (default/weak) has been added * Docs fixes (external weak linkage, visibility, formatting) llvm-svn: 33136	2007-01-12 19:20:47 +00:00
Evan Cheng	df277336b8	- FCOPYSIGN custom lowering bug. Clear the sign bit of operand 0 first before or'ing in the sign bit of operand 1. - Tweaking: rather than left shift the sign bit, fp_extend operand 1 first before taking its sign bit if its type is smaller than that of operand 0. llvm-svn: 32932	2007-01-05 21:37:56 +00:00
Evan Cheng	c273c9c830	Typo llvm-svn: 32902	2007-01-05 08:32:24 +00:00
Evan Cheng	bcf3d2bd15	With SSE2, expand FCOPYSIGN to a series of SSE bitwise operations. llvm-svn: 32900	2007-01-05 07:55:56 +00:00
Reid Spencer	c2da5d3d97	Fix a comment that referred to the now defunct ubyte type. llvm-svn: 32840	2007-01-03 17:24:59 +00:00
Anton Korobeynikov	2b39939053	Really big cleanup. - New target type "mingw" was introduced - Same things for both mingw & cygwin are marked as "cygming" (as in gcc) - .lcomm is supported here, so allow LLVM to use it - Correctly use underscored versions of setjmp & _longjmp for both mingw & cygwin llvm-svn: 32833	2007-01-03 11:43:14 +00:00
Reid Spencer	dda168599d	For PR950: Three changes: 1. Convert signed integer types to signless versions. 2. Implement the @sext and @zext parameter attributes. Previously the type of an function parameter was used to determine whether it should be sign extended or zero extended before the call. This information is now communicated via the function type's parameter attributes. 3. The interface to LowerCallTo had to be changed in order to accommodate the parameter attribute information. Although it would have been convenient to pass in the FunctionType itself, there isn't always one present in the caller. Consequently, a signedness indication for the result type and for each parameter was provided for in the interface to this method. All implementations were changed to make the adjustment necessary. llvm-svn: 32788	2006-12-31 05:55:36 +00:00
Anton Korobeynikov	3a6faf0b96	Refactored JIT codegen for mingw32. Now we're using standart relocation type for distinguish JIT & non-JIT instead of "dirty" hacks :) llvm-svn: 32745	2006-12-22 22:29:05 +00:00
Evan Cheng	5effab79f3	f64 <-> i64 bit_convert using movq in 64-bit mode. llvm-svn: 32587	2006-12-14 21:55:39 +00:00
Anton Korobeynikov	e76b69846d	Cleaned setjmp/longjmp lowering interfaces. Now we're producing right code (both asm & cbe) for Mingw32 target. Removed autoconf checks for underscored versions of setjmp/longjmp. llvm-svn: 32415	2006-12-10 23:12:42 +00:00
Chris Lattner	6a9de21df5	If we have ScalarSSE, we can select bitconvert into single instructions. This compiles bitcast.ll:test3/test4 into: _test3: movd %xmm0, %eax ret _test4: movd %edi, %xmm0 ret llvm-svn: 32230	2006-12-05 18:45:06 +00:00
Chris Lattner	4ad68ab4d7	Fix PR1033 and CodeGen/X86/bitcast.ll, by expanding bitcast to a load/store pair. This could be better, readme entry pending. llvm-svn: 32228	2006-12-05 18:22:22 +00:00
Chris Lattner	7368984a3d	Fix typo noticed by Lauro Ramos Venancio, thanks! llvm-svn: 32223	2006-12-05 17:29:40 +00:00
Evan Cheng	2c35691a02	- Fix X86-64 JIT by temporarily disabling code that treats GV address as 32-bit immediate in small code model. The JIT cannot ensure GV's are placed in the lower 4G. - Some preliminary support for large code model. llvm-svn: 32215	2006-12-05 04:01:03 +00:00
Chris Lattner	0be88afd90	Add support for 64-bit 'r' constraint. Patch by by Rafael ~Avila de EspÃndol! This fixes PR1029 and Regression/CodeGen/X86/x86-64-asm.ll llvm-svn: 32206	2006-12-04 22:38:21 +00:00
Evan Cheng	456101ebb9	- Use a different wrapper node for RIP-relative GV, etc. - Proper support for both small static and PIC modes under X86-64 - Some (non-optimal) support for medium modes. llvm-svn: 32046	2006-11-30 21:55:46 +00:00
Evan Cheng	f64e341522	In PIC mode, GV not requiring an extra load can be used as address immediate. llvm-svn: 32028	2006-11-29 23:48:14 +00:00
Evan Cheng	7e20347607	Fix for PR1018 - Better support for X86-64 Linux in small code model. llvm-svn: 32026	2006-11-29 23:19:46 +00:00
Evan Cheng	f968824cdb	Custom lower READCYCLECOUNTER for x86-64. llvm-svn: 32017	2006-11-29 08:28:13 +00:00
Evan Cheng	98fa7ab4d7	Change MachineInstr ctor's to take a TargetInstrDescriptor reference instead of opcode and number of operands. llvm-svn: 31947	2006-11-27 23:37:22 +00:00
Anton Korobeynikov	ddb73b4f1e	Refactored *GVRequiresExtraLoad() to Subtarget method. llvm-svn: 31887	2006-11-21 00:01:06 +00:00
Anton Korobeynikov	2449b29bdd	We should perform extra load for dllimported calls llvm-svn: 31874	2006-11-20 10:46:14 +00:00
Evan Cheng	0e82270ff2	Matches MachineInstr changes. llvm-svn: 31712	2006-11-13 23:36:35 +00:00
Evan Cheng	b9e2ae9e37	Add implicit use / def operands to created MI's. llvm-svn: 31676	2006-11-11 10:21:44 +00:00
Evan Cheng	ae1f3758bd	Don't dag combine floating point select to max and min intrinsics. Those take v4f32 / v2f64 operands and may end up causing larger spills / restores. Added X86 specific nodes X86ISD::FMAX, X86ISD::FMIN instead. This fixes PR996. llvm-svn: 31645	2006-11-10 21:43:37 +00:00
Anton Korobeynikov	23ffdb1971	Fixing PR990: http://llvm.org/PR990 . This should unbreak csretcc on Linux & mingw targets. Several tests from llvm-test should be also restored (fftbench, bigfib). llvm-svn: 31613	2006-11-10 00:48:11 +00:00
Evan Cheng	7ca1f47a96	Fixed a bug which causes x86 be to incorrectly match shuffle v, undef, <2, ?, 3, ?> to movhlps It should match to unpckhps instead. Added proper matching code for shuffle v, undef, <2, 3, 2, 3> llvm-svn: 31519	2006-11-07 22:14:24 +00:00
Reid Spencer	4bafa71dc1	For PR786: Turn on -Wunused and -Wno-unused-parameter. Clean up most of the resulting fall out by removing unused variables. Remaining warnings have to do with unused functions (I didn't want to delete code without review) and unused variables in generated code. Maintainers should clean up the remaining issues when they see them. All changes pass DejaGnu tests and Olden. llvm-svn: 31380	2006-11-02 20:25:50 +00:00
Chris Lattner	def30d3eda	allow the address of a global to be used with the "i" constraint when in -static mode. This implements PR882. llvm-svn: 31326	2006-10-31 20:13:11 +00:00
Chris Lattner	3bed109ed9	handle "st" as "st(0)" llvm-svn: 31320	2006-10-31 19:42:44 +00:00
Anton Korobeynikov	e6ba8a819c	1. Clean up code due to changes in SwitchTo*Section(2) 2. Added partial debug support for mingw\cygwin targets (the same as Linux\ELF). Please note, that currently mingw\cygwin uses 'stabs' format for storing debug info by default, thus many (runtime) libraries has this information included. These formats shouldn't be mixed in one binary ('stabs' & 'DWARF'), otherwise binutils tools will be confused. llvm-svn: 31311	2006-10-31 08:31:24 +00:00
Reid Spencer	db06ed9156	Add debug support for X86/ELF targets (Linux). This allows llvm-gcc4 generated object modules to be debugged with gdb. Hopefully this helps pre-release debugging. llvm-svn: 31299	2006-10-30 22:32:30 +00:00
Evan Cheng	5766dd6455	All targets expand BR_JT for now. llvm-svn: 31294	2006-10-30 08:02:39 +00:00
Evan Cheng	090e9abaee	Fixed a significant bug where unpcklpd is incorrectly used to extract element 1 from a v2f64 value. llvm-svn: 31228	2006-10-27 21:08:32 +00:00
Evan Cheng	a1ce4523e5	Fix for PR968: expand vector sdiv, udiv, srem, urem. llvm-svn: 31220	2006-10-27 18:49:08 +00:00
Evan Cheng	1abe8bd233	During vector shuffle lowering, we sometimes commute a vector shuffle to try to match MOVL (movss, movsd, etc.). Don't forget to commute it back and try unpck* and shufp* if that doesn't pan out. llvm-svn: 31186	2006-10-25 21:49:50 +00:00
Evan Cheng	01529405c9	Remove -disable-x86-shuffle-opti llvm-svn: 31183	2006-10-25 20:48:19 +00:00
Chris Lattner	62a0f00312	Implement branch analysis/xform hooks required by the branch folding pass. llvm-svn: 31065	2006-10-20 17:42:20 +00:00
Evan Cheng	ca5eaf4020	Avoid getting into an infinite loop when -disable-x86-shuffle-opti is specified. llvm-svn: 30974	2006-10-16 06:36:00 +00:00
Evan Cheng	fe5bb5dbe6	Merge ISD::TRUNCSTORE to ISD::STORE. Switch to using StoreSDNode. llvm-svn: 30945	2006-10-13 21:14:26 +00:00
Evan Cheng	d07e2f081a	Some X86ISD::CMP were created with wrong ValueType's. llvm-svn: 30913	2006-10-12 19:12:56 +00:00
Evan Cheng	8f6c6b19e6	Don't convert to MOVLP if using shufps etc. may allow load folding. llvm-svn: 30847	2006-10-09 21:39:25 +00:00
Evan Cheng	d22f3dd3ed	Reflects ISD::LOAD / ISD::LOADX / LoadSDNode changes. llvm-svn: 30844	2006-10-09 20:57:25 +00:00
Evan Cheng	275825195a	Make use of getStore(). llvm-svn: 30759	2006-10-05 23:01:46 +00:00
Chris Lattner	7f98896c02	Lower some min/max idioms to minss/maxss when unsafe fp math is enabled. llvm-svn: 30748	2006-10-05 04:11:26 +00:00
Evan Cheng	a77dd83caf	Added option -disable-x86-shuffle-opti to disable X86 specific vector shuffle optimizations. llvm-svn: 30723	2006-10-04 18:33:38 +00:00
Chris Lattner	4ec6c298d8	Pattern match min/max nodes when we have sse. This implements CodeGen/X86/scalar_sse_minmax.ll llvm-svn: 30719	2006-10-04 06:57:07 +00:00
Evan Cheng	494e8e6971	Combine ISD::EXTLOAD, ISD::SEXTLOAD, ISD::ZEXTLOAD into ISD::LOADX. Add an extra operand to LOADX to specify the exact value extension type. llvm-svn: 30714	2006-10-04 00:56:09 +00:00
Chris Lattner	72276ae651	Fix PR933 and CodeGen/X86/2006-10-02-BoolRetCrash.ll llvm-svn: 30703	2006-10-03 17:18:42 +00:00
Chris Lattner	0501525c0e	silence warnings in release build llvm-svn: 30631	2006-09-27 18:29:38 +00:00
Chris Lattner	b826d65e01	Various random and minor code cleanups. llvm-svn: 30608	2006-09-26 03:57:53 +00:00
Nick Lewycky	401794f2a7	Fix compile error. llvm-svn: 30553	2006-09-21 02:08:31 +00:00
Anton Korobeynikov	59ef7e94eb	Adding codegeneration for StdCall & FastCall calling conventions llvm-svn: 30549	2006-09-20 22:03:51 +00:00
Anton Korobeynikov	7c2118575c	Added some eye-candy for Subtarget type checking Added X86 StdCall & FastCall calling conventions. Codegen will follow. llvm-svn: 30446	2006-09-17 20:25:45 +00:00
Anton Korobeynikov	b2b7c2f8b9	Small fixes for supporting dll* linkage types llvm-svn: 30441	2006-09-17 13:06:18 +00:00
Anton Korobeynikov	6e19f80688	Adding dllimport, dllexport and external weak linkage types. DLL* linkages got full (I hope) codegeneration support in C & both x86 assembler backends. External weak linkage added for future use, we don't provide any codegeneration, etc. support for it. llvm-svn: 30374	2006-09-14 18:23:27 +00:00
Chris Lattner	5a00b32eec	Turn X < 0 -> TEST X,X js llvm-svn: 30294	2006-09-13 17:04:54 +00:00
Chris Lattner	d2b6dcd4bf	The sense of this branch was inverted :( llvm-svn: 30293	2006-09-13 16:56:12 +00:00
Chris Lattner	806ef17e5b	Compile X > -1 -> text X,X; js dest This implements CodeGen/X86/jump_sign.ll. llvm-svn: 30283	2006-09-13 03:22:10 +00:00
Evan Cheng	dd52a60189	Reflects MachineConstantPoolEntry changes. llvm-svn: 30279	2006-09-12 21:04:05 +00:00
Evan Cheng	cfd7b147cf	X86ISD::CMP now produces a chain as well as a flag. Make that the chain operand of a conditional branch to allow load folding into CMP / TEST instructions. llvm-svn: 30241	2006-09-11 02:19:56 +00:00
Evan Cheng	15dd42884e	Committing X86-64 support. llvm-svn: 30177	2006-09-08 06:48:29 +00:00
Evan Cheng	da0f93c5af	- Identify a vector_shuffle that can be turned into an undef, e.g. shuffle V1, <undef>, <undef, undef, 4, 5> - Fix some suspicious logic into LowerVectorShuffle that cause less than optimal code by failing to identify MOVL (move to lowest element of a vector). llvm-svn: 30171	2006-09-08 01:50:06 +00:00
Chris Lattner	9c7673ffca	Eliminate X86ISD::TEST, using X86ISD::CMP instead. Match X86ISD::CMP patterns using test, which provides nice simplifications like: - movl %edi, %ecx - andl $2, %ecx - cmpl $0, %ecx + testl $2, %edi je LBB1_11 #cond_next90 There are a couple of dagiselemitter deficiencies that this exposes, they will be handled later. llvm-svn: 30156	2006-09-07 20:33:45 +00:00
Chris Lattner	f92824aa27	Revert this patch, the front-end has been fixed to make it unneccesary. llvm-svn: 29752	2006-08-17 18:43:24 +00:00
Chris Lattner	989f26b766	'g' is handled by the front-end. llvm-svn: 29751	2006-08-17 18:12:28 +00:00
Andrew Lenharth	e720e5c444	Fix handling of 'g'. Closes 883 llvm-svn: 29750	2006-08-17 17:50:12 +00:00
Andrew Lenharth	bea8a84eb1	Add the 'c' constraint as needed by the linux kernel llvm-svn: 29747	2006-08-17 16:07:50 +00:00
Andrew Lenharth	7fdc3bbbcf	Add support for S and D constraints, as needed to compile the linux kernel. llvm-svn: 29746	2006-08-17 15:35:43 +00:00
Chris Lattner	8ca6e82bce	Eliminate use of getNode that takes a vector. llvm-svn: 29614	2006-08-11 17:38:39 +00:00
Evan Cheng	6053206580	Match tablegen changes. llvm-svn: 29604	2006-08-11 09:08:15 +00:00
Evan Cheng	131c832304	Convert more calls of getNode() that takes a vector to pass in the start of an array. llvm-svn: 29601	2006-08-11 07:35:45 +00:00
Chris Lattner	7b1362fa52	Start eliminating temporary vectors used to create DAG nodes. Instead, pass in the start of an array and a count of operands where applicable. In many cases, the number of operands is known, so this static array can be allocated on the stack, avoiding the heap. In many other cases, a SmallVector can be used, which has the same benefit in the common cases. I updated a lot of code calling getNode that takes a vector, but ran out of time. The rest of the code should be updated, and these methods should be removed. We should also do the same thing to eliminate the methods that take a vector of MVT::ValueTypes. It would be extra nice to convert the dagiselemitter to avoid creating vectors for operands when calling getTargetNode. llvm-svn: 29566	2006-08-08 02:23:42 +00:00
Chris Lattner	26ff12f7f5	Fix PR850 and CodeGen/X86/2006-07-31-SingleRegClass.ll. The CFE refers to all single-register constraints (like "A") by their 16-bit name, even though the 8 or 32-bit version of the register may be needed. The X86 backend should realize what is going on and redecode the name back to its proper form. llvm-svn: 29420	2006-07-31 23:26:50 +00:00
Chris Lattner	b4165c39d7	Rename RelocModel::PIC to PIC_, to avoid conflicts with -DPIC. llvm-svn: 29307	2006-07-26 21:12:04 +00:00
Evan Cheng	ed1c019899	This opt is now handled in DAG combine. llvm-svn: 29243	2006-07-21 08:26:46 +00:00
Evan Cheng	56434b7578	A splat of a vector constant of all zero or all one is the vector constant. llvm-svn: 29234	2006-07-20 23:09:47 +00:00
Chris Lattner	d0202bbed3	Add information preventing several register class constraints from working. This implements PR828 and CodeGen/X86/2006-07-12-InlineAsmQConstraint.ll llvm-svn: 29118	2006-07-12 16:59:49 +00:00
Chris Lattner	b75fe307e1	Implement the inline asm 'A' constraint. This implements PR825 and CodeGen/X86/2006-07-10-InlineAsmAConstraint.ll llvm-svn: 29101	2006-07-11 02:54:03 +00:00
Evan Cheng	b5e6a4a74d	Fixed stack objects do not specify alignments, but their offsets are known. Use that information when doing the transformation to merge multiple loads into a 128-bit load. llvm-svn: 29090	2006-07-10 21:37:44 +00:00
Chris Lattner	18c77b92a9	Mark internal function static llvm-svn: 29085	2006-07-10 19:53:12 +00:00
Evan Cheng	1d48a494a2	X86 target specific DAG combine: turn build_vector (load x), (load x+4), (load x+8), (load x+12), <0, 1, 2, 3> to a single 128-bit load (aligned and unaligned). e.g. __m128 test(float a, float b, float c, float d) { return _mm_set_ps(d, c, b, a); } _test: movups 4(%esp), %xmm0 ret llvm-svn: 29042	2006-07-07 08:33:52 +00:00
Evan Cheng	801ea78096	Reorg. No functionality change. llvm-svn: 28999	2006-07-05 22:17:51 +00:00
Evan Cheng	db5c7909f5	Simplify X86CompilationCallback: always align to 16-byte boundary; don't save EAX/EDX if unnecessary. llvm-svn: 28910	2006-06-24 08:36:10 +00:00
Evan Cheng	bc79e5f0e4	Type of vector extract / insert index operand should be iPTR. llvm-svn: 28796	2006-06-15 08:14:54 +00:00
Evan Cheng	242ebc9ab3	Add argument registers to the end of call operand list (partial fix). llvm-svn: 28783	2006-06-14 18:17:40 +00:00
Evan Cheng	8f46dba83d	Minor compilation speed improvement. llvm-svn: 28736	2006-06-09 06:24:42 +00:00
Evan Cheng	2da9d803a4	Added X86FunctionInfo subclass of MachineFunction to record whether the function that is being lowered is forced to use FP. Currently this is only true for main() / Cygwin. llvm-svn: 28703	2006-06-06 23:30:24 +00:00
Evan Cheng	4966cad9b5	Typos llvm-svn: 28617	2006-06-01 05:53:27 +00:00
Evan Cheng	15c1f84762	Remove a warning llvm-svn: 28607	2006-06-01 00:30:39 +00:00
Evan Cheng	82db95cd32	Remove dead code. llvm-svn: 28581	2006-05-31 00:50:42 +00:00
Evan Cheng	de0f25081a	Change RET node to include signness information of the return values. i.e. RET chain, value1, sign1, value2, sign2, ... llvm-svn: 28510	2006-05-26 23:10:12 +00:00
Evan Cheng	25db1a52d2	Vector argument must be passed in memory location aligned on 16-byte boundary. llvm-svn: 28505	2006-05-26 20:37:47 +00:00
Evan Cheng	7f468901bb	Mac OS X ABI document lied. The first four XMM registers are used to pass vector arguments, not three. llvm-svn: 28504	2006-05-26 19:22:06 +00:00
Evan Cheng	629df0afb2	Minor update to make the code more clear llvm-svn: 28499	2006-05-26 18:39:59 +00:00
Evan Cheng	bde90b2732	Update more comments. llvm-svn: 28498	2006-05-26 18:37:16 +00:00
Evan Cheng	414c909954	Fix some comments. llvm-svn: 28497	2006-05-26 18:25:43 +00:00
Evan Cheng	a4b6f4749d	No need to handle illegal types. llvm-svn: 28496	2006-05-26 18:22:49 +00:00
Evan Cheng	f610c3e318	Consistency llvm-svn: 28488	2006-05-25 23:31:23 +00:00
Evan Cheng	fb5e64ff9e	Some clean up. llvm-svn: 28483	2006-05-25 22:38:31 +00:00
Evan Cheng	ee280ac5d0	Remove some dead code. llvm-svn: 28481	2006-05-25 22:25:52 +00:00
Evan Cheng	abb909e3bb	Build breakage. llvm-svn: 28475	2006-05-25 18:56:34 +00:00
Evan Cheng	bb17ad5ffa	Switch X86 over to a call-selection model where the lowering code creates the copyto/fromregs instead of making the X86ISD::CALL selection code create them. llvm-svn: 28463	2006-05-25 00:59:30 +00:00
Chris Lattner	ac9665d682	Fix file header comment llvm-svn: 28441	2006-05-23 23:20:42 +00:00
Evan Cheng	3ef176507d	Better way to check for vararg. llvm-svn: 28440	2006-05-23 21:08:24 +00:00
Evan Cheng	65c3f3f26b	Remove PreprocessCCCArguments and PreprocessFastCCArguments now that FORMAL_ARGUMENTS nodes include a token operand. llvm-svn: 28439	2006-05-23 21:06:34 +00:00
Chris Lattner	9aec97df10	Implement an annoying part of the Darwin/X86 abi: the callee of a struct return argument pops the hidden struct pointer if present, not the caller. For example, in this testcase: struct X { int D, E, F, G; }; struct X bar() { struct X a; a.D = 0; a.E = 1; a.F = 2; a.G = 3; return a; } void foo(struct X P) { P = bar(); } We used to emit: _foo: subl $28, %esp movl 32(%esp), %eax movl %eax, (%esp) call _bar addl $28, %esp ret _bar: movl 4(%esp), %eax movl $0, (%eax) movl $1, 4(%eax) movl $2, 8(%eax) movl $3, 12(%eax) ret This is correct on Linux/X86 but not Darwin/X86. With this patch, we now emit: _foo: subl $28, %esp movl 32(%esp), %eax movl %eax, (%esp) call _bar * addl $24, %esp ret _bar: movl 4(%esp), %eax movl $0, (%eax) movl $1, 4(%eax) movl $2, 8(%eax) movl $3, 12(%eax) * ret $4 For the record, GCC emits (which is functionally equivalent to our new code): _bar: movl 4(%esp), %eax movl $3, 12(%eax) movl $2, 8(%eax) movl $1, 4(%eax) movl $0, (%eax) ret $4 _foo: pushl %esi subl $40, %esp movl 48(%esp), %esi leal 16(%esp), %eax movl %eax, (%esp) call _bar subl $4, %esp movl 16(%esp), %eax movl %eax, (%esi) movl 20(%esp), %eax movl %eax, 4(%esi) movl 24(%esp), %eax movl %eax, 8(%esi) movl 28(%esp), %eax movl %eax, 12(%esi) addl $40, %esp popl %esi ret This fixes SingleSource/Benchmarks/CoyoteBench/fftbench with LLC and the JIT, and fixes the X86-backend portion of PR729. The CBE still needs to be updated. llvm-svn: 28438	2006-05-23 18:50:38 +00:00
Chris Lattner	6d6a4d0e49	CSRet allows varargs llvm-svn: 28409	2006-05-19 21:34:04 +00:00
Evan Cheng	d282cb8542	Should pass by reference. llvm-svn: 28357	2006-05-17 19:07:40 +00:00
Chris Lattner	c04371da56	Implement the custom lowering hook right, returning values for all of the arguments at once. llvm-svn: 28327	2006-05-16 17:14:26 +00:00
Chris Lattner	f501a979ec	Fix a bug I introduced yesterday, which broke functions with no arguments. llvm-svn: 28326	2006-05-16 17:08:35 +00:00
Evan Cheng	dc9b5f5fc0	X86 integer register classes naming changes. Make them consistent with FP, vector classes. llvm-svn: 28324	2006-05-16 07:21:53 +00:00
Chris Lattner	ba1dfc1da7	Add a chain to FORMAL_ARGUMENTS. This is a minimal port of the X86 backend, it doesn't currently use/maintain the chain properly. Also, make the X86ISelLowering.cpp file 80-col clean. llvm-svn: 28320	2006-05-16 06:45:34 +00:00
Chris Lattner	db8caed257	Dead variable llvm-svn: 28265	2006-05-12 21:12:22 +00:00
Chris Lattner	89fa42b51e	Teach the X86 backend about non-i32 inline asm register classes. llvm-svn: 28139	2006-05-06 00:29:37 +00:00
Chris Lattner	a03676690b	Teach the code generator to use cvtss2sd as extload f32 -> f64 llvm-svn: 28131	2006-05-05 21:35:18 +00:00
Owen Anderson	71bc529dfa	Refactor TargetMachine, pushing handling of TargetData into the target-specific subclasses. This has one caller-visible change: getTargetData() now returns a pointer instead of a reference. This fixes PR 759. llvm-svn: 28074	2006-05-03 01:29:57 +00:00
Evan Cheng	a33feb51db	Initial caller side support (for CCC only, not FastCC) of 128-bit vector passing by value. llvm-svn: 28015	2006-04-28 21:29:37 +00:00
Evan Cheng	d577ce4c4a	Implement four-wide shuffle with 2 shufps if no more than two elements come from each vector. e.g. shuffle(G1, G2, 7, 1, 5, 2) ==> movaps _G2, %xmm0 shufps $151, _G1, %xmm0 shufps $216, %xmm0, %xmm0 llvm-svn: 28011	2006-04-28 07:03:38 +00:00
Evan Cheng	f843942504	TargetLowering::LowerArguments should return a VBIT_CONVERT of FORMAL_ARGUMENTS SDOperand in the return result vector. llvm-svn: 28009	2006-04-28 05:25:15 +00:00
Evan Cheng	11e3cec8bd	Make x86 isel lowering produce tailcall nodes. They are match to normal calls for now. Patch contributed by Alexander Friedman. llvm-svn: 27994	2006-04-27 08:40:39 +00:00
Evan Cheng	24795120e1	Support for passing 128-bit vector arguments via XMM registers. llvm-svn: 27992	2006-04-27 08:31:10 +00:00
Evan Cheng	1e065ae594	Oops llvm-svn: 27989	2006-04-27 05:44:50 +00:00
Evan Cheng	a0e0eabc07	Bug fix: not updating NumIntRegs. llvm-svn: 27988	2006-04-27 05:35:28 +00:00
Evan Cheng	a1f9f34f35	- Clean up formal argument lowering code. Prepare for vector pass by value work. - Fixed vararg support. llvm-svn: 27985	2006-04-27 01:32:22 +00:00
Evan Cheng	3abec16563	Fix fastcc failures. llvm-svn: 27980	2006-04-26 18:21:31 +00:00
Evan Cheng	58d4133b60	Switching over FORMAL_ARGUMENTS mechanism to lower call arguments. llvm-svn: 27975	2006-04-26 01:20:17 +00:00
Evan Cheng	09112df9d3	Separate LowerOperation() into multiple functions, one per opcode. llvm-svn: 27972	2006-04-25 20:13:52 +00:00
Evan Cheng	0282b48ec2	Special case handling two wide build_vector(0, x). llvm-svn: 27961	2006-04-24 22:58:52 +00:00
Evan Cheng	1eae7398a6	A little bit more build_vector enhancement for v8i16 cases. llvm-svn: 27959	2006-04-24 18:01:45 +00:00
Evan Cheng	4812ce5035	MOVL shuffle (i.e. movd or movss / movsd from memory) of undef, V2 == V2 llvm-svn: 27953	2006-04-23 06:35:19 +00:00
Nate Begeman	7ed816f900	JumpTable support! What this represents is working asm and jit support for x86 and ppc for 100% dense switch statements when relocations are non-PIC. This support will be extended and enhanced in the coming days to support PIC, and less dense forms of jump tables. llvm-svn: 27947	2006-04-22 18:53:45 +00:00
Evan Cheng	1c33e83af5	Don't do all the lowering stuff for 2-wide build_vector's. Also, minor optimization for shuffle of undef. llvm-svn: 27946	2006-04-22 08:34:05 +00:00
Evan Cheng	ec33bd04fb	Fix a performance regression. Use {p}shuf* when there are only two distinct elements in a build_vector. llvm-svn: 27945	2006-04-22 06:21:46 +00:00
Evan Cheng	5cb5fdd8eb	Revamp build_vector lowering to take advantage of movss and movd instructions. movd always clear the top 96 bits and movss does so when it's loading the value from memory. The net result is codegen for 4-wide shuffles is much improved. It is near optimal if one or more elements is a zero. e.g. __m128i test(int a, int b) { return _mm_set_epi32(0, 0, b, a); } compiles to _test: movd 8(%esp), %xmm1 movd 4(%esp), %xmm0 punpckldq %xmm1, %xmm0 ret compare to gcc: _test: subl $12, %esp movd 20(%esp), %xmm0 movd 16(%esp), %xmm1 punpckldq %xmm0, %xmm1 movq %xmm1, %xmm0 movhps LC0, %xmm0 addl $12, %esp ret or icc: _test: movd 4(%esp), %xmm0 #5.10 movd 8(%esp), %xmm3 #5.10 xorl %eax, %eax #5.10 movd %eax, %xmm1 #5.10 punpckldq %xmm1, %xmm0 #5.10 movd %eax, %xmm2 #5.10 punpckldq %xmm2, %xmm3 #5.10 punpckldq %xmm3, %xmm0 #5.10 ret #5.10 There are still room for improvement, for example the FP variant of the above example: __m128 test(float a, float b) { return _mm_set_ps(0.0, 0.0, b, a); } _test: movss 8(%esp), %xmm1 movss 4(%esp), %xmm0 unpcklps %xmm1, %xmm0 xorps %xmm1, %xmm1 movlhps %xmm1, %xmm0 ret The xorps and movlhps are unnecessary. This will require post legalizer optimization to handle. llvm-svn: 27939	2006-04-21 23:03:30 +00:00
Evan Cheng	e0289de5ab	Now generating perfect (I think) code for "vector set" with a single non-zero scalar value. e.g. _mm_set_epi32(0, a, 0, 0); ==> movd 4(%esp), %xmm0 pshufd $69, %xmm0, %xmm0 _mm_set_epi8(0, 0, 0, 0, 0, a, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0); ==> movzbw 4(%esp), %ax movzwl %ax, %eax pxor %xmm0, %xmm0 pinsrw $5, %eax, %xmm0 llvm-svn: 27923	2006-04-21 01:05:10 +00:00
Evan Cheng	41f2933444	- Added support to turn "vector clear elements", e.g. pand V, <-1, -1, 0, -1> to a vector shuffle. - VECTOR_SHUFFLE lowering change in preparation for more efficient codegen of vector shuffle with zero (or any splat) vector. llvm-svn: 27875	2006-04-20 08:58:49 +00:00
Evan Cheng	9dcd046bbd	Handle v2i64 BUILD_VECTOR custom lowering correctly. v2i64 is a legal type, but i64 is not. If possible, change a i64 op to a f64 (e.g. load, constant) and then cast it back. llvm-svn: 27849	2006-04-20 00:11:39 +00:00
Evan Cheng	d79f6a9f5a	isSplatMask() bug: first element can be an undef. llvm-svn: 27847	2006-04-19 23:28:59 +00:00
Evan Cheng	019dea6886	- Added support to do aribitrary 4 wide shuffle with no more than three instructions. - Fixed a commute vector_shuff bug. llvm-svn: 27845	2006-04-19 22:48:17 +00:00
Evan Cheng	265831aa45	Commute vector_shuffle to match more movlhps, movlp{s\|d} cases. llvm-svn: 27840	2006-04-19 20:35:22 +00:00
Evan Cheng	98b1ca65dd	Use movss to insert_vector_elt(v, s, 0). llvm-svn: 27782	2006-04-17 22:45:49 +00:00
Evan Cheng	ecf13c5d79	Use two pinsrw to insert an element into v4i32 / v4f32 vector. llvm-svn: 27779	2006-04-17 22:04:06 +00:00
Evan Cheng	4de1805c84	Implement v8i16, v16i8 splat using unpckl + pshufd. llvm-svn: 27768	2006-04-17 20:43:08 +00:00
Chris Lattner	e1d38ad84b	implement returns of a vector, testcase here: CodeGen/X86/vec_return.ll llvm-svn: 27767	2006-04-17 20:32:50 +00:00
Evan Cheng	eb739d0355	FP SETOLT, SETOLT, SETUGE, SETUGT conditions were implemented incorrectly llvm-svn: 27755	2006-04-17 07:24:10 +00:00
Evan Cheng	32e5d4f6bc	Silly bug llvm-svn: 27719	2006-04-15 05:37:34 +00:00
Evan Cheng	f9a93a1d3f	Do not use movs{h\|l}dup for a shuffle with a single non-undef node. llvm-svn: 27718	2006-04-15 03:13:24 +00:00
Evan Cheng	32c4470374	Last few SSE3 intrinsics. llvm-svn: 27711	2006-04-14 21:59:03 +00:00
Evan Cheng	25fcfb9f2d	X86 SSE2 supports v8i16 multiplication llvm-svn: 27644	2006-04-13 05:10:25 +00:00
Evan Cheng	2c2d734efd	All "integer" logical ops (pand, por, pxor) are now promoted to v2i64. Clean up and fix various logical ops issues. llvm-svn: 27633	2006-04-12 21:21:57 +00:00
Evan Cheng	66fb7beed7	Promote v4i32, v8i16, v16i8 load to v2i64 load. llvm-svn: 27612	2006-04-12 17:12:36 +00:00
Evan Cheng	da283be867	Added support for _mm_move_ss and _mm_move_sd. llvm-svn: 27575	2006-04-11 00:19:04 +00:00
Evan Cheng	2b6c899eb2	Conditional move of vector types. llvm-svn: 27556	2006-04-10 07:23:14 +00:00
Evan Cheng	281a7abddf	Code clean up. llvm-svn: 27501	2006-04-07 21:53:05 +00:00
Evan Cheng	9f27046dc9	- movlp{s\|d} and movhp{s\|d} support. - Normalize shuffle nodes so result vector lower half elements come from the first vector, the rest come from the second vector. (Except for the exceptions :-). - Other minor fixes. llvm-svn: 27474	2006-04-06 23:23:56 +00:00
Evan Cheng	6d470008c8	Support for comi / ucomi intrinsics. llvm-svn: 27444	2006-04-05 23:38:46 +00:00
Evan Cheng	056e0af55a	Handle canonical form of e.g. vector_shuffle v1, v1, <0, 4, 1, 5, 2, 6, 3, 7> This is turned into vector_shuffle v1, <undef>, <0, 0, 1, 1, 2, 2, 3, 3> by dag combiner. It would match a {p}unpckl on x86. llvm-svn: 27437	2006-04-05 07:20:06 +00:00
Evan Cheng	d562dfa0db	Bogus assert llvm-svn: 27434	2006-04-05 06:11:20 +00:00
Evan Cheng	9e56e97205	Fallthrough to expand if a VECTOR_SHUFFLE cannot be custom lowered. llvm-svn: 27433	2006-04-05 06:09:26 +00:00
Evan Cheng	849a726354	Handle v8i16 shuffle that must be broken into a pair of pshufhw / pshuflw. llvm-svn: 27427	2006-04-05 01:47:37 +00:00
Evan Cheng	7ff32cd571	Use movlpd to: store lower f64 extracted from v2f64. Use movhpd to: store upper f64 extracted from v2f64. llvm-svn: 27382	2006-04-03 22:30:54 +00:00
Evan Cheng	169240beb7	- More efficient extract_vector_elt with shuffle and movss, movsd, movd, etc. - Some bug fixes and naming inconsistency fixes. llvm-svn: 27377	2006-04-03 20:53:28 +00:00
Evan Cheng	4623ebd3d0	Use a X86 target specific node X86ISD::PINSRW instead of a mal-formed INSERT_VECTOR_ELT to insert a 16-bit value in a 128-bit vector. llvm-svn: 27314	2006-03-31 21:55:24 +00:00
Evan Cheng	7b9a0c6d7a	Add support to use pextrw and pinsrw to extract and insert a word element from a 128-bit vector. llvm-svn: 27304	2006-03-31 19:22:53 +00:00
Evan Cheng	4ca9bbc1bb	Expand all INSERT_VECTOR_ELT (obviously bad) for now. llvm-svn: 27275	2006-03-31 01:30:39 +00:00
Evan Cheng	5d9fc9fdd0	Typo llvm-svn: 27272	2006-03-31 00:33:57 +00:00
Evan Cheng	c55052da81	Ok for vector_shuffle mask to contain undef elements. llvm-svn: 27271	2006-03-31 00:30:29 +00:00
Evan Cheng	d3c692650f	Make sure all possible shuffles are matched. Use pshufd, pshuhw, and pshulw to shuffle v4f32 if shufps doesn't match. Use shufps to shuffle v4f32 if pshufd, pshuhw, and pshulw don't match. llvm-svn: 27259	2006-03-30 19:54:57 +00:00
Evan Cheng	7bc3bc8246	- Added some SSE2 128-bit packed integer ops. - Added SSE2 128-bit integer pack with signed saturation ops. - Added pshufhw and pshuflw ops. llvm-svn: 27252	2006-03-29 23:07:14 +00:00
Evan Cheng	d0d3eade59	Need to special case splat after all. Make the second operand of splat vector_shuffle undef. llvm-svn: 27250	2006-03-29 19:02:40 +00:00
Evan Cheng	02b5de9b3e	- More shuffle related bug fixes. - Whenever possible use ops of the right packed types for vector shuffles / splats. llvm-svn: 27246	2006-03-29 03:04:49 +00:00
Evan Cheng	5194a37602	- Only use pshufd for v4i32 vector shuffles. - Other shuffle related fixes. llvm-svn: 27244	2006-03-29 01:30:51 +00:00
Evan Cheng	178e36174a	Fixing buggy code. llvm-svn: 27239	2006-03-28 23:41:33 +00:00
Jim Laskey	fa6dfa9212	Added missing paren on behalf of Ramana Radhakrishnan. llvm-svn: 27223	2006-03-28 10:17:11 +00:00
Evan Cheng	0305cec743	Missed X86::isUNPCKHMask llvm-svn: 27222	2006-03-28 08:27:15 +00:00
Evan Cheng	fb4b2bfc7d	* Prefer using operation of matching types. e.g unpcklpd rather than movlhps. * Bug fixes. llvm-svn: 27218	2006-03-28 06:50:32 +00:00
Evan Cheng	4d554dae17	- Clean up / consoladate various shuffle masks. - Some misc. bug fixes. - Use MOVHPDrm to load from m64 to upper half of a XMM register. llvm-svn: 27210	2006-03-28 02:43:26 +00:00
Evan Cheng	d8d7ec47bd	Model unpack lower and interleave as vector_shuffle so we can lower the intrinsics as such. llvm-svn: 27200	2006-03-28 00:39:58 +00:00
Evan Cheng	0865274fa5	Use pcmpeq to generate vector of all ones. llvm-svn: 27167	2006-03-27 07:00:16 +00:00
Nate Begeman	3d518334b9	SelectionDAGISel can now natively handle Switch instructions, in the same manner that the LowerSwitch LLVM to LLVM pass does: emitting a binary search tree of basic blocks. The new approach has several advantages: it is faster, it generates significantly smaller code in many cases, and it paves the way for implementing dense switch tables as a jump table by handling switches directly in the instruction selector. This functionality is currently only enabled on x86, but should be safe for every target. In anticipation of making it the default, the cfg is now properly updated in the x86, ppc, and sparc select lowering code. llvm-svn: 27156	2006-03-27 01:32:24 +00:00
Evan Cheng	1dfbede1d1	Remove X86:isZeroVector, use ISD::isBuildVectorAllZeros instead; some fixes / cleanups llvm-svn: 27150	2006-03-26 09:53:12 +00:00
Evan Cheng	e5807f6b47	Build arbitrary vector with more than 2 distinct scalar elements with a series of unpack and interleave ops. llvm-svn: 27119	2006-03-25 09:37:23 +00:00
Evan Cheng	234090b386	Added 128-bit packed integer subtraction. llvm-svn: 27096	2006-03-25 01:33:37 +00:00
Evan Cheng	bdb85b387f	Support for scalar to vector with zero extension. llvm-svn: 27091	2006-03-24 23:15:12 +00:00
Evan Cheng	d58d54cf3e	Handle BUILD_VECTOR with all zero elements. llvm-svn: 27056	2006-03-24 07:29:27 +00:00
Chris Lattner	ace2d0d227	Gabor points out that we can't spell. :) llvm-svn: 27049	2006-03-24 07:12:19 +00:00
Evan Cheng	8507228441	All v2f64 shuffle cases can be handled. llvm-svn: 27044	2006-03-24 06:40:32 +00:00
Evan Cheng	3028b04057	More efficient v2f64 shuffle using movlhps, movhlps, unpckhpd, and unpcklpd. llvm-svn: 27040	2006-03-24 02:58:06 +00:00
Evan Cheng	68410804f0	Handle more shuffle cases with SHUFP* instructions. llvm-svn: 27024	2006-03-24 01:18:28 +00:00
Evan Cheng	daa75ed684	Typo llvm-svn: 26997	2006-03-23 20:26:04 +00:00
Evan Cheng	5f7cf963db	Add 128-bit integer vector load and add (for testing). llvm-svn: 26967	2006-03-23 01:57:24 +00:00
Evan Cheng	54215cd1ea	Added a ValueType operand to isShuffleMaskLegal(). For now, x86 will not do 64-bit vector shuffle. llvm-svn: 26964	2006-03-22 22:07:06 +00:00
Evan Cheng	7cb4e14749	Some clean up. llvm-svn: 26957	2006-03-22 19:22:18 +00:00
Evan Cheng	ae6a39ea92	- Supposely movlhps is faster / better than unpcklpd. - Don't forget pshufd is only available with sse2. llvm-svn: 26956	2006-03-22 19:16:21 +00:00
Evan Cheng	cff38e19c3	- Implement X86ISelLowering::isShuffleMaskLegal(). We currently only support splat and PSHUFD cases. - Clean up shuffle / splat matching code. llvm-svn: 26954	2006-03-22 18:59:22 +00:00
Evan Cheng	f6dc0a7f5e	- VECTOR_SHUFFLE of v4i32 / v4f32 with undef second vector always matches PSHUFD. We can make permutes entries which point to the undef pointing anything we want. - Change some names to appease Chris. llvm-svn: 26951	2006-03-22 08:01:21 +00:00
Chris Lattner	1554bf155e	fix a warning llvm-svn: 26941	2006-03-22 04:18:34 +00:00
Evan Cheng	7aac4350c7	Some splat and shuffle support. llvm-svn: 26940	2006-03-22 02:53:00 +00:00
Evan Cheng	47dd756c72	- Use movaps to store 128-bit vector integers. - Each scalar to vector v8i16 and v16i8 is a any_extend followed by a movd. llvm-svn: 26932	2006-03-21 23:01:21 +00:00
Chris Lattner	31a93c7740	These targets don't support EXTRACT_VECTOR_ELT, though, in time, X86 will. llvm-svn: 26930	2006-03-21 20:51:05 +00:00
Chris Lattner	09ede9ec9f	Add a build_vector node llvm-svn: 26895	2006-03-20 06:18:01 +00:00
Chris Lattner	1bd0aaf2b8	rename these nodes llvm-svn: 26848	2006-03-19 01:13:28 +00:00

... 5 6 7 8 9 ...

761 Commits