llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-11-22 18:54:02 +01:00

Author	SHA1	Message	Date
Simon Pilgrim	bf9d74ce23	[X86][SSE] Add clflush scheduling test llvm-svn: 316925	2017-10-30 17:20:50 +00:00
Jina Nahias	85b1b56f46	[X86][AVX512] Adding a pattern for broadcastm intrinsic. Differential Revision: https://reviews.llvm.org/D38312 Change-Id: I71c8605a8e4c98013ef25289694afc5cfd46bb0b llvm-svn: 316921	2017-10-30 16:37:28 +00:00
Rafael Espindola	6b525c8f51	Move isDSOLocal check and add a comment. llvm-svn: 316920	2017-10-30 16:32:31 +00:00
Fangrui Song	52eea02d9d	[PPC CodeGen] Fix the bitreverse.i64 intrinsic. Summary: The two 32-bit words were swapped. Update a test omitted in reverted r316270. Reviewers: jtony, aaron.ballman Subscribers: nemanjai, kbarton Differential Revision: https://reviews.llvm.org/D39163 llvm-svn: 316916	2017-10-30 16:03:44 +00:00
Craig Topper	403b0416bf	[X86] Make sure we don't create locked inc/dec instructions when the carry flag is being used. Summary: INC/DEC don't update the carry flag so we need to make sure we don't try to use it. This patch introduces new X86ISD opcodes for locked INC/DEC. Teaches lowerAtomicArithWithLOCK to emit these nodes if INC/DEC is not slow or the function is being optimized for size. An additional flag is added that allows the INC/DEC to be disabled if the caller determines that the carry flag is being requested. The test_sub_1_cmp_1_setcc_ugt test is currently showing this bug. The other test case changes are recovering cases that were regressed in r316860. This should fully fix PR35068 finishing the fix started in r316860. Reviewers: RKSimon, zvi, spatel Reviewed By: zvi Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D39411 llvm-svn: 316913	2017-10-30 14:51:37 +00:00
Craig Topper	eb5a0ad6c7	[X86] Remove AVX512 early out from X86FastISel::X86SelectCmp. This shouldn't be needed anymore since i1 isn't a legal type. llvm-svn: 316912	2017-10-30 14:50:11 +00:00
Craig Topper	ab3031ecea	[X86] Regenerate test using update_llc_test_checks.py llvm-svn: 316911	2017-10-30 14:50:10 +00:00
Sanjay Patel	2f425f444d	[PassManager, SimplifyCFG] add test for PR34603 / D38566; NFC Sinking common insts and converting to select early can inhibit better folds in other passes. llvm-svn: 316908	2017-10-30 14:34:30 +00:00
Yaxun Liu	3938c6fc0d	[AMDGPU] Emit metadata for hidden arguments for kernel enqueue Identifies kernels which performs device side kernel enqueues and emit metadata for the associated hidden kernel arguments. Such kernels are marked with calls-enqueue-kernel function attribute by AMDGPUOpenCLEnqueueKernelLowering pass and later on hidden kernel arguments metadata HiddenDefaultQueue and HiddenCompletionAction are emitted for them. Differential Revision: https://reviews.llvm.org/D39255 llvm-svn: 316907	2017-10-30 14:30:28 +00:00
Clement Courbet	32210b316a	[CodeGen][ExpandMemcmp] Allow memcmp to expand to vector loads (2). - Targets that want to support memcmp expansions now return the list of supported load sizes. - Expansion codegen does not assume that all power-of-two load sizes smaller than the max load size are valid. For examples, this is not the case for x86(32bit)+sse2. Fixes PR34887. llvm-svn: 316905	2017-10-30 14:19:33 +00:00
Krzysztof Parzyszek	84befe7d52	[Hexagon] Allow the RDF optimizations to be run in .mir testcases llvm-svn: 316904	2017-10-30 14:11:52 +00:00
Javed Absar	9816d74485	[GlobalISel\|ARM] : Allow legalizing G_FSUB Adding support for VSUB. Reviewed by: @rovka Differential Revision: https://reviews.llvm.org/D39261 llvm-svn: 316902	2017-10-30 13:51:56 +00:00
Andrew V. Tischenko	b7183e72ac	Invalid used of 'w' suffix on push and pop using 64-bit register. Differential Revision: https://reviews.llvm.org/D38626 llvm-svn: 316898	2017-10-30 12:02:06 +00:00
Diana Picus	661d30992e	[ARM GlobalISel] Fixup r316572. NFC Just missed a few spots... llvm-svn: 316897	2017-10-30 11:58:09 +00:00
Jina Nahias	2a382da3c2	Revert "[X86][AVX512] Adding a pattern for broadcastm intrinsic." This reverts commit r316890. Change-Id: I683cceee9848ef309b452293086b1f26a941950d llvm-svn: 316894	2017-10-30 10:35:53 +00:00
Florian Hahn	c686f0d8f2	Recommit r315288: [SCCP] Propagate integer range info for parameters in IPSCCP. This version of the patch includes a fix addressing a stage2 LTO buildbot failure and addressed some additional nits. Original commit message: This updates the SCCP solver to use of the ValueElement lattice for parameters, which provides integer range information. The range information is used to remove unneeded icmp instructions. For the following function, f() can be optimized to ret i32 2 with this change source_filename = "sccp.c" target datalayout = "e-m:e-i64:64-f80:128-n8:16:32:64-S128" target triple = "x86_64-unknown-linux-gnu" ; Function Attrs: norecurse nounwind readnone uwtable define i32 @main() local_unnamed_addr #0 { entry: %call = tail call fastcc i32 @f(i32 1) %call1 = tail call fastcc i32 @f(i32 47) %add3 = add nsw i32 %call, %call1 ret i32 %add3 } ; Function Attrs: noinline norecurse nounwind readnone uwtable define internal fastcc i32 @f(i32 %x) unnamed_addr #1 { entry: %c1 = icmp sle i32 %x, 100 %cmp = icmp sgt i32 %x, 300 %. = select i1 %cmp, i32 1, i32 2 ret i32 %. } attributes #1 = { noinline } Reviewers: davide, sanjoy, efriedma, dberlin Reviewed By: davide, dberlin Subscribers: mcrosier, gberry, mssimpso, dberlin, llvm-commits Differential Revision: https://reviews.llvm.org/D36656 llvm-svn: 316891	2017-10-30 10:07:42 +00:00
Jina Nahias	7b5457cd96	[X86][AVX512] Adding a pattern for broadcastm intrinsic. Differential Revision: https://reviews.llvm.org/D38312 Change-Id: I6551fb13879e098aed74de410e29815cf37d9ab5 llvm-svn: 316890	2017-10-30 09:59:52 +00:00
Max Kazantsev	543c6dac33	[IRCE][NFC] Store Length as SCEV in RangeCheck instead of Value llvm-svn: 316889	2017-10-30 09:35:16 +00:00
Florian Hahn	bb32dc55ba	Revert r316887 to fix buildbot failures. llvm-svn: 316888	2017-10-30 09:21:50 +00:00
Florian Hahn	341bfb0a9b	Recommit r315288: [SCCP] Propagate integer range info for parameters in IPSCCP. This version of the patch includes a fix addressing a stage2 LTO buildbot failure and addressed some additional nits. Original commit message: This updates the SCCP solver to use of the ValueElement lattice for parameters, which provides integer range information. The range information is used to remove unneeded icmp instructions. For the following function, f() can be optimized to ret i32 2 with this change source_filename = "sccp.c" target datalayout = "e-m:e-i64:64-f80:128-n8:16:32:64-S128" target triple = "x86_64-unknown-linux-gnu" ; Function Attrs: norecurse nounwind readnone uwtable define i32 @main() local_unnamed_addr #0 { entry: %call = tail call fastcc i32 @f(i32 1) %call1 = tail call fastcc i32 @f(i32 47) %add3 = add nsw i32 %call, %call1 ret i32 %add3 } ; Function Attrs: noinline norecurse nounwind readnone uwtable define internal fastcc i32 @f(i32 %x) unnamed_addr #1 { entry: %c1 = icmp sle i32 %x, 100 %cmp = icmp sgt i32 %x, 300 %. = select i1 %cmp, i32 1, i32 2 ret i32 %. } attributes #1 = { noinline } Reviewers: davide, sanjoy, efriedma, dberlin Reviewed By: davide, dberlin Subscribers: mcrosier, gberry, mssimpso, dberlin, llvm-commits Differential Revision: https://reviews.llvm.org/D36656 llvm-svn: 316887	2017-10-30 09:04:18 +00:00
Max Kazantsev	9d3ed5c225	[GVN][NFC] Mark instruction for deletion instead of immediate erasing in LoadPRE It is done to uniformly handle instructions removal. Differential Revision: https://reviews.llvm.org/D39369 llvm-svn: 316884	2017-10-30 04:48:34 +00:00
Craig Topper	bd4dda833e	[X86] Rearrange code in X86InstrInfo.cpp to put all the foldMemoryOperandImpl methods together without partial/undef register handling in the middle. NFC I have a future patch that wants to make use of the one of the partial functions in one of the earlier memory folding methods and the current ordering prevents that. llvm-svn: 316883	2017-10-30 04:39:18 +00:00
Craig Topper	80732260a3	[X86] Simplify code by removing an unnecessary temporary variable. NFC llvm-svn: 316882	2017-10-30 03:35:44 +00:00
Craig Topper	7728309a1a	[X86] Move some EVEX->VEX code to a helper function to prepare for a future patch. NFC llvm-svn: 316881	2017-10-30 03:35:43 +00:00
Simon Pilgrim	9020d12789	[SelectionDAG] Add SEXT/AND/XOR/Or demanded elts support to ComputeNumSignBits llvm-svn: 316875	2017-10-29 22:03:37 +00:00
Simon Pilgrim	ddd0b4efdb	[X86][SSE] Split ComputeNumSignBits SEXT/AND/XOR/OR demandedelts test Max depth was being exceeded which could prevent some combines working llvm-svn: 316871	2017-10-29 21:35:28 +00:00
Sanjay Patel	52b396ea52	[(new) Pass Manager] instantiate SimplifyCFG with the same options as the old PM The old PM sets the options of what used to be known as "latesimplifycfg" on the instantiation after the vectorizers have run, so that's what we'redoing here. FWIW, there's a later SimplifyCFGPass instantiation in both PMs where we do not set the "late" options. I'm not sure if that's intentional or not. Differential Revision: https://reviews.llvm.org/D39407 llvm-svn: 316869	2017-10-29 20:49:31 +00:00
Simon Pilgrim	faa66d0dbc	[X86][SSE] ComputeNumSignBits tests showing missing SEXT/AND/XOR/OR demandedelts support llvm-svn: 316868	2017-10-29 20:49:27 +00:00
Simon Pilgrim	630e28612c	[SelectionDAG] Add SRA/SHL demanded elts support to ComputeNumSignBits Introduce a isConstOrDemandedConstSplat helper function that can recognise a constant splat build vector for at least the demanded elts we care about. llvm-svn: 316866	2017-10-29 18:19:37 +00:00
Simon Pilgrim	25ed3d3242	[X86][SSE] ComputeNumSignBits tests showing missing SHL/SRA demandedelts support llvm-svn: 316865	2017-10-29 18:01:31 +00:00
Craig Topper	4d831251c3	[X86] Add a slow-incdec command line to atomic-eflags-reuse.ll I believe the test_sub_1_cmp_1_setcc_ugt test case is being miscompiled in the fast inc/dec case. llvm-svn: 316864	2017-10-29 17:15:09 +00:00
Craig Topper	c69c243db6	[X86] Remove combine that turns X86ISD::LSUB into X86ISD::LADD. Update patterns that depended on this. If the carry flag is being used, this transformation isn't safe. This does prevent some test cases from using DEC now, but I'll try to look into that separately. Fixes PR35068. llvm-svn: 316860	2017-10-29 06:51:04 +00:00
Craig Topper	9f26fb3d96	[X86] Fix typo in comment. NFC llvm-svn: 316859	2017-10-29 06:51:02 +00:00
Craig Topper	7261914ac3	[X86] Use the extended vector register classes in fast isel with AVX512F/VL. llvm-svn: 316857	2017-10-29 05:14:26 +00:00
Craig Topper	b7aca7ce59	[X86] Add AVX512 support to X86FastISel::X86SelectFPExt and X86FastISel::X86SelectFPTrunc. llvm-svn: 316856	2017-10-29 02:50:31 +00:00
Craig Topper	623fbac7c6	[X86] Use update_llc_test_checks.py to regenerate fast-isel-int-float-conversion.ll llvm-svn: 316855	2017-10-29 02:25:48 +00:00
Craig Topper	6754a6777f	[X86] Use update_llc_test_checks.py to regenerate fast-isel-fptrunc-fpext.ll llvm-svn: 316854	2017-10-29 02:18:43 +00:00
Craig Topper	17fe6db486	[X86] Add AVX512 support to X86FastISel::X86MaterializeFP llvm-svn: 316853	2017-10-29 02:18:41 +00:00
Craig Topper	8a057e40ff	[X86] Remove invalid code from LowerVSELECT. This code attempted to say that v8i16/v16i16 VSELECT is legal if BWI and VLX are enabled, but the only way we could reach this point is if the condition was not a vXi1 type. Which means it really wasn't legal. We don't have any tests that exercise this code. So I'm hoping it wasn't really reachable. llvm-svn: 316851	2017-10-28 23:10:13 +00:00
Simon Pilgrim	612223546a	[SelectionDAG] Add support for INSERT_SUBVECTOR to computeKnownBits llvm-svn: 316847	2017-10-28 22:10:40 +00:00
Simon Pilgrim	a7fd382c14	[X86][SSE] Combine 128-bit target shuffles to PACKSS/PACKUS. llvm-svn: 316845	2017-10-28 20:51:27 +00:00
Simon Pilgrim	70261faac1	[X86][SSE] Split off matchVectorShuffleWithPACK. NFCI. Split matchVectorShuffleWithPACK from lowerVectorShuffleWithPACK so that we can reuse it for target shuffle combines llvm-svn: 316844	2017-10-28 20:27:22 +00:00
Craig Topper	2cb3773070	[X86] Fix a mistake in the X86ISelDAGToDAG.cpp code for MUL8r/IMUL8r. I think this code is unreachable due to some promotions that occur elsewhere. I'll look into that to be sure, but for now I thought I should at least fix the obvious typo. llvm-svn: 316840	2017-10-28 19:56:57 +00:00
Craig Topper	54a084614e	[X86] Replace some default cases in X86SelectShift with llvm_unreachable. llvm-svn: 316839	2017-10-28 19:56:56 +00:00
Saleem Abdulrasool	92084746fc	ADT: add a helper to check if the Triple is ARM64 Add a trivial helper for checking if the architecture is AArch64 Little Endian or Big Endian. llvm-svn: 316837	2017-10-28 19:15:05 +00:00
Sanjay Patel	ab3266d1be	[SimplifyCFG] use pass options and remove the latesimplifycfg pass This is no-functional-change-intended. This is repackaging the functionality of D30333 (defer switch-to-lookup-tables) and D35411 (defer folding unconditional branches) with pass parameters rather than a named "latesimplifycfg" pass. Now that we have individual options to control the functionality, we could decouple when these fire (but that's an independent patch if desired). The next planned step would be to add another option bit to disable the sinking transform mentioned in D38566. This should also make it clear that the new pass manager needs to be updated to limit simplifycfg in the same way as the old pass manager. Differential Revision: https://reviews.llvm.org/D38631 llvm-svn: 316835	2017-10-28 18:43:07 +00:00
Simon Pilgrim	ceb99fd633	[X86][SSE] Rename truncateVectorCompareWithPACKSS to truncateVectorWithPACKSS. NFC. We no longer rely on the vector source being a comparison result, just have sufficient sign bits. llvm-svn: 316834	2017-10-28 17:59:56 +00:00
Craig Topper	5107597825	[X86] Correct the alignments on the aligned test cases in fast-isel-vecload.ll to make sure they test selection of aligned loads. llvm-svn: 316833	2017-10-28 17:37:51 +00:00
Simon Pilgrim	a120e99e55	[SelectionDAG] Support 'bit preserving' floating points bitcasts on computeKnownBits/ComputeNumSignBits For cases where we know the floating point representations match the bitcasted integer equivalent, allow bitcasting to these types. This is especially useful for the X86 floating point compare results which return all/zero bits but as a floating point type. Differential Revision: https://reviews.llvm.org/D39289 llvm-svn: 316831	2017-10-28 14:27:53 +00:00
Craig Topper	262a2f9079	[X86] Add avx command lines to fast-isel-constpool.ll to improve coverage. llvm-svn: 316829	2017-10-28 06:31:48 +00:00

... 2 3 4 5 6 ...

156173 Commits