llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-10-23 13:02:52 +02:00

Author	SHA1	Message	Date
Matt Arsenault	d374291e12	AMDGPU: Fix handling of div_scale with undef inputs The src0 register must match src1 or src2, but if these were undefined they could end up using different implicit_defed virtual registers. Force these to use one undef vreg or pick the defined other register. Also fixes producing invalid nodes without the right number of inputs when src2 is undef. llvm-svn: 309743	2017-08-01 20:49:41 +00:00
Nirav Dave	d961d16bde	[DAG] Factor out common expressions. NFC. llvm-svn: 309740	2017-08-01 20:30:52 +00:00
Chad Rosier	e36216c004	[Value Tracking] Default argument to true and rename accordingly. NFC. IMHO this is a bit more readable. llvm-svn: 309739	2017-08-01 20:18:54 +00:00
Matt Arsenault	64cece2cf6	AMDGPU: Add test for r308774 llvm-svn: 309733	2017-08-01 19:54:58 +00:00
Matt Arsenault	508f6988b6	AMDGPU: Initial implementation of calls Includes a hack to fix the type selected for the GlobalAddress of the function, which will be fixed by changing the default datalayout to use generic pointers for 0. llvm-svn: 309732	2017-08-01 19:54:18 +00:00
Reid Kleckner	3e7aea4dcc	[DebugInfo] Don't turn dbg.declare into DBG_VALUE for static allocas Summary: We already have information about static alloca stack locations in our side table. Emitting instructions for them is inefficient, and it only happens when the address of the alloca has been materialized within the current block, which isn't often. Reviewers: aprantl, probinson, dblaikie Subscribers: jfb, dschuff, sbc100, jgravelle-google, hiraditya, llvm-commits, aheejin Differential Revision: https://reviews.llvm.org/D36117 llvm-svn: 309729	2017-08-01 19:45:09 +00:00
Chad Rosier	13c60f0534	[Value Tracking] Refactor and/or logic into helper. NFC. llvm-svn: 309726	2017-08-01 19:22:36 +00:00
Davide Italiano	53e0e0b017	[AMDGPU] Put a function used only inside assert() under NDEBUG. llvm-svn: 309723	2017-08-01 19:07:20 +00:00
Jacques Pienaar	c6cb382a99	[lanai] Add getIntImmCost in LanaiTargetTransformInfo. Add simple int immediate cost function. llvm-svn: 309721	2017-08-01 18:40:08 +00:00
Nirav Dave	ce20a18119	Pull out VectorNumElements value. NFC. llvm-svn: 309719	2017-08-01 18:19:56 +00:00
Simon Pilgrim	a055ec8f83	[X86][SSE3] Add scheduler tests for MONITOR/MWAIT llvm-svn: 309718	2017-08-01 18:16:44 +00:00
Nirav Dave	2076795f08	Revert "[DAG] Extend visitSCALAR_TO_VECTOR optimization to truncated vector." This reverts commit r309680 which appears to be raising an assertion in the test-suite. llvm-svn: 309717	2017-08-01 18:09:25 +00:00
Kostya Serebryany	d056e401d1	[libFuzzer] temporarty remove pc-tables and disable test/fuzzer-printcovpcs.test until this can be fixed on Windows llvm-svn: 309716	2017-08-01 18:02:19 +00:00
Simon Pilgrim	1abbd42308	[X86][SSE] Added missing vector logic intrinsic schedules Improves atom scheduler test coverage (to make it easier to upgrade them for PR32431). Merged SSE_VEC_BIT_ITINS_P + SSE_BIT_ITINS_P as we were interchanging between them. llvm-svn: 309715	2017-08-01 17:51:20 +00:00
Sanjay Patel	4d111202e7	[CGP] use narrower types in memcmp expansion when possible This only affects very small memcmp on x86 for now, but it will become more important if we allow vector-sized load and compares. llvm-svn: 309711	2017-08-01 17:24:54 +00:00
Nirav Dave	4f93d4b76f	[DAG] Convert extload check to equivalent type check. NFC. Replace check with check that consuming store has the same type. llvm-svn: 309708	2017-08-01 17:19:41 +00:00
Craig Topper	a8dd496fff	[X86] Use BEXTR/BEXTRI for 64-bit 'and' with a large mask Summary: The 64-bit 'and' with immediate instruction only supports a 32-bit immediate. So for larger constants we have to load the constant into a register first. If the immediate happens to be a mask we can use the BEXTRI instruction to perform the masking. We already do something similar using the BZHI instruction from the BMI2 instruction set. Reviewers: RKSimon, spatel Reviewed By: RKSimon Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D36129 llvm-svn: 309706	2017-08-01 17:18:14 +00:00
Simon Pilgrim	b385b3e7e3	[X86][SSE] Added missing PACKSS/PACKUS intrinsic schedules Improves atom scheduler test coverage (to make it easier to upgrade them for PR32431). Checked on Agner that these actually match the UNPACK schedules, but better to include a separate class llvm-svn: 309701	2017-08-01 16:47:48 +00:00
Craig Topper	da3779ac8d	[X86] Split bmi.ll into a bmi test and a bmi2 test. This moves all the bmi2 specific intrinsics to a separate test file and adds a bmi1 only command line to the existing bmi test. This will allow us to see the missed opportunity to use bextr to handle 64-bit 'and' with a large mask. This will be improved in an upcoming patch. llvm-svn: 309700	2017-08-01 16:45:11 +00:00
Simon Pilgrim	be72c0a440	[X86][SSSE3] Added missing PHADDS/PHSUBS/PSIGN intrinsic schedules llvm-svn: 309699	2017-08-01 16:18:25 +00:00
Nirav Dave	25c69df529	[DAG] Move extload check in store merge. NFC. Move candidate check from later check to initial candidate check. llvm-svn: 309698	2017-08-01 16:00:47 +00:00
Manoj Gupta	6b73c7cf07	[X86] Fix a crash in FEntryInserter Pass. Summary: FEntryInserter pass unconditionally derefs the first Instruction in the first Basic Block. The pass crashes when the first BasicBlock is empty. Fix the crash by not dereferencing the basic Block iterator. This fixes an issue observed when building Linux kernel 4.4 with clang. Fixes PR33971. Reviewers: hfinkel, niravd, dblaikie Reviewed By: niravd Subscribers: davide, llvm-commits Differential Revision: https://reviews.llvm.org/D35979 llvm-svn: 309694	2017-08-01 15:39:12 +00:00
Craig Topper	acb202221c	[AVX-512] Don't use unmasked VMOVDQU8/16 for 8-bit or 16-bit element stores even when BWI instructions are supported. Always use VMOVDQA32/VMOVDQU32. We were already using the 32 bit element opcode if BWI isn't enabled, but there's no reason to change opcode if we have BWI. We will still use the 8/16 opcodes for masked stores though. This allows us to use the aligned opcode when we can which makes our test output more consistent between different modes. It also reduces the number of isel patterns we need. This is a slight inconsistency with loads which default to 64 bit element opcodes. I'll probably rectify that in a future patch. Differential Revision: https://reviews.llvm.org/D35978 llvm-svn: 309693	2017-08-01 15:31:24 +00:00
Simon Pilgrim	1d232548cf	[X86][SSSE3] Fix typos in pabsw/pmulhrsw tests for load folding scheduling. llvm-svn: 309692	2017-08-01 15:31:24 +00:00
Simon Pilgrim	62fdb3333a	[X86] Added missing cpu to fix generic scheduling model tests llvm-svn: 309691	2017-08-01 15:14:35 +00:00
Craig Topper	83316a77e5	[InstCombine] Remove explicit check for impossible condition. Replace with assert Summary: As far as I can tell the earlier call getLimitedValue will guaranteed ShiftAmt is saturated to BitWidth-1 preventing it from ever being equal or greater than BitWidth. At one point in the past the getLimitedValue call was only passed BitWidth not BitWidth - 1. This would have allowed the equality case to get here. And in fact this check was initially added as just BitWidth == ShiftAmt, but was changed shortly after to include > which should have never been possible. Reviewers: spatel, majnemer, davide Reviewed By: davide Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D36123 llvm-svn: 309690	2017-08-01 15:10:25 +00:00
Daniel Sanders	07522e0ec1	[globalisel][tablegen] Removed unnecessary typedef pointed out in post-commit review for r308599. NFC llvm-svn: 309687	2017-08-01 14:55:34 +00:00
David Blaikie	4ad59d196c	DebugInfo: Update flag description that'd been copypasted from another Post-commit review feedback from Paul Robinson on r309630. Thanks Paul! llvm-svn: 309685	2017-08-01 14:50:50 +00:00
Tobias Grosser	b8b90d548c	[PostDom] document the current handling of infinite loops and unreachables Summary: As we are in the process of changing the behavior of how the post-dominator tree is computed, make sure we have some more test coverage in this area. Current inconsistencies: - Newly unreachable nodes are not added as new roots, in case the PDT is updated but not rebuilt. - Newly unreachable loops are not added to the CFG at all (neither when building from scratch nor when updating the CFG). This is inconsistent with the fact that unreachables are added to the PDT, but unreachable loops not. On the other side, PDT relationships are not loosened at the moment in cases where new unreachable loops are built. This commit is providing additional test coverage for https://reviews.llvm.org/D35851 Reviewers: dberlin, kuhar Reviewed By: kuhar Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D36107 llvm-svn: 309684	2017-08-01 14:40:55 +00:00
Benjamin Kramer	f215791080	[DebugInfo] Use shrink_to_fit to simplify code. NFCI. llvm-svn: 309683	2017-08-01 14:38:08 +00:00
Nirav Dave	6741080a11	[DAG] Extend visitSCALAR_TO_VECTOR optimization to truncated vector. Summary: Allow SCALAR_TO_VECTOR of EXTRACT_VECTOR_ELT to reduce to EXTRACT_SUBVECTOR of vector shuffle when output is smaller. Marginally improves vector shuffle computations. Reviewers: efriedma, RKSimon, spatel Subscribers: javed.absar, llvm-commits Differential Revision: https://reviews.llvm.org/D35566 llvm-svn: 309680	2017-08-01 13:45:35 +00:00
Strahinja Petrovic	6c965b443f	[Mips] Fix for BBIT octeon instruction This patch enables control flow optimization for variations of BBIT instruction. In this case optimization removes unnecessary branch after BBIT instruction. Differential Revision: https://reviews.llvm.org/D35359 llvm-svn: 309679	2017-08-01 13:42:45 +00:00
Krzysztof Parzyszek	cfa276ba08	[Hexagon] Convert HVX vector constants of i1 to i8 Certain operations require vector of i1 values. However, for Hexagon architecture compatibility, they need to be represented as vector of i8. Patch by Suyog Sarda. llvm-svn: 309677	2017-08-01 13:12:53 +00:00
Simon Pilgrim	aefa20c864	[X86] Regenerate big structure return test and check on x86_64 as well. llvm-svn: 309676	2017-08-01 13:12:15 +00:00
Tom Stellard	773b728b26	AMDGPU/GlobalISel: Add support for amdgpu_vs calling convention Reviewers: arsenm Reviewed By: arsenm Subscribers: kzhuravl, wdng, nhaehnle, yaxunl, rovka, kristof.beyls, igorb, dstuttard, tpr, llvm-commits, t-tye Differential Revision: https://reviews.llvm.org/D35916 llvm-svn: 309675	2017-08-01 12:38:33 +00:00
Tobias Grosser	a5516eec2b	[PostDom] Fix typo in comment [NFC] llvm-svn: 309673	2017-08-01 11:01:28 +00:00
Max Kazantsev	4aae2d2719	[NFC] Remove obsolete profiling data from eq_ne test llvm-svn: 309670	2017-08-01 10:13:29 +00:00
Andrew V. Tischenko	3fc6c47540	Support itineraries in TargetSubtargetInfo::getSchedInfoStr - Now if the given instr does not have sched model then we try to calculate the latecy/throughput with help of itineraries. Differential Revision https://reviews.llvm.org/D35997 llvm-svn: 309666	2017-08-01 09:15:43 +00:00
Max Kazantsev	f2bbb1101e	[IRCE][NFC] Add another assert that AddRecExpr's step is not zero One more assertion of this kind. It is a preparation step for generalizing to the case of stride not equal to +1/-1. llvm-svn: 309663	2017-08-01 06:49:29 +00:00
Chandler Carruth	7c07404bc3	[PM] Add a comment clarifying what a particular predicate is doing. This came up as a point of confusion while working on a fundamental problem with the combination of CGSCC iteration and the inliner. llvm-svn: 309662	2017-08-01 06:40:11 +00:00
Max Kazantsev	3d9db69707	[IRCE][NFC] Add assert that AddRecExpr's step is not zero We should never return zero steps, ensure this fact by adding a sanity check when we are analyzing the induction variable. llvm-svn: 309661	2017-08-01 06:27:51 +00:00
Petr Hosek	855ab9f801	Revert "[llvm][llvm-objcopy] Added support for outputting to binary in llvm-objcopy" The change seems to be failing on bots which are using gcc and bfd.ld as a host compiler and linker. This reverts commit r309658. llvm-svn: 309660	2017-08-01 05:31:50 +00:00
Daniel Jasper	f281349b64	Revert r309415: "[LVI] Constant-propagate a zero extension of the switch condition value through case edges" This causes assertion failures in (a somewhat old version of) SpiderMonkey. I have already forwarded reproduction instructions to the patch author. llvm-svn: 309659	2017-08-01 05:30:49 +00:00
Petr Hosek	f01a57df58	[llvm][llvm-objcopy] Added support for outputting to binary in llvm-objcopy This change adds the "-O binary" flag which directs llvm-objcopy to output the object file to the same format as GNU objcopy does when given the flag "-O binary". This was done by splitting the Object class into two subclasses ObjectELF and ObjectBianry which each output a different format but relay on the same code to read in the Object in Object. Patch by Jake Ehrlich Differential Revision: https://reviews.llvm.org/D34480 llvm-svn: 309658	2017-08-01 05:18:30 +00:00
Davide Italiano	da86042768	[MetaRenamer] Leave `@main` alone. To the best of my knowledge -metarenamer is used in two cases: 1) obfuscate names, when e.g. they contain informations that can't be shared. 2) Improve clarity of the textual IR for testcases. One of the usecases if getting the output of `opt` and passing it to the lli interpreter to run the test. If metarenamer renames @main, lli can't find an entry point. llvm-svn: 309657	2017-08-01 05:14:45 +00:00
Craig Topper	a53dfb3621	[MathExtras] Remove unnecessary cast of a constant 1 in a subtract. Pretty sure this will automatically promoted to match the type of the other operand of the subtract. There's plenty of other similar code around here without this cast. llvm-svn: 309653	2017-08-01 04:18:34 +00:00
Hiroshi Inoue	71cfb62124	[StackColoring] Update AliasAnalysis information in stack coloring pass Stack coloring pass need to maintain AliasAnalysis information when merging stack slots of different types. Actually, there is a FIXME comment in StackColoring.cpp // FIXME: In order to enable the use of TBAA when using AA in CodeGen, // we'll also need to update the TBAA nodes in MMOs with values // derived from the merged allocas. But, TBAA has been already enabled in CodeGen without fixing this pass. The incorrect TBAA metadata results in recent failures in bootstrap test on ppc64le (PR33928) by allowing unsafe instruction scheduling. Although we observed the problem on ppc64le, this is a platform neutral issue. This patch makes the stack coloring pass maintains AliasAnalysis information when merging multiple stack slots. llvm-svn: 309651	2017-08-01 03:32:15 +00:00
Kostya Serebryany	66066b6fb8	[libFuzzer] implement more correct way of computing feature index for Inline8bitCounters llvm-svn: 309647	2017-08-01 01:16:26 +00:00
Kostya Serebryany	cdca55c896	[libFuzzer] enable -fsanitize-coverage=pc-table for all tests llvm-svn: 309646	2017-08-01 00:48:44 +00:00
Alina Sbirlea	73dde02cf6	Default MemoryLocation passed to getModRefInfo should be None (D35441) llvm-svn: 309645	2017-08-01 00:47:17 +00:00

1 2 3 4 5 ...

152452 Commits