llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2025-01-31 20:51:52 +01:00

Author	SHA1	Message	Date
Qiu Chaofan	d6e49b7d70	[PowerPC] Pass nofpexcept flag to custom lowered constrained ops This is a follow-up of D86605. For strict DAG FP node, if its FP exception behavior metadata is ignore, it should have nofpexcept flag. But during custom lowering, this flag isn't passed down. This is also seen on X86 target. Reviewed By: uweigand Differential Revision: https://reviews.llvm.org/D87390	2020-09-21 10:44:25 +08:00
Fangrui Song	4dabbec93d	[XRay] Change mips to use version 2 sled (PC-relative address) Follow-up to D78590. All targets use PC-relative addresses now. Reviewed By: atanasyan, dberris Differential Revision: https://reviews.llvm.org/D87977	2020-09-20 17:59:57 -07:00
wlei	e56a524153	[llvm-profdata]Fix llvm-profdata crash on compact binary profile llvm-profdata `show` and `overlap` will crash in `getFuncName` on compact binary profile. This change fixed this by switching to use `getName`. `getFuncName` is misused in llvm-profdata. As showed below, `GUIDToFuncNameMap` is only supported in compilation mode, there is no initialization in llvm-profdata. Compact profile whose MD5 is true would try to query `GUIDToFuncNameMap` then caused the crash. So fix this by switching to `getName` Reviewed By: MaskRay, wmi, wenlei, weihe, hoy Differential Revision: https://reviews.llvm.org/D87740	2020-09-20 16:58:34 -07:00
Craig Topper	634488ce52	[X86] Make reduceMaskedLoadToScalarLoad/reduceMaskedStoreToScalarStore work for avx512 after type legalization. The scalar elements of the vXi1 build_vector will have been type legalized to i8 by padding with 0s. So we can't check for all ones. Instead we should just look at bit 0 of the constant. Differential Revision: https://reviews.llvm.org/D87863	2020-09-20 13:54:20 -07:00
Craig Topper	5cf94a6388	[X86] Pre-commit test cases for D87863. NFC	2020-09-20 13:53:05 -07:00
Craig Topper	aa3ccaf198	[X86] Stop reduceMaskedLoadToScalarLoad/reduceMaskedStoreToScalarStore from creating scalar i64 load/stores in 32-bit mode If we emit a scalar i64 load/store it will get type legalized to two i32 load/stores. Differential Revision: https://reviews.llvm.org/D87862	2020-09-20 13:46:59 -07:00
Craig Topper	b826feba37	[X86] Add 32-bit command lines to masked_store.ll and masked_load.ll	2020-09-20 13:46:59 -07:00
David Green	aa89df6c4c	[ARM] Constant fold VMOVrh This adds simple constant folding for VMOVrh, to constant fold fp16 constants to integer values. It can help especially with soft calling conventions, but some of the results are not optimal as we end up loading using a vldr. This will be improved in a follow up patch. Differential Revision: https://reviews.llvm.org/D87789	2020-09-20 21:32:51 +01:00
Nikita Popov	f5e86efea8	[CVP] Additional tests for comparison with offset (NFC) Both icmps have an additional offset here. We would fold this if the second one didn't.	2020-09-20 22:10:34 +02:00
Nikita Popov	766062255a	[LVI] Get value range from mask comparison InstCombine likes to canonicalize comparisons of the form X == C \|\| X == C+1 into (X & -2) == C'. Make sure LVI can still recover the value range from this. Can of course also be useful for proper mask comparisons. For the sake of clarity, the implementation goes through KnownBits to compute the range.	2020-09-20 21:13:57 +02:00
Nikita Popov	cdfdf3ce65	[CVP] Add tests for mask comparisons (NFC)	2020-09-20 21:13:57 +02:00
Nikita Popov	e0082e6dcf	[LVI] Refactor getValueFromICmpCondition (NFC) Rewrite this in a way where the core logic is in a separate function, that is invoked with swapped operands. This makes it easier to add handling for additional icmp patterns.	2020-09-20 21:13:57 +02:00
Simon Pilgrim	228d97b26d	[X86][SSE] Fold EXTEND_VECTOR_INREG(EXTRACT_SUBVECTOR(EXTEND(X),0)) -> EXTEND_VECTOR_INREG(X)	2020-09-20 18:39:12 +01:00
Simon Pilgrim	19d99cd594	[X86][SSE] Fold SIGN_EXTEND(SIGN_EXTEND_VECTOR_INREG(X)) -> SIGN_EXTEND_VECTOR_INREG(X) It should be possible to make this generic, but we're not great at checking legality of *_EXTEND_VECTOR_INREG ops so I'm conservatively putting this inside X86ISelLowering.cpp	2020-09-20 18:39:12 +01:00
Sanjay Patel	79a8f1c79c	[InstCombine] factorize left shifts of add/sub We do similar factorization folds in SimplifyUsingDistributiveLaws, but that drops no-wrap properties. Propagating those optimally may help solve: https://llvm.org/PR47430 The propagation is all-or-nothing for these patterns: when all 3 incoming ops have nsw or nuw, the 2 new ops should have the same no-wrap property: https://alive2.llvm.org/ce/z/Dv8wsU This also solves: https://llvm.org/PR47584	2020-09-20 12:55:24 -04:00
Sanjay Patel	2991bca580	[InstCombine] replace zombie unreachable values with 'undef' before erasing The test (currently crashing) is reduced from the example provided in the post-commit discussion in D87149. Differential Revision: https://reviews.llvm.org/D87965	2020-09-20 12:25:08 -04:00
Simon Pilgrim	178cee9086	[X86][SSE] Fold EXTEND_VECTOR_INREG(EXTEND_VECTOR_INREG(X)) -> EXTEND_VECTOR_INREG(X) It should be possible to make this generic, but we're not great at checking legality of *_EXTEND_VECTOR_INREG ops so I'm conservatively putting this inside X86ISelLowering.cpp	2020-09-20 16:33:02 +01:00
Simon Pilgrim	66a7dfdc0c	[X86][SSE] Enable ZERO_EXTEND_VECTOR_INREG shuffle combining on SSE41 targets. Allows ZERO_EXTEND_VECTOR_INREG to be shuffle combined on all targets where it is legal.	2020-09-20 16:05:10 +01:00
Simon Pilgrim	c44544f40e	[X86] Rename getExtendInVec to getEXTEND_VECTOR_INREG. NFCI. Make it easier to find the method by naming it after the ops it actually handles. We already do this for lowering/combining.	2020-09-20 15:19:39 +01:00
Simon Pilgrim	c10f7482f0	DWARFYAML::emitDebugSections - fix use after std::move warnings. NFCI. We were using Err after it had been moved into cantFail - avoid this by calling cantFail with Error::success() directly.	2020-09-20 14:42:36 +01:00
Simon Pilgrim	b9b03a71fd	[X86] combineX86ShufflesRecursively - fix use after move warning. NFCI. After moving WidenedMask is in an undefined state, so reduce scope of the variable so its reinitialized every iteration - we should still retain any memory allocation savings.	2020-09-20 14:06:50 +01:00
Dávid Bolvanský	fe83462295	[MemLoc] Support lllvm.memcpy.inline in MemoryLocation::getForArgument Reviewed By: fhahn Differential Revision: https://reviews.llvm.org/D87971	2020-09-20 14:01:48 +02:00
Simon Pilgrim	0e960b0556	[X86] Rename combineExtInVec to combineEXTEND_VECTOR_INREG. NFCI. Make it easier to find the method by naming it after the ops it actually handles. We already do this for lowering.	2020-09-20 12:16:00 +01:00
Rainer Orth	4d957ca492	[tools][remarks-shlib] Don't build libRemarks.so without PIC A build on `sparcv9-sun-solaris2.11` with `-DLLVM_ENABLE_PIC=Off` failed linking `libRemarks.so`: [27/2297] Linking CXX shared library lib/libRemarks.so.12git FAILED: lib/libRemarks.so.12git [...] ld: fatal: relocation error: R_SPARC_H44: file lib/libLLVMRemarks.a(Remark.cpp.o): symbol _ZTVN4llvm18raw_string_ostreamE: invalid shared object relocation type: ABS44 code model unsupported [...] On Solaris/sparcv9 as on many other targets you cannot link non-PIC objects into a shared object. The following patch avoids this by not building the library with PIC. It allowed the build to complete and `ninja check-all` showed no errors. Differential Revision: https://reviews.llvm.org/D85626	2020-09-20 12:40:21 +02:00
Fangrui Song	acda854fa5	[FunctionAttrs] Inline setDoesNotRecurse() and delete it. NFC It always returns true, which may lead to confusion. Inline it because it is trivial and only called twice.	2020-09-19 22:24:52 -07:00
Fangrui Song	2485eebd4f	[FunctionAttrs] Remove redundant check. NFC	2020-09-19 20:46:18 -07:00
Fangrui Song	64799c106c	Fix some clang-tidy bugprone-argument-comment issues	2020-09-19 20:41:25 -07:00
Lang Hames	6ab9d78a52	[ORC][examples] Add an OrcV2 example for IR optimization via IRTransformLayer. Shows how to write a custom IR transform to apply a legacy::PassManager pipeline.	2020-09-19 18:59:52 -07:00
Nikita Popov	2428bdcb45	[Local] Clean up enforceKnownAlignment() (NFC) I want to export this function, and the current API was a bit weird: It took an additional Alignment argument that didn't really have anything to do with what the function does. Drop it, and perform a max at the callsite. Also rename it to tryEnforceAlignment().	2020-09-19 22:29:40 +02:00
Nikita Popov	5a88990a71	[InstCombine] Regenerate test checks (NFC)	2020-09-19 21:07:54 +02:00
Roman Lebedev	4c7d81897d	[NFC][PhaseOrdering] Add test showing SROA not being performed after loop unrolling	2020-09-19 21:18:35 +03:00
Dávid Bolvanský	7aaf91a742	[BasicAA] Regenerate test checks	2020-09-19 19:36:10 +02:00
Florian Hahn	03e16aab4f	[SCEVExpander] Support expanding nonintegral pointers with constant base. Currently SCEVExpander creates inttoptr for non-integral pointers if the base is a null constant for example. This results in invalid IR. This patch changes InsertNoopCastOfTo to emit a GEP & bitcast to convert to a non-integral pointer. First, a GEP of i8* null is generated and the integral value is used as index. The GEP is then bitcasted to the target type. This was exposed by D71539. Reviewed By: efriedma Differential Revision: https://reviews.llvm.org/D87827	2020-09-19 17:19:53 +01:00
Dávid Bolvanský	0126c354e7	[MemLoc] Support bcmp in MemoryLocation::getForArgument Reviewed By: fhahn Differential Revision: https://reviews.llvm.org/D87964	2020-09-19 17:12:43 +02:00
Sanjay Patel	7799c41807	[InstCombine] auto-generate test checks; NFC	2020-09-19 11:06:47 -04:00
Sanjay Patel	c6580be685	[InstCombine] regenerate test checks; NFC	2020-09-19 10:43:18 -04:00
Sanjay Patel	883809adba	[ConstantFolding] add undef handling for fmin/fmax intrinsics The output here may not be optimal (yet), but it should be consistent for commuted operands (it was not before) and correct. We can do better by checking FMF and NaN if needed. Code in InstSimplify generally assumes that we have already folded code like this, so it was not handling 2 constant inputs by commuting consistently.	2020-09-19 10:31:01 -04:00
Paul C. Anagnostopoulos	58f51bc97c	Change name of Record::TheInit to CorrespondingDefInit to make code clearer. Differential Revision: https://reviews.llvm.org/D87919	2020-09-19 09:18:44 -04:00
Nico Weber	fa66d9af58	[gn build] Port 2124ca1d5cb	2020-09-19 08:34:58 -04:00
Nico Weber	d4bdf8df2e	[gn build] (manually) merge 2124ca1d5	2020-09-19 08:29:55 -04:00
Nico Weber	9723fed320	Revert "Revert "[gn build] (manually) port 9b6765e784b3" anf follow-ups" This reverts commit 90fffdd0f705bfb480810cef087305567dc4f6cf. The original change relanded.	2020-09-19 08:28:38 -04:00
Simon Pilgrim	c3b840e49c	InstCombiner.h - remove unnecessary KnownBits forward declaration. NFCI. We already include KnownBits.h	2020-09-19 12:53:07 +01:00
Joachim Meyer	d06a70e02e	Add -Wno-error=unknown flag to clang-format. Currently newer clang-format options cannot be included in .clang-format files, if not all users can be forced to use an updated version. This patch tries to solve this by adding an option to clang-format, enabling to ignore unknown (newer) options. Differential Revision: https://reviews.llvm.org/D86137	2020-09-19 10:17:57 +02:00
Craig Topper	cba5d6bed8	[X86] Return from SimplifyDemandedBitsForTargetNode after calculating known bits for VSHLI/VSRAI/VSRLI. We were breaking out of the switch which falls into the default implementation of SimplifyDemandedBitsForTargetNode which is a wrapper around computeKnownBits. So we end up doing the recursion and known bits calculation all over again. Instead we should return with the known bits we calculated in the switch.	2020-09-18 23:57:01 -07:00
Amara Emerson	3e4cfa99b8	[AArch64][GlobalISel] Add legalization and selection support for <4 x s16> G_SHL.	2020-09-18 23:32:01 -07:00
Xun Li	bc3710453a	[ASAN] Properly deal with musttail calls in ASAN When address sanitizing a function, stack unpinsoning code is inserted before each ret instruction. However if the ret instruciton is preceded by a musttail call, such transformation broke the musttail call contract and generates invalid IR. This patch fixes the issue by moving the insertion point prior to the musttail call if there is one. Differential Revision: https://reviews.llvm.org/D87777	2020-09-18 23:10:34 -07:00
Andrew Litteken	c53dab65b4	[IRSim] Adding ilist for IRInstructionData. The IRInstructionData structs are a different representation of the program. This list treats the program as if it was "flattened" and the only parent is this list. This lets us easily create ranges of instructions. Differential Revision: https://reviews.llvm.org/D86969	2020-09-19 00:18:39 -05:00
Craig Topper	e9b884a465	[X86] Fix copy paste mistake in @ccnp flag. We were treating @ccp and @ccnp the same.	2020-09-18 21:28:01 -07:00
Craig Topper	2ad28d9322	[X86] Invert the compares in inline-asm-flag-output.ll so that the setcc instruction condition matches the test name. NFC Also add nounwind to the tests to remove cfi directives.	2020-09-18 21:23:53 -07:00
David Blaikie	4e8a4f6d62	DebugInfo: Cleanup RLE dumping, using a length-constrained DataExtractor rather than carrying the end offset separately	2020-09-18 19:32:38 -07:00

1 2 3 4 5 ...

203859 Commits