llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-11-23 19:23:23 +01:00

Author	SHA1	Message	Date
Nirav Dave	6060bdbe57	reenable accidentally disabled test NFC. llvm-svn: 296266	2017-02-25 19:11:53 +00:00
Craig Topper	fa12331384	[AVX-512] Remove unnecessary masked versions of VCVTSS2SD and VCVTSD2SS using the scalar register class. We only have patterns for the masked intrinsics. llvm-svn: 296264	2017-02-25 18:43:42 +00:00
Craig Topper	6b16d17f5d	[ExecutionDepsFix] Don't make copies of LiveReg objects when collecting operands for soft instructions Summary: While collecting operands we make copies of the LiveReg objects which are stored in the LiveRegs array. If the instruction uses the same register multiple times we end up with multiple copies. Later we iterate through the collected list of LiveReg objects and merge DomainValues. In the process of doing this the merge function can change the contents of the original LiveReg object in the LiveRegs array, but not the copies that have been made. So when we get to the second usage of the register we end up seeing a stale copy of the LiveReg object. To fix this I've stopped copying and now just store a pointer to the original LiveReg object. Another option might be to avoid adding the same register to the Regs array twice, but this approach seemed simpler. The included test case exposes this bug due to an AVX-512 masked OR instruction using the same register for the passthru operand and one of the inputs to the OR operation. Fixes PR30284. Reviewers: RKSimon, stoklund, MatzeB, spatel, myatsina Reviewed By: RKSimon Subscribers: llvm-commits Differential Revision: https://reviews.llvm.org/D30242 llvm-svn: 296260	2017-02-25 18:12:25 +00:00
Artyom Skrobov	af8c5cfaf6	No need to copy the variable [NFC] llvm-svn: 296259	2017-02-25 17:18:09 +00:00
NAKAMURA Takumi	046a844fb7	Revert r296215, "[PDB] General improvements to Stream library." and followings. r296215, "[PDB] General improvements to Stream library." r296217, "Disable BinaryStreamTest.StreamReaderObject temporarily." r296220, "Re-enable BinaryStreamTest.StreamReaderObject." r296244, "[PDB] Disable some tests that are breaking bots." r296249, "Add static_cast to silence -Wc++11-narrowing." std::errc::no_buffer_space should be used for OS-oriented errors for socket transmission. (Seek discussions around llvm/xray.) I could substitute s/no_buffer_space/others/g, but I revert whole them ATM. Could we define and use LLVM errors there? llvm-svn: 296258	2017-02-25 17:04:23 +00:00
Amaury Sechet	76dcd27058	Update various test's codegen. NFC llvm-svn: 296257	2017-02-25 16:46:47 +00:00
Amaury Sechet	2d17bc2f21	Add test for known bits in uaddo and saddo. llvm-svn: 296255	2017-02-25 15:58:34 +00:00
Artyom Skrobov	e038fcb62b	The automatic CHECK: to CHECK-LABEL: conversion, back in 2013, had missed most labels in this test because they didn't end with a colon. llvm-svn: 296254	2017-02-25 15:17:16 +00:00
Victor Leschuk	f8eb958ed7	[DebugInfo] Skip implicit_const attributes when dumping .debug_info. NFC. When dumping .debug_info section we loop through all attributes mentioned in .debug_abbrev section and dump values using DWARFFormValue::extractValue(). We need to skip implicit_const attributes here as their values are not really located in .debug_info but directly in .debug_abbrev. This patch fixes triggered assert() in DWARFFormValue::extractValue() caused by trying to access implicit_const values from .debug_info. llvm-svn: 296253	2017-02-25 13:15:57 +00:00
Nirav Dave	0d7bce1241	In visitSTORE, always use FindBetterChain, rather than only when UseAA is enabled. Recommiting after fixup of 32-bit aliasing sign offset bug in DAGCombiner. * Simplify Consecutive Merge Store Candidate Search Now that address aliasing is much less conservative, push through simplified store merging search and chain alias analysis which only checks for parallel stores through the chain subgraph. This is cleaner as the separation of non-interfering loads/stores from the store-merging logic. When merging stores search up the chain through a single load, and finds all possible stores by looking down from through a load and a TokenFactor to all stores visited. This improves the quality of the output SelectionDAG and the output Codegen (save perhaps for some ARM cases where we correctly constructs wider loads, but then promotes them to float operations which appear but requires more expensive constant generation). Some minor peephole optimizations to deal with improved SubDAG shapes (listed below) Additional Minor Changes: 1. Finishes removing unused AliasLoad code 2. Unifies the chain aggregation in the merged stores across code paths 3. Re-add the Store node to the worklist after calling SimplifyDemandedBits. 4. Increase GatherAllAliasesMaxDepth from 6 to 18. That number is arbitrary, but seems sufficient to not cause regressions in tests. 5. Remove Chain dependencies of Memory operations on CopyfromReg nodes as these are captured by data dependence 6. Forward loads-store values through tokenfactors containing {CopyToReg,CopyFromReg} Values. 7. Peephole to convert buildvector of extract_vector_elt to extract_subvector if possible (see CodeGen/AArch64/store-merge.ll) 8. Store merging for the ARM target is restricted to 32-bit as some in some contexts invalid 64-bit operations are being generated. This can be removed once appropriate checks are added. This finishes the change Matt Arsenault started in r246307 and jyknight's original patch. Many tests required some changes as memory operations are now reorderable, improving load-store forwarding. One test in particular is worth noting: CodeGen/PowerPC/ppc64-align-long-double.ll - Improved load-store forwarding converts a load-store pair into a parallel store and a memory-realized bitcast of the same value. However, because we lose the sharing of the explicit and implicit store values we must create another local store. A similar transformation happens before SelectionDAG as well. Reviewers: arsenm, hfinkel, tstellarAMD, jyknight, nhaehnle llvm-svn: 296252	2017-02-25 11:43:58 +00:00
Piotr Padlewski	b4b081f454	[Doc] Modernize programmers manual Summary: Fixed bunch of for loops to range based for loop and bunch of rendundat types with auto. Reviewers: echristo, silvas, chandlerc Subscribers: mehdi_amini, llvm-commits Differential Revision: https://reviews.llvm.org/D30338 llvm-svn: 296251	2017-02-25 10:33:37 +00:00
Xin Tong	383471ea7a	Empty line. NFCI llvm-svn: 296250	2017-02-25 08:10:28 +00:00
Daniel Jasper	25222ad6ce	Add static_cast to silence -Wc++11-narrowing. llvm-svn: 296249	2017-02-25 07:53:36 +00:00
Zachary Turner	f46bb38e2b	[PDB] Disable some tests that are breaking bots. This has to do with big endian, but I can't fix it until Monday. The code itself is fine, just the tests are wrong. Disabling 3 tests for now. llvm-svn: 296244	2017-02-25 05:57:57 +00:00
Jan Vesely	ab021283cf	AMDGPU/SI: export s_waitcnt builtin Differential Revision: https://reviews.llvm.org/D30358 llvm-svn: 296228	2017-02-25 02:13:32 +00:00
Junmo Park	4cdb3dfaea	Minor code cleanup. NFC. llvm-svn: 296222	2017-02-25 01:50:45 +00:00
Zachary Turner	58ec587afc	Re-enable BinaryStreamTest.StreamReaderObject. I had an invalid pointer / size calculation that was causing a stack smash. Should be fixed now. llvm-svn: 296220	2017-02-25 01:20:08 +00:00
Akira Hatanaka	a2fe827525	Remove redundant code. NFC. llvm-svn: 296219	2017-02-25 00:59:49 +00:00
Akira Hatanaka	7642ff31c9	Clean up ObjCARCOpts.cpp. NFC. I removed unused functions and variables and moved variables closer to their uses. llvm-svn: 296218	2017-02-25 00:53:38 +00:00
Zachary Turner	f6089ca357	Disable BinaryStreamTest.StreamReaderObject temporarily. This is crashing on some bots, so I need some time to investigate. llvm-svn: 296217	2017-02-25 00:52:59 +00:00
Zachary Turner	c0166260b8	[PDB] General improvements to Stream library. This adds various new functionality and cleanup surrounding the use of the Stream library. Major changes include: * Renaming of all classes for more consistency / meaningfulness * Addition of some new methods for reading multiple values at once. * Full suite of unit tests for reader / writer functionality. * Full set of doxygen comments for all classes. * Streams now store their own endianness. * Fixed some bugs in a few of the classes that were discovered by the unit tests. llvm-svn: 296215	2017-02-25 00:44:30 +00:00
Zachary Turner	5260228d29	[PDB] Rename Stream related source files. This is part of a larger effort to get the Stream code moved up to Support. I don't want to do it in one large patch, in part because the changes are so big that it will treat everything as file deletions and add, losing history in the process. Aside from that though, it's just a good idea in general to make small changes. So this change only changes the names of the Stream related source files, and applies necessary source fix ups. llvm-svn: 296211	2017-02-25 00:33:34 +00:00
Dean Michael Berris	b3879b8ab0	[XRAY] A Color Choosing helper for XRay Graph Summary: In Preparation for graph comparison, this patch breaks out the color choice code from xray-graph into a library and adds polynomials for the Sequential and Difference sets from ColorBrewer. Depends on D29005 Reviewers: dblaikie, chandlerc, dberris Reviewed By: dberris Subscribers: chandlerc, llvm-commits, mgorny Differential Revision: https://reviews.llvm.org/D29363 llvm-svn: 296210	2017-02-25 00:26:42 +00:00
Easwaran Raman	14f904421f	[InlineCost] Move the code in isGEPOffsetConstant to a lambda. Differential revision: https://reviews.llvm.org/D30112 llvm-svn: 296208	2017-02-25 00:10:22 +00:00
Junmo Park	8a41f5785f	Minor code cleanup. NFC. llvm-svn: 296207	2017-02-25 00:08:53 +00:00
Rong Xu	7823d72232	[PGO] Directory name stripping in global identifier for static functions Current internal option -static-func-full-module-prefix keeps all the directory path the profile counter names for static functions. The default of this option is false. This strips the directory names from the source filename which is problematic: (1) it creates linker errors for profile-generation compilation, exposed in our internal benchmarks. We are seeing messages like "warning: relocation refers to discarded section". This is due to the name conflicts after the stripping. (2) the stripping only applies to getPGOFuncName. Current Thin-LTO module importing for the indirect-calls assumes the source directory name not being stripped. Current default value for this option can potentially prevent some inter-module indirect-call-promotions. This patch turns the default value for -static-func-full-module-prefix to true. The second part of the patch is to have an alternative implementation under the internal option -static-func-strip-dirname-prefix=<value> This options specifies level of directories to be stripped from the source filename. Using a large value as the parameter has the same effect as -static-func-full-module-prefix. Differential Revision: http://reviews.llvm.org/D29512 llvm-svn: 296206	2017-02-25 00:00:36 +00:00
Mike Aizatsky	dae9b41484	[sancov] extending sancov --help documentation Differential Revision: https://reviews.llvm.org/D30361 llvm-svn: 296205	2017-02-24 23:55:18 +00:00
Dan Gohman	0d910ecf63	[WebAssembly] Add support for using a wasm global for the stack pointer. This replaces the __stack_pointer variable which was allocated in linear memory. llvm-svn: 296201	2017-02-24 23:46:05 +00:00
Krzysztof Parzyszek	11263e176d	[Hexagon] Undo shift folding where it could simplify addressing mode For example, avoid (single shift): r0 = and(##536870908,lsr(r0,#3)) r0 = memw(r1+r0<<#0) in favor of (two shifts): r0 = lsr(r0,#5) r0 = memw(r1+r0<<#2) llvm-svn: 296196	2017-02-24 23:34:24 +00:00
Dan Gohman	d6675baf7a	[WebAssembly] Basic support for Wasm object file encoding. With the "wasm32-unknown-unknown-wasm" triple, this allows writing out simple wasm object files, and is another step in a larger series toward migrating from ELF to general wasm object support. Note that this code and the binary format itself is still experimental. llvm-svn: 296190	2017-02-24 23:18:00 +00:00
Chris Bieneman	f7edf18649	[.gitignore] Update .gitignore to ignore a nested build directory Summary: A number of tools and common workflows include putting a build directory inside the source checkout under the folder "build". Adding this to .gitignore seems useful. As an example, the CMake Tools plugin for VSCode does this. Reviewers: chandlerc, echristo, zturner Reviewed By: zturner Subscribers: MatzeB, mehdi_amini, llvm-commits, jgosnell Differential Revision: https://reviews.llvm.org/D30346 llvm-svn: 296188	2017-02-24 23:09:30 +00:00
Krzysztof Parzyszek	ccb439d06e	[Hexagon] Prettify code in HexagonDAGToDAGISel::Select llvm-svn: 296187	2017-02-24 23:00:40 +00:00
Wei Ding	eeeaa7242c	AMDGPU : Replace FMAD with FMA when denormals are enabled. Differential Revision: http://reviews.llvm.org/D29958 llvm-svn: 296186	2017-02-24 23:00:29 +00:00
Stanislav Mekhanoshin	14ca3863d3	Revert "Correct register pressure calculation in presence of subregs" This reverts commit r296009. It broke one out of tree target and also does not account for all partial lines added or removed when calculating PressureDiff. llvm-svn: 296182	2017-02-24 21:56:16 +00:00
Sanjay Patel	26ce1715ae	[utils] allow auto-generation of checks for thumb triples If there's some reason not to do this, feel free to revert and/or fix, but for the cases I'm looking at, the script appears to do fine for these targets. llvm-svn: 296181	2017-02-24 21:47:44 +00:00
Evgeniy Stepanov	1170493c8c	Disallow redefinition of section symbols. Differential Revision: https://reviews.llvm.org/D30235 llvm-svn: 296180	2017-02-24 21:44:58 +00:00
Evgeniy Stepanov	7c963cf691	Initialize MCContext::InlineSrcMgr in the constructor. Found with ASan (and a local source change) on test/CodeGen/XCore/section-name.ll. llvm-svn: 296179	2017-02-24 21:44:52 +00:00
Sanjay Patel	de9324159f	[ARM] add tests for alternate forms of select-of-constants; NFC llvm-svn: 296178	2017-02-24 21:36:34 +00:00
Dan Gohman	f54d23aa0f	[WebAssembly] Define an initial set of relocation types for Wasm. This set will likely evolve, along with the Wasm linking ABI. llvm-svn: 296177	2017-02-24 21:21:44 +00:00
Tim Northover	e8e7f2c369	GlobalISel: check for CImm rather than Imm on G_CONSTANTs. All G_CONSTANTS created by the MachineIRBuilder have an operand of type CImm (i.e. a ConstantInt), so that's what the selector needs to look for. llvm-svn: 296176	2017-02-24 21:21:38 +00:00
Sanjay Patel	2027771bad	[ARM] auto-generate complete checks; NFC The affected test may change with a patch I'm looking at for DAGCombiner, so I want to make sure it's not a regression. llvm-svn: 296175	2017-02-24 21:19:09 +00:00
Dan Gohman	8ef94867cd	[WebAssembly] Handle f16 in fast-isel. llvm-svn: 296172	2017-02-24 21:05:35 +00:00
Xin Tong	b158f6e04a	Fix Indentation. NFCI llvm-svn: 296169	2017-02-24 20:59:26 +00:00
Lang Hames	35ded6ae85	[Orc][RPC] Accept both const char* and char* arguments for string serialization. llvm-svn: 296168	2017-02-24 20:56:43 +00:00
Eli Friedman	ab9ea9bab7	[CodeGenPrepare] Make -addr-sink-using-gep work with address spaces. When we construct addressing modes, we use isNoopAddrSpaceCast to ignore addrspacecast instructions. Make sure we insert the correct addrspacecast when we reconstruct the addressing mode. Differential Revision: https://reviews.llvm.org/D30114 llvm-svn: 296167	2017-02-24 20:51:36 +00:00
Yaxun Liu	cb3ae36b91	[InstCombine] Fix bug in pointer replacement This optimisation was crashing when there was a chain of more than one bitcast instruction to replace, as a result of the changes in D27283. Patch by James Price. Differential Revision: https://reviews.llvm.org/D30347 llvm-svn: 296163	2017-02-24 20:27:25 +00:00
Davide Italiano	130d797922	[Target/MIPS] Kill dead code, no functional change intended. Hopefully placates gcc with -Werror. llvm-svn: 296153	2017-02-24 18:48:10 +00:00
Michael Kuperstein	c87d81a5a0	[CGP] Split some critical edges coming out of indirect branches Splitting critical edges when one of the source edges is an indirectbr is hard in general (because it requires changing the memory the indirectbr reads). But if a block only has a single indirectbr predecessor (which is the common case), we can simulate splitting that edge by splitting the destination block, and retargeting the direct branches. This is motivated by the use of computed gotos in python 2.7: PyEval_EvalFrame() ends up using an indirect branch with ~100 successors, and passing a constant to each of those. Since MachineSink can't break indirect critical edges on demand (and doing this in MIR doesn't look feasible), this causes us to emit about ~100 defs of registers containing constants, which we in the predecessor block, where only one of those constants is used in each successor. So, at each computed goto, we needlessly spill about a 100 constants to stack. The end result is that a clang-compiled python interpreter can be about ~2.5x slower on a simple python reduction loop than a gcc-compiled interpreter. Differential Revision: https://reviews.llvm.org/D29916 llvm-svn: 296149	2017-02-24 18:41:32 +00:00
Simon Pilgrim	643050a88e	Revert: r296141 [APInt] Add APInt::extractBits() method to extract APInt subrange The current pattern for extract bits in range is typically: Mask.lshr(BitOffset).trunc(SubSizeInBits); Which can be particularly slow for large APInts (MaskSizeInBits > 64) as they require the allocation of memory for the temporary variable. This is another of the compile time issues identified in PR32037 (see also D30265). This patch adds the APInt::extractBits() helper method which avoids the temporary memory allocation. Differential Revision: https://reviews.llvm.org/D30336 llvm-svn: 296147	2017-02-24 18:31:04 +00:00
Matthew Simpson	5f0e8af735	[LV] Merge floating-point and integer induction widening code This patch merges the existing floating-point induction variable widening code into the integer induction variable widening code, creating a single set of functions for both kinds of inductions. The primary motivation for doing this is to enable vector phi node creation for floating-point induction variables. Differential Revision: https://reviews.llvm.org/D30211 llvm-svn: 296145	2017-02-24 18:20:12 +00:00

1 2 3 4 5 ...

145426 Commits