llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2024-10-30 07:22:55 +01:00

Author	SHA1	Message	Date
Bill Wendling	d1f1ad97d3	Testcase for PR17964 llvm-svn: 194961	2013-11-17 10:53:19 +00:00
Bill Wendling	ee9e9bca00	Revert "Micro-optimization" This reverts commit f1d9fe9d04ce93f6d5dcebbd2cb6a07414d7a029. This was causing PR17964. We need to use thread data before regular data. llvm-svn: 194960	2013-11-17 10:53:13 +00:00
Benjamin Kramer	61051e1fa4	DAGCombiner: Partially revert r192795, getNOT was fixed not to create illegal constants. llvm-svn: 194959	2013-11-17 10:40:03 +00:00
Yaron Keren	0d735d41b6	DebugLoc defines LineCol as 32 bit in comment but unsigned in code. This patch modifies LineCol to be a uint32_t. See http://llvm.org/bugs/show_bug.cgi?id=17957 llvm-svn: 194957	2013-11-17 09:47:39 +00:00
Michael Gottesman	e011e8ef82	[block-freq] Add BlockFrequency::scale that returns a remainder from the division and make the private scale in BlockFrequency more performant. This change is the first in a series of changes improving LLVM's Block Frequency propogation implementation to not lose probability mass in branchy code when propogating block frequency information from a basic block to its successors. This patch is a simple infrastructure improvement that does not actually modify the block frequency algorithm. The specific changes are: 1. Changes the division algorithm used when scaling block frequencies by branch probabilities to a short division algorithm. This gives us the remainder for free as well as provides a nice speed boost. When I benched the old routine and the new routine on a Sandy Bridge iMac with disabled turbo mode performing 8192 iterations on an array of length 32768, I saw ~600% increase in speed in mean/median performance. 2. Exposes a scale method that returns a remainder. This is important so we can ensure that when we scale a block frequency by some branch probability BP = N/D, the remainder from the division by D can be retrieved and propagated to other children to ensure no probability mass is lost (more to come on this). llvm-svn: 194950	2013-11-17 03:25:24 +00:00
Chandler Carruth	07fce80a5a	[PM] Completely remove support for explicit 'require' methods on the AnalysisManager. All this method did was assert something and we have a perfectly good way to trigger that assert from the query path. llvm-svn: 194947	2013-11-17 03:18:05 +00:00
Matt Arsenault	6b010b095e	Use more getZExtOrTruncs llvm-svn: 194945	2013-11-17 02:31:26 +00:00
Matt Arsenault	93a5aa0436	Use getZExtOrTrunc instead of repeating the same logic. llvm-svn: 194944	2013-11-17 02:24:21 +00:00
Hal Finkel	f09058ab9e	Add the cold attribute to error-reporting call sites Generally speaking, control flow paths with error reporting calls are cold. So far, error reporting calls are calls to perror and calls to fprintf, fwrite, etc. with stderr as the stream. This can be extended in the future. The primary motivation is to improve block placement (the cold attribute affects the static branch prediction heuristics). llvm-svn: 194943	2013-11-17 02:06:35 +00:00
Andrew Trick	bd486c29f4	Added a size field to the stack map record to handle subregister spills. Implementing this on bigendian platforms could get strange. I added a target hook, getStackSlotRange, per Jakob's recommendation to make this as explicit as possible. llvm-svn: 194942	2013-11-17 01:36:23 +00:00
Hal Finkel	5d880eed19	Fix ndebug-build unused variable in loop rerolling llvm-svn: 194941	2013-11-17 01:21:54 +00:00
Matt Arsenault	ae406d5aa1	Use right address space pointer size llvm-svn: 194940	2013-11-17 00:06:39 +00:00
Hal Finkel	cc70e01f05	Add a loop rerolling pass This adds a loop rerolling pass: the opposite of (partial) loop unrolling. The transformation aims to take loops like this: for (int i = 0; i < 3200; i += 5) { a[i] += alpha * b[i]; a[i + 1] += alpha * b[i + 1]; a[i + 2] += alpha * b[i + 2]; a[i + 3] += alpha * b[i + 3]; a[i + 4] += alpha * b[i + 4]; } and turn them into this: for (int i = 0; i < 3200; ++i) { a[i] += alpha * b[i]; } and loops like this: for (int i = 0; i < 500; ++i) { x[3i] = foo(0); x[3i+1] = foo(0); x[3*i+2] = foo(0); } and turn them into this: for (int i = 0; i < 1500; ++i) { x[i] = foo(0); } There are two motivations for this transformation: 1. Code-size reduction (especially relevant, obviously, when compiling for code size). 2. Providing greater choice to the loop vectorizer (and generic unroller) to choose the unrolling factor (and a better ability to vectorize). The loop vectorizer can take vector lengths and register pressure into account when choosing an unrolling factor, for example, and a pre-unrolled loop limits that choice. This is especially problematic if the manual unrolling was optimized for a machine different from the current target. The current implementation is limited to single basic-block loops only. The rerolling recognition should work regardless of how the loop iterations are intermixed within the loop body (subject to dependency and side-effect constraints), but the significant restriction is that the order of the instructions in each iteration must be identical. This seems sufficient to capture all current use cases. This pass is not currently enabled by default at any optimization level. llvm-svn: 194939	2013-11-16 23:59:05 +00:00
Juergen Ributzka	01930f65b5	The WebKit_JS CC preserves the same registers as the C CC. llvm-svn: 194936	2013-11-16 22:08:58 +00:00
Hal Finkel	79b1387151	Apply the InstCombine fptrunc sqrt optimization to llvm.sqrt InstCombine, in visitFPTrunc, applies the following optimization to sqrt calls: (fptrunc (sqrt (fpext x))) -> (sqrtf x) but does not apply the same optimization to llvm.sqrt. This is a problem because, to enable vectorization, Clang generates llvm.sqrt instead of sqrt in fast-math mode, and because this optimization is being applied to sqrt and not applied to llvm.sqrt, sometimes the fast-math code is slower. This change makes InstCombine apply this optimization to llvm.sqrt as well. This fixes the specific problem in PR17758, although the same underlying issue (optimizations applied to libcalls are not applied to intrinsics) exists for other optimizations in SimplifyLibCalls. llvm-svn: 194935	2013-11-16 21:29:08 +00:00
Matt Arsenault	3f72b0ae69	Fix assert on unaligned access to global with different address space size. llvm-svn: 194934	2013-11-16 20:50:54 +00:00
Matt Arsenault	82257ae18e	Fix codegen for null different sized pointer. llvm-svn: 194932	2013-11-16 20:24:41 +00:00
Benjamin Kramer	891fca3708	ScalarEvolution: Warn if the result of setFlags/clearFlags is unused. This was a source of bugs in the past. llvm-svn: 194929	2013-11-16 16:25:47 +00:00
Benjamin Kramer	56648addca	Annotate APInt methods where it's not clear whether they are in place with warn_unused_result. Fix ScalarEvolution bugs uncovered by this. llvm-svn: 194928	2013-11-16 16:25:41 +00:00
Vincent Lejeune	2a45033d9c	R600: Make dot_4 instructions predicable llvm-svn: 194927	2013-11-16 16:24:41 +00:00
Duncan P. N. Exon Smith	9d5d4717ed	Use array_pod_sort instead of std::sort Per Rafael's review of r194514. llvm-svn: 194926	2013-11-16 16:15:56 +00:00
Benjamin Kramer	0519e29d1b	InstCombine: fold (A >> C) == (B >> C) --> (A^B) < (1 << C) for constant Cs. This is common in bitfield code. llvm-svn: 194925	2013-11-16 16:00:48 +00:00
Duncan P. N. Exon Smith	c331c75e8e	Fix filename in header comment llvm-svn: 194924	2013-11-16 15:40:54 +00:00
NAKAMURA Takumi	83d2235e42	gtest-death-test.cc: Move ~DeathTestFactory() to unbreak cygming build since r194865. llvm-svn: 194918	2013-11-16 05:26:49 +00:00
Manman Ren	73208636e1	Debug Info Verifier: remove un-used argument in verifyDebugInfo. No functionality change. llvm-svn: 194917	2013-11-16 02:34:57 +00:00
Jim Grosbach	56d800bb1e	X86: Encode the 'h' cpu subtype in the MachO header for x86. llvm-svn: 194906	2013-11-16 00:52:57 +00:00
Matt Arsenault	733e6d6386	Mention address space related changes in release notes. llvm-svn: 194904	2013-11-16 00:36:46 +00:00
Matt Arsenault	4b9d0ada44	Use correct size for address space in BasicAA. The tests just hit this with a different sized address space since I haven't figured out how to use this to break it. I thought I committed this a long time ago, and I'm not sure why missing this hasn't caused any problems. llvm-svn: 194903	2013-11-16 00:36:43 +00:00
David Blaikie	f1587e80a0	DwarfCompileUnit: Push type safety of DIDescriptor through CompileUnit::createAndAddDIE. llvm-svn: 194902	2013-11-16 00:29:01 +00:00
David Blaikie	5ae5d9e0a8	DwarfCompileUnit: Remove unnecessary OwningPtr<T>::get() call llvm-svn: 194901	2013-11-16 00:28:15 +00:00
Owen Anderson	04f460b741	Small improvement to InstrinsicEmitter::EmitAttributes. This change removes the “pushing” and “clearing” of the SmallVector and instead uses const arrays to pass the attributeKinds to AttributeSet::get . Patch by Aditya Nandakumar. llvm-svn: 194899	2013-11-16 00:20:01 +00:00
Eric Christopher	61a58988fa	For dwarf4 use the correct form for referencing debug_loc locations, and update test cases accordingly. This doesn't affect the output dumped using llvm-dwarfdump, but readelf does now dump the debug_loc section. llvm-svn: 194898	2013-11-16 00:18:40 +00:00
David Blaikie	b67e6e21c9	DwarfCompileUnit: Add type safety to CompileUnit::getNode by returning DICompileUnit instead of a raw MDNode*. llvm-svn: 194895	2013-11-15 23:54:45 +00:00
David Blaikie	325663f20a	DwarfCompileUnit: Add type safety by using DICompileUnit rather than raw MDNode* for the CU metadata node llvm-svn: 194893	2013-11-15 23:52:02 +00:00
David Blaikie	bc96dfab21	DwarfCompileUnit: Simplify getLanguage() calls to use existing member function llvm-svn: 194892	2013-11-15 23:50:53 +00:00
Ana Pazos	b1568fd504	Implemented aarch64 Neon scalar vmulx_lane intrinsics Implemented aarch64 Neon scalar vfma_lane intrinsics Implemented aarch64 Neon scalar vfms_lane intrinsics Implemented legacy vmul_n_f64, vmul_lane_f64, vmul_laneq_f64 intrinsics (v1f64 parameter type) using Neon scalar instructions. Implemented legacy vfma_lane_f64, vfms_lane_f64, vfma_laneq_f64, vfms_laneq_f64 intrinsics (v1f64 parameter type) using Neon scalar instructions. llvm-svn: 194888	2013-11-15 23:32:10 +00:00
Adrian Prantl	f7219be43e	Replace the dangling context hotfix with an assertion. llvm-svn: 194883	2013-11-15 23:21:39 +00:00
Lang Hames	37fe732f75	Remove unused arguments. llvm-svn: 194882	2013-11-15 23:19:01 +00:00
Lang Hames	7a23518af7	During folding for patchpoint/stackmap instructions, defer creation of new MIs until we know that folding will be successful. No functional change. llvm-svn: 194880	2013-11-15 23:13:21 +00:00
David Blaikie	2353a7536a	DwarfDebug: Push DISubprogram through updateSubprogramScopeDIE llvm-svn: 194879	2013-11-15 23:13:08 +00:00
Arnold Schwaighofer	01b6f1cc9a	LoopVectorizer: Use abi alignment for accesses with no alignment When we vectorize a scalar access with no alignment specified, we have to set the target's abi alignment of the scalar access on the vectorized access. Using the same alignment of zero would be wrong because most targets will have a bigger abi alignment for vector types. This probably fixes PR17878. llvm-svn: 194876	2013-11-15 23:09:33 +00:00
David Blaikie	8be401c053	DwarfCompileUnit: Push DIDescriptors through a getDIE/insertDIE llvm-svn: 194875	2013-11-15 23:09:13 +00:00
Juergen Ributzka	a72441fa71	Fix previous commit (r194865) llvm-svn: 194874	2013-11-15 23:02:56 +00:00
David Blaikie	6e1f63e3a7	DwarfCompileUnit: Push DIDescriptor usage out from isShareableAcrossCUs This is the first of a few similar patches. We'll see how far it goes/makes sense. llvm-svn: 194871	2013-11-15 22:59:36 +00:00
Matt Arsenault	81cef6643a	Fix typos. I somehow didn't notice before that the examples for addrspacecast use the wrong syntax for addrspace. llvm-svn: 194868	2013-11-15 22:43:50 +00:00
Juergen Ributzka	ee3af15269	[weak vtables] Remove a bunch of weak vtables This patch removes most of the trivial cases of weak vtables by pinning them to a single object file. Differential Revision: http://llvm-reviews.chandlerc.com/D2068 Reviewed by Andy llvm-svn: 194865	2013-11-15 22:34:48 +00:00
Matt Arsenault	4bd044bad2	Fix confusing machine verifier error. The error reported the number of explicit operands, but that isn't what is checked. In my case, this resulted in the confusing errors "Too few operands." followed shortly by "8 operands expected, but 8 given." llvm-svn: 194862	2013-11-15 22:18:19 +00:00
Andrew Kaylor	978daf6104	Fix a problem in MCJIT identifying the module containing a global variable. Patch by Keno Fischer! llvm-svn: 194859	2013-11-15 22:10:21 +00:00
Matt Arsenault	afd9f31433	Make method static llvm-svn: 194858	2013-11-15 22:02:28 +00:00
Chandler Carruth	90456d2ef5	[PM] Fix an iterator problem spotted by the MSVC debug iterators and AaronBallman. Thanks for the excellent review. llvm-svn: 194857	2013-11-15 21:56:44 +00:00

1 2 3 4 5 ...

97602 Commits