llvm-mirror

mirror of https://github.com/RPCS3/llvm-mirror.git synced 2025-01-09 21:32:49 +00:00

Author	SHA1	Message	Date
Colin LeMahieu	07e42127a4	[Hexagon] Adding sub/and/or reg, imm forms llvm-svn: 223522	2014-12-05 21:38:29 +00:00
Rafael Espindola	702e9736ec	Remove dead code. We are only lazy about functions with bodies. llvm-svn: 223521	2014-12-05 21:36:06 +00:00
Kuba Brecka	9027246b4b	Reverting r223513 and r223514. llvm-svn: 223520	2014-12-05 21:32:46 +00:00
Sanjay Patel	88b824a8d3	Optimize merging of scalar loads for 32-byte vectors [X86, AVX] Fix the poor codegen seen in PR21710 ( http://llvm.org/bugs/show_bug.cgi?id=21710 ). Before we crack 32-byte build vectors into smaller chunks (and then subsequently glue them back together), we should look for the easy case where we can just load all elements in a single op. An example of the codegen change is: From: vmovss 16(%rdi), %xmm1 vmovups (%rdi), %xmm0 vinsertps $16, 20(%rdi), %xmm1, %xmm1 vinsertps $32, 24(%rdi), %xmm1, %xmm1 vinsertps $48, 28(%rdi), %xmm1, %xmm1 vinsertf128 $1, %xmm1, %ymm0, %ymm0 retq To: vmovups (%rdi), %ymm0 retq Differential Revision: http://reviews.llvm.org/D6536 llvm-svn: 223518	2014-12-05 21:28:14 +00:00
Peter Collingbourne	58b6120d03	[DFSAN][MIPS][LLVM] Defining ShadowPtrMask variable for MIPS64 Patch by Kumar Sukhani! corresponding compiler-rt patch: http://reviews.llvm.org/D6437 clang patch: http://reviews.llvm.org/D6147 Differential Revision: http://reviews.llvm.org/D6459 llvm-svn: 223516	2014-12-05 21:22:32 +00:00
Colin LeMahieu	4f286f5f23	[Hexagon] Updating mux_ir/ri/ii/rr with encoding bits llvm-svn: 223515	2014-12-05 21:09:27 +00:00
Kuba Brecka	fb6741fc3b	AddressSanitizer - Don't instrument globals from cstring_literals sections. (llvm part) Reviewed at http://reviews.llvm.org/D6488 llvm-svn: 223513	2014-12-05 21:04:43 +00:00
Rafael Espindola	6941835aef	Simplify the loop linking function bodies. NFC. llvm-svn: 223512	2014-12-05 21:04:36 +00:00
Jan Wen Voung	b856ac92dc	Use 32-bit ebp for NaCl64 in a limited case: llvm.frameaddress. Summary: Follow up to [x32] "Use ebp/esp as frame and stack pointer": http://reviews.llvm.org/D4617 In that earlier patch, NaCl64 was made to always use rbp. That's needed for most cases because rbp should hold a full 64-bit address within the NaCl sandbox so that load/stores off of rbp don't require sandbox adjustment (zeroing the top 32-bits, then filling those by adding r15). However, llvm.frameaddress returns a pointer and pointers are 32-bit for NaCl64. In this case, use ebp instead, which will make the register copy type check. A similar mechanism may be needed for llvm.eh.return, but is not added in this change. Test Plan: test/CodeGen/X86/frameaddr.ll Reviewers: dschuff, nadav Subscribers: jfb, llvm-commits Differential Revision: http://reviews.llvm.org/D6514 llvm-svn: 223510	2014-12-05 20:55:53 +00:00
Bill Seurer	ba76ed3a63	[PowerPC]Update Power VSX test cases to also test fast-isel Update of some of the VSX test cases for Power to check fast-isel codegen as well as the regular codegen. http://reviews.llvm.org/D6357 llvm-svn: 223509	2014-12-05 20:32:05 +00:00
Bill Seurer	b4d665d454	[PowerPC]Add VSX loads/stores to fastisel for PPC target This patch adds VSX floating point loads and stores to fastisel. Along with the change to tablegen (D6220), VSX instructions are now fully supported in fastisel. http://reviews.llvm.org/D6274 llvm-svn: 223507	2014-12-05 20:15:56 +00:00
Colin LeMahieu	ea4694aa6c	[Hexagon] Adding tfrih/l instructions. llvm-svn: 223506	2014-12-05 20:07:19 +00:00
Andrea Di Biagio	38a80209d6	[X86] Improved lowering of packed vector shifts to vpsllq/vpsrlq. SSE2/AVX non-constant packed shift instructions only use the lower 64-bit of the shift count. This patch teaches function 'getTargetVShiftNode' how to deal with shifts where the shift count node is of type MVT::i64. Before this patch, function 'getTargetVShiftNode' only knew how to deal with shift count nodes of type MVT::i32. This forced the backend to wrongly truncate the shift count to MVT::i32, and then zero-extend it back to MVT::i64. llvm-svn: 223505	2014-12-05 20:02:22 +00:00
Colin LeMahieu	873ff2f4a3	[Hexagon] Adding add reg, imm form with encoding bits and test. llvm-svn: 223504	2014-12-05 19:51:23 +00:00
Rafael Espindola	21c3ddd148	Remove unused arguments. NFC. llvm-svn: 223503	2014-12-05 19:35:07 +00:00
Eric Christopher	c5e1feb235	These two calls were grabbing the same register info. Unify them. llvm-svn: 223502	2014-12-05 19:23:55 +00:00
Duncan P. N. Exon Smith	c2d918ec4e	BFI: Saturate when combining edges to a successor When a loop gets bundled up, its outgoing edges are quite large, and can just barely overflow 64-bits. If one successor has multiple incoming edges -- and that successor is getting all the incoming mass -- combining just its edges can overflow. Handle that by saturating rather than asserting. This fixes PR21622. llvm-svn: 223500	2014-12-05 19:13:42 +00:00
Colin LeMahieu	5f7eada35a	[Hexagon] Adding DoubleRegs decoder. Moving C2_mux and A2_nop. Adding combine imm-imm form. llvm-svn: 223494	2014-12-05 18:24:06 +00:00
Adrian Prantl	387890f688	Fix a bug when pretty-printing DW_OP_deref. llvm-svn: 223493	2014-12-05 18:19:38 +00:00
Adrian Prantl	ee8fd497f7	Regenerate this stale testcase from source. llvm-svn: 223492	2014-12-05 18:19:32 +00:00
Ahmed Bougacha	59bdb1e517	[CodeGenPrepare] Use variables for reused values. NFC. llvm-svn: 223491	2014-12-05 18:04:40 +00:00
Colin LeMahieu	65798940e5	[Hexagon] [NFC] Rearranging patterns and mux instruction. llvm-svn: 223488	2014-12-05 17:58:06 +00:00
Colin LeMahieu	71d62a88df	[Hexagon] [NFC] Rearranging def order. llvm-svn: 223487	2014-12-05 17:55:51 +00:00
Rafael Espindola	80aaf1eaea	Refactor duplicated code. NFC. llvm-svn: 223486	2014-12-05 17:53:15 +00:00
Colin LeMahieu	9d15eacd68	[Hexagon] Adding combine reg-reg forms. llvm-svn: 223485	2014-12-05 17:38:36 +00:00
Colin LeMahieu	77e5a6b190	[Hexagon] Marking several instructions as isCodeGenOnly=0 and adding direct disassembly tests for many instructions. llvm-svn: 223482	2014-12-05 17:27:39 +00:00
Rafael Espindola	ce9d67c91b	Be less conservative about when we build the gold plugin. It is only build if LLVM_BINUTILS_INCDIR is explicitly given, so there is no point in having extra restrictions. llvm-svn: 223481	2014-12-05 17:25:52 +00:00
Benjamin Kramer	d446c909b7	LLVMContext: Store APInt/APFloat directly into the ConstantInt/FP DenseMaps. Required some APInt massaging to get proper empty/tombstone values. Apart from making the code a bit simpler this also reduces the bucket size of the ConstantInt map from 32 to 24 bytes. llvm-svn: 223478	2014-12-05 17:03:01 +00:00
Asiri Rathnayake	4984b53708	Improvements to ARM assembler tests No functional changes. Got myself bitten in r223113 when adding support for modified immediate syntax (regressions reported by joerg@britannica.bec.de, fixes in r223366 and r223381). Our assembler tests did not cover serveral different syntax variants. This patch expands the test coverage to check for the following cases: 1. Modified immediate operands may be expressed with expressions, as in #(4 * 2) instead of #8. 2. Modified immediate operands may be _optionally_ prefixed by a '#' symbol or a '$' symbol. 3. Certain instructions (e.g. ADD) support single input register variants; [ADD r0, #mod_imm] is same as [ADD r0, r0, #mod_imm]. 4. Certain instructions have aliases which convert plain immediates to modified immediates. For an example, [ADD r0, -10] is not valid because -10 (in two's complement) cannot be encoded as a modified immediate, but ARMInstrInfo.td defines an alias which can transform this into a [SUB r0, 10]. llvm-svn: 223475	2014-12-05 16:33:56 +00:00
Rafael Espindola	c114e956ae	Small cleanup on how we clear constant variables. NFC. llvm-svn: 223474	2014-12-05 16:05:19 +00:00
Chad Rosier	00d129cc80	Update TargetTriple format info. Phabricator revision: http://reviews.llvm.org/D6543 llvm-svn: 223473	2014-12-05 16:05:14 +00:00
Chad Rosier	f95d5347c0	Fix typos in llvm/IR/Module.h Phabricator revision: http://reviews.llvm.org/D6535 llvm-svn: 223472	2014-12-05 16:02:06 +00:00
Rafael Espindola	15b21effe3	Use an early return. NFC. llvm-svn: 223470	2014-12-05 15:42:30 +00:00
Evgeniy Stepanov	c900d55e15	[msan] Avoid extra origin address realignment. Do not realign origin address if the corresponding application address is at least 4-byte-aligned. Saves 2.5% code size in track-origins mode. llvm-svn: 223464	2014-12-05 14:34:03 +00:00
Andrea Di Biagio	549dad7c4c	[X86] Avoid introducing extra shuffles when lowering packed vector shifts. When lowering a vector shift node, the backend checks if the shift count is a shuffle with a splat mask. If so, then it introduces an extra dag node to extract the splat value from the shuffle. The splat value is then used to generate a shift count of a target specific shift. However, if we know that the shift count is a splat shuffle, we can use the splat index 'I' to extract the I-th element from the first shuffle operand. The advantage is that the splat shuffle may become dead since we no longer use it. Example: ;; define <4 x i32> @example(<4 x i32> %a, <4 x i32> %b) { %c = shufflevector <4 x i32> %b, <4 x i32> undef, <4 x i32> zeroinitializer %shl = shl <4 x i32> %a, %c ret <4 x i32> %shl } ;; Before this patch, llc generated the following code (-mattr=+avx): vpshufd $0, %xmm1, %xmm1 # xmm1 = xmm1[0,0,0,0] vpxor %xmm2, %xmm2 vpblendw $3, %xmm1, %xmm2, %xmm1 # xmm1 = xmm1[0,1],xmm2[2,3,4,5,6,7] vpslld %xmm1, %xmm0, %xmm0 retq With this patch, the redundant splat operation is removed from the code. vpxor %xmm2, %xmm2 vpblendw $3, %xmm1, %xmm2, %xmm1 # xmm1 = xmm1[0,1],xmm2[2,3,4,5,6,7] vpslld %xmm1, %xmm0, %xmm0 retq llvm-svn: 223461	2014-12-05 12:13:30 +00:00
Charlie Turner	1fb529a753	Add missing FP build attribute tests. The test file test/CodeGen/ARM/build-attributes.ll was missing several floating-point build attribute tests. The intention of this commit is that for each CPU / architecture currently tested, there are now tests that make sure the following attributes are sufficiently checked, * Tag_ABI_FP_rounding * Tag_ABI_FP_denormal * Tag_ABI_FP_exceptions * Tag_ABI_FP_user_exceptions * Tag_ABI_FP_number_model Also in this commit, the -unsafe-fp-math flag has been augmented with the full suite of flags Clang sends to LLVM when you pass -ffast-math to Clang. That is, `-unsafe-fp-math' has been changed to `-enable-unsafe-fp-math -disable-fp-elim -enable-no-infs-fp-math -enable-no-nans-fp-math -fp-contract=fast' Change-Id: I35d766076bcbbf09021021c0a534bf8bf9a32dfc llvm-svn: 223454	2014-12-05 08:22:47 +00:00
Hal Finkel	bf9e347b44	Revert "r223440 - Consider subregs when calling MI::registerDefIsDead for phys deps" Reverting this because, while it fixes the problem in the reduced test case, it does not fix the problem in the full test case from the bug report. llvm-svn: 223442	2014-12-05 02:07:35 +00:00
Hal Finkel	c4a8e3d745	Consider subregs when calling MI::registerDefIsDead for phys deps The scheduling dependency graph is built bottom-up within each scheduling region, and ScheduleDAGInstrs::addPhysRegDeps is called to add output/anti dependencies, based on physical registers, to the SUs for instructions based on those that come before them. In the test case, we start before post-RA scheduling with a block that looks like this: ... INLINEASM <... andc $0,$0,$2 stdcx. $0,0,$3 bne- 1b > [sideeffect] [mayload] [maystore] [attdialect], $0:[regdef-ec:G8RC], %X6<earlyclobber,def,dead>, $1:[mem], %X3<kill>, $2:[reguse:G8RC], %X5<kill>, $3:[reguse:G8RC], %X3, $4:[mem], %X3, $5:[clobber], %CC<earlyclobber,imp-def,dead>, <<badref>> ... %X4<def,dead> = ANDIo8 %X4<kill>, 1, %CR0<imp-def,dead>, %CR0GT<imp-def> ... %R29<def> = ISEL %R3<undef>, %R4<kill>, %CR0GT<kill> where it is relevant that %CC is an alias to %CR0, and that %CR0GT is a subregister of %CR0. However, for post-RA scheduling, no dependency was added to prevent the INLINEASM from being scheduled in between the ANDIo8 and the ISEL (which communicate via the %CR0GT register). In ScheduleDAGInstrs::addPhysRegDeps, when called for the %CC operand, we'd iterate over all of its aliases (which include %CC itself and also %CR0), and look for previously-encountered defs of those registers. We'd find the ANDIo8, but decide not to add a dependency between the INLINEASM and the ANDIo8 because both the INLINEASM's def of %CC is dead, and also the ANDIo8 def of %CR0 is dead. This ignores, however, that ANDIo8 has a non-dead def of %CR0GT, a subregister of %CR0, and thus a dependency still must exist. To fix this problem, when calling registerDefIsDead on the SU with the def, we also check all subregisters for possible non-dead defs, and add the dependency if any are found. Fixes PR21742. llvm-svn: 223440	2014-12-05 01:57:22 +00:00
Duncan P. N. Exon Smith	7e245a07bf	ADT: Remove GetStringMapEntryFromValue() It relies on undefined behaviour, since `StringMapEntry<>` is not a standard layout type. There are no users anyway. llvm-svn: 223439	2014-12-05 01:41:36 +00:00
Duncan P. N. Exon Smith	ba32a2cc2a	IR: Stop relying on GetStringMapEntryFromValue() It relies on undefined behaviour. llvm-svn: 223438	2014-12-05 01:41:34 +00:00
Adrian Prantl	a9cacdf274	Cleanup: Calls to getDwarfRegNum() may actually fail, if there is no DWARF register number mapping, or if the register was a virtual register that was never materialized. Previously, we would just emit a bogus location, after this patch we don't emit a location at all by doing an early exit. After my bugfix in r223401 today, this doesn't actually happen on any target that I tested this with, but it's still preferable to make the possibility of a failure explicit. llvm-svn: 223428	2014-12-05 01:02:46 +00:00
Adrian Prantl	039b9f8d4e	Add a comment. llvm-svn: 223427	2014-12-05 01:02:36 +00:00
Paul Robinson	cd39adad02	Make GetSVN.cmake do its VCS queries with native CMake code. This lets the queries work on Windows as well as Linux. This does mean make and cmake aren't using the same scripts to do the queries (again), but at least GetSVN.cmake understands git and git-svn as well as svn now. llvm-svn: 223425	2014-12-05 00:50:15 +00:00
Rafael Espindola	4dea429380	linkGlobalVariableProto never returns null. Simplify the caller. NFC. llvm-svn: 223424	2014-12-05 00:30:47 +00:00
Eric Christopher	f0ccd8ac1c	Rename the x86 isTargetMacho to isTargetMachO for uniformity. llvm-svn: 223421	2014-12-05 00:22:38 +00:00
Eric Christopher	8ab71c495e	Both of these subtargets have functions that check whether or not the target is mach-o. Use them. llvm-svn: 223420	2014-12-05 00:22:35 +00:00
Rafael Espindola	2b60849d37	Move merging of alignment to a central location. NFC. llvm-svn: 223418	2014-12-05 00:09:02 +00:00
Rafael Espindola	f1bc0a9da3	Add a few extra cases to the test. NFC. llvm-svn: 223417	2014-12-05 00:02:42 +00:00
Kevin Enderby	291f346eff	Re-add support to llvm-objdump for Mach-O universal files and archives with -macho with fixes. Includes the move of tests for llvm-objdump for universal files to an X86 directory. And the fix where it was failing on linux Rafael tracked down with asan. I had both Jim Grosbach and Adam Hemet look over the second fix since I could not set up asan to reproduce with the old version but not with the fix. llvm-svn: 223416	2014-12-04 23:56:27 +00:00
Ahmed Bougacha	18eba7c45f	[X86] Delete dead code in fcopysign lowering. NFC. r32900 introduced custom lowering for fcopysign, with two checks to change the magnitude value's type if it's larger/smaller than the sign value's type. r32932 replaced that code for the smaller case. r43205 did the same for the larger case, but left the old code, now dead. llvm-svn: 223415	2014-12-04 23:52:15 +00:00

1 2 3 4 5 ...

110417 Commits