7 years ago[PBQP] Fix comment wording. NFC
[PBQP] Fix comment wording. NFC

7 years ago[X86] Add assembler and disassembler test cases for clflushopt, clwb, pcommit, xsaves...
[X86] Add assembler and disassembler test cases for clflushopt, clwb, pcommit, xsaves, xrstors, xsavec

7 years ago[X86] Remove a ton of duplicate test cases for the assembler.
[X86] Remove a ton of duplicate test cases for the assembler.

7 years agoR600/SI: Amend a test to ensure WQM is enabled for LDS in pixel shaders
R600/SI: Amend a test to ensure WQM is enabled for LDS in pixel shaders

Reviewed-by: Tom Stellard <tom@stellard.net>
7 years agoR600/SI: Don't enable WQM for V_INTERP_* instructions v2
R600/SI: Don't enable WQM for V_INTERP_* instructions v2

Doesn't seem necessary anymore. I think this was mostly compensating for
not enabling WQM for texture sampling instructions.

v2: Add test coverage
Reviewed-by: Tom Stellard <tom@stellard.net>
7 years agoR600/SI: Also enable WQM for image opcodes which calculate LOD v3
R600/SI: Also enable WQM for image opcodes which calculate LOD v3

If whole quad mode isn't enabled for these, the level of detail is
calculated incorrectly for pixels along diagonal triangle edges, causing

v2: Use a TSFlag instead of lots of switch cases
v3: Add test coverage

Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=88642
Reviewed-by: Tom Stellard <tom@stellard.net>
7 years agoIntroduce print-memderefs to test isDereferenceablePointer
Introduce print-memderefs to test isDereferenceablePointer

Since testing the function indirectly is tricky, introduce a direct
print-memderefs pass, in the same spirit as print-memdeps, which prints
dereferenceability information matched by FileCheck.

Differential Revision: http://reviews.llvm.org/D7075

7 years agoAArch64: Make test more robust.
AArch64: Make test more robust.

Avoid the creation of select instructions which can result in different
scheduling of the selects.

I also added a bunch of additional store volatiles. Those avoid A
CodeGen problem (bug?) where normalizes and denomarlizing the control
moves all shift instructions into the first block where ISel can't match
them together with the cmps.

7 years agoX86: Test cleanup
X86: Test cleanup

Use FileCheck, make it more consistent and do not rely on unoptimized
or(cmp,cmp) getting combined for max to be matched.

7 years agoSmall cleanup of MachineLICM.cpp
Small cleanup of MachineLICM.cpp

- Calculate the loop pre-header once at the stat of HoistOutOfLoop, so:
  - We don't-DFS walk the MachineDomTree if we aren't going to do anything
  - Don't call getCurPreheader for each Scope
- Don't needlessly use a do-while loop
- Use early exit for Scopes.size() == 0

No functional changes intended.

7 years ago[Hexagon] Renaming v4 compare-and-jump instructions.
[Hexagon] Renaming v4 compare-and-jump instructions.

7 years ago[Hexagon] Deleting unused patterns.
[Hexagon] Deleting unused patterns.

7 years ago[Hexagon] Simplifying and formatting several patterns. Changing a pattern multiply...
[Hexagon] Simplifying and formatting several patterns.  Changing a pattern multiply to be expanded.

7 years ago[BasicAA] Add datalayouts to make some tests more useful. NFC.
[BasicAA] Add datalayouts to make some tests more useful.  NFC.

Fixes PR22462: two of the tests have regressed for a while,
but were using CHECK-NOT to match "May:".  The actual output
was changed to "MayAlias:" at some point, which made the tests
Two others return MayAlias only because of a lack of analysis;
BasicAA returns PartialAlias in those cases, when a datalayout
is present.

7 years ago[Hexagon] Factoring a class out of some store patterns, deleting unused definitions...
[Hexagon] Factoring a class out of some store patterns, deleting unused definitions and reformatting some patterns.

7 years ago[Hexagon] Factoring out a class for immediate transfers and cleaning up formatting.
[Hexagon] Factoring out a class for immediate transfers and cleaning up formatting.

7 years agoInstrProf: Avoid using std::to_string
InstrProf: Avoid using std::to_string

Apparently std::to_string doesn't exist in mingw32:


7 years ago[ASan] Enable -asan-stack-dynamic-alloca by default.
[ASan] Enable -asan-stack-dynamic-alloca by default.

By default, store all local variables in dynamic alloca instead of
static one. It reduces the stack space usage in use-after-return mode
(dynamic alloca will not be called if the local variables are stored
in a fake stack), and improves the debug info quality for local
variables (they will not be described relatively to %rbp/%rsp, which
are assumed to be clobbered by function calls).

7 years agoRemove the use of getSubtarget in the creation of the X86
Remove the use of getSubtarget in the creation of the X86
PassManager instance. In one case we can make the determination
from the Triple, in the other (execution dependency pass) the
pass will avoid running if we don't have any code that uses that
register class so go ahead and add it to the pipeline.

7 years agoUse cached subtargets inside X86FixupLEAs.
Use cached subtargets inside X86FixupLEAs.

7 years agoMigrate the X86 AsmPrinter away from using the subtarget when
Migrate the X86 AsmPrinter away from using the subtarget when
dealing with module level emission. Currently this is using
the Triple to determine, but eventually the logic should
probably migrate to TLOF.

7 years agoFix an incorrect identifier
Fix an incorrect identifier

EIEIO is not a correct declaration and breaks the build under Debian HURD.
Instead, E_IEIO is used.

Some additional classes of identifier names are reserved for future
extensions to the C language or the POSIX.1 environment. While using
these names for your own purposes right now might not cause a problem,
they do raise the possibility of conflict with future versions of the C
or POSIX standards, so you should avoid these names.
Names beginning with a capital ‘E’ followed a digit or uppercase letter
may be used for additional error code names. See Error Reporting.//

Reported here:
And patch wrote by Svante Signell
With this patch, LLVM, Clang & LLDB build under Debian HURD:

Reviewers: hfinkel

Reviewed By: hfinkel

Subscribers: llvm-commits

Differential Revision: http://reviews.llvm.org/D7437

7 years ago[Hexagon] Renaming Y2_barrier. Fixing issues where doubleword variants of instructio...
[Hexagon] Renaming Y2_barrier.  Fixing issues where doubleword variants of instructions can't be newvalue producers.

7 years ago[PowerPC] Prepare loops for pre-increment loads/stores
[PowerPC] Prepare loops for pre-increment loads/stores

PowerPC supports pre-increment load/store instructions (except for Altivec/VSX
vector load/stores). Using these on embedded cores can be very important, but
most loops are not naturally set up to use them. We can often change that,
however, by placing loops into a non-canonical form. Generically, this means
transforming loops like this:

  for (int i = 0; i < n; ++i)
    array[i] = c;

to look like this:

  T *p = array[-1];
  for (int i = 0; i < n; ++i)
    *++p = c;

the key point is that addresses accessed are pulled into dedicated PHIs and
"pre-decremented" in the loop preheader. This allows the use of pre-increment
load/store instructions without loop peeling.

A target-specific late IR-level pass (running post-LSR), PPCLoopPreIncPrep, is
introduced to perform this transformation. I've used this code out-of-tree for
generating code for the PPC A2 for over a year. Somewhat to my surprise,
running the test suite + externals on a P7 with this transformation enabled
showed no performance regressions, and one speedup:

-2.32514% +/- 1.03736%

So I'm going to enable it on everything for now. I was surprised by this
because, on the POWER cores, these pre-increment load/store instructions are
cracked (and, thus, harder to schedule effectively). But seeing no regressions,
and feeling that it is generally easier to split instructions apart late than
it is to combine them late, this might be the better approach regardless.

In the future, we might want to integrate this functionality into LSR (but
currently LSR does not create new PHI nodes, so (for that and other reasons)
significant work would need to be done).

7 years ago[PowerPC] Generate pre-increment floating-point ld/st instructions
[PowerPC] Generate pre-increment floating-point ld/st instructions

PowerPC supports pre-increment floating-point load/store instructions, both r+r
and r+i, and we had patterns for them, but they were not marked as legal. Mark
them as legal (and add a test case).

7 years ago[Hexagon] Renaming A2_subri, A2_andir, A2_orir. Fixing formatting.
[Hexagon] Renaming A2_subri, A2_andir, A2_orir.  Fixing formatting.

7 years ago[CodeGen] Add hook/combine to form vector extloads, enabled on X86.
[CodeGen] Add hook/combine to form vector extloads, enabled on X86.

The combine that forms extloads used to be disabled on vector types,
because "None of the supported targets knows how to perform load and
sign extend on vectors in one instruction."

That's not entirely true, since at least SSE4.1 X86 knows how to do
those sextloads/zextloads (with PMOVS/ZX).
But there are several aspects to getting this right.
First, vector extloads are controlled by a profitability callback.
For instance, on ARM, several instructions have folded extload forms,
so it's not always beneficial to create an extload node (and trying to
match extloads is a whole 'nother can of worms).

The interesting optimization enables folding of s/zextloads to illegal
(splittable) vector types, expanding them into smaller legal extloads.

It's not ideal (it introduces some legalization-like behavior in the
combine) but it's better than the obvious alternative: form illegal
extloads, and later try to split them up.  If you do that, you might
generate extloads that can't be split up, but have a valid ext+load
expansion.  At vector-op legalization time, it's too late to generate
this kind of code, so you end up forced to scalarize. It's better to
just avoid creating egregiously illegal nodes.

This optimization is enabled unconditionally on X86.

Note that the splitting combine is happy with "custom" extloads. As
is, this bypasses the actual custom lowering, and just unrolls the
extload. But from what I've seen, this is still much better than the
current custom lowering, which does some kind of unrolling at the end
anyway (see for instance load_sext_4i8_to_4i64 on SSE2, and the added

Also note that the existing combine that forms extloads is now also
enabled on legal vectors.  This doesn't have a big effect on X86
(because sext+load is usually combined to sext_inreg+aextload).
On ARM it fires on some rare occasions; that's for a separate commit.

Differential Revision: http://reviews.llvm.org/D6904

7 years ago[CodeGen] Add isLoadExtLegalOrCustom helper to TargetLowering.
[CodeGen] Add isLoadExtLegalOrCustom helper to TargetLowering.

7 years agoX86 ABI fix for return values > 24 bytes.
X86 ABI fix for return values > 24 bytes.

The return value's address must be returned in %rax.
i.e. the callee needs to copy the sret argument (%rdi)
into the return value (%rax).

This probably won't manifest as a bug when the caller is LLVM-compiled
code. But it is an ABI guarantee and tools expect it.

7 years ago[Hexagon] Renaming A2_addi and formatting.
[Hexagon] Renaming A2_addi and formatting.

7 years agomove fold comments to the corresponding fold; NFC
move fold comments to the corresponding fold; NFC

7 years ago[Hexagon] Since decoding conflicts have been resolved, isCodeGenOnly = 0 by default...
[Hexagon] Since decoding conflicts have been resolved, isCodeGenOnly = 0 by default and remove explicitly setting it.

7 years agoIdentical code for different branches (CID 1254883)
Identical code for different branches (CID 1254883)

Reviewers: kledzik, rafael

Reviewed By: rafael

Subscribers: llvm-commits

Differential Revision: http://reviews.llvm.org/D6303

7 years agoLowerSwitch: Use ConstantInt for CaseRange::{Low,High}
LowerSwitch: Use ConstantInt for CaseRange::{Low,High}

Case values are always ConstantInt. This allows us to remove
a bunch of casts. NFC.

7 years agoLowerSwitch: remove default args from CaseRange ctor; NFC
LowerSwitch: remove default args from CaseRange ctor; NFC

7 years agorevert 228308. The code has changed since the review
revert 228308. The code has changed since the review

7 years agoIdentical code for different branches (CID 1254883)
Identical code for different branches (CID 1254883)

Reviewers: kledzik, rafael

Reviewed By: rafael

Subscribers: llvm-commits

Differential Revision: http://reviews.llvm.org/D6303

7 years agoR600/SI: Fix bug in TTI loop unrolling preferences
R600/SI: Fix bug in TTI loop unrolling preferences

We should be setting UnrollingPreferences::MaxCount to MAX_UINT instead
of UnrollingPreferences::Count.

Count is a 'forced unrolling factor', while MaxCount sets an upper
limit to the unrolling factor.

Setting Count to MAX_UINT was causing the loop in the testcase to be
unrolled 15 times, when it only had a maximum of 4 iterations.

7 years agoR600/SI: Fix bug from insertion of llvm.SI.end.cf into loop headers
R600/SI: Fix bug from insertion of llvm.SI.end.cf into loop headers

The llvm.SI.end.cf intrinsic is used to mark the end of if-then blocks,
if-then-else blocks, and loops.  It is responsible for updating the
exec mask to re-enable threads that had been masked during the preceding
control flow block.  For example:

s_mov_b64 exec, 0x3                 ; Initial exec mask
s_mov_b64 s[0:1], exec              ; Saved exec mask
v_cmpx_gt_u32 exec, s[2:3], v0, 0   ; llvm.SI.if
s_or_b64 exec, exec, s[0:1]         ; llvm.SI.end.cf

The bug fixed by this patch was one where the llvm.SI.end.cf intrinsic
was being inserted into the header of loops.  This would happen when
an if block terminated in a loop header and we would end up with
code like this:

s_mov_b64 exec, 0x3                 ; Initial exec mask
s_mov_b64 s[0:1], exec              ; Saved exec mask
v_cmpx_gt_u32 exec, s[2:3], v0, 0   ; llvm.SI.if

LOOP:                       ; Start of loop header
s_or_b64 exec, exec, s[0:1] ; llvm.SI.end.cf <-BUG: The exec mask has the
                              same value at the beginning of each loop
s_cbranch_execnz LOOP

The fix is to create a new basic block before the loop and insert the
llvm.SI.end.cf there.  This way the exec mask is restored before the
start of the loop instead of at the beginning of each iteration.

7 years ago[PowerPC] Implement the vclz instructions for PWR8
[PowerPC] Implement the vclz instructions for PWR8

Patch by Kit Barton.

Add the vector count leading zeros instruction for byte, halfword,
word, and doubleword sizes.  This is a fairly straightforward addition
after the changes made for vpopcnt:

 1. Add the correct definitions for the various instructions in
 2. Make the CTLZ operation legal on vector types when using P8Altivec
    in PPCISelLowering.cpp

Test Plan

Created new test case in test/CodeGen/PowerPC/vec_clz.ll to check the
instructions are being generated when the CTLZ operation is used in

Check the encoding and decoding in test/MC/PowerPC/ppc_encoding_vmx.s
and test/Disassembler/PowerPC/ppc_encoding_vmx.txt respectively.

7 years agoAdd a FIXME.
Add a FIXME.

Thanks to Eric for the suggestion.

7 years agoRemoving an unused variable warning I accidentally introduced with my last warning...
Removing an unused variable warning I accidentally introduced with my last warning fix; NFC.

7 years agoSilencing an MSVC warning about a switch statement with no cases; NFC.
Silencing an MSVC warning about a switch statement with no cases; NFC.

7 years ago[X86][MMX] Handle i32->mmx conversion using movd
[X86][MMX] Handle i32->mmx conversion using movd

Implement a BITCAST dag combine to transform i32->mmx conversion patterns
into a X86 specific node (MMX_MOVW2D) and guarantee that moves between
i32 and x86mmx are better handled, i.e., don't use store-load to do the

7 years ago[X86][MMX] Add several bitcast tests
[X86][MMX] Add several bitcast tests

Avoid regression in previously supported MMX code by adding different
combinations of tests which exercise MMX bitcasts. Small improvements
to these patterns should come next.

7 years ago[X86][MMX] Move MMX DAG node to proper file
[X86][MMX] Move MMX DAG node to proper file

7 years agoTeach isDereferenceablePointer() to look through bitcast constant expressions.
Teach isDereferenceablePointer() to look through bitcast constant expressions.
This fixes a LICM regression due to the new load+store pair canonicalization.

Differential Revision: http://reviews.llvm.org/D7411

7 years ago[X86] Add xrstors/xsavec/xsaves/clflushopt/clwb/pcommit instructions
[X86] Add xrstors/xsavec/xsaves/clflushopt/clwb/pcommit instructions

7 years ago[X86] Remove two feature flags that covered sets of instructions that have no pattern...
[X86] Remove two feature flags that covered sets of instructions that have no patterns or intrinsics. Since we don't check feature flags in the assembler parser for any instruction sets, these flags don't provide any value. This frees up 2 of the fully utilized feature flags.

7 years agoR600/SI: Fix i64 truncate to i1
R600/SI: Fix i64 truncate to i1

git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228273 91177308-0d34-0410-b5e6-96231b3b80d8

7 years agoDisable enumeral mismatch warning when compiling llvm with gcc.
Disable enumeral mismatch warning when compiling llvm with gcc.
Tested with gcc 4.9.2.
Compiling with -Werror was producing:
.../llvm/lib/Target/X86/X86ISelLowering.cpp: In function 'llvm::SDValue lowerVectorShuffleAsBitMask(llvm::SDLoc, llvm::MVT, llvm::SDValue, llvm::SDValue, llvm::ArrayRef<int>, llvm::SelectionDAG&)':
.../llvm/lib/Target/X86/X86ISelLowering.cpp:7771:40: error: enumeral mismatch in conditional expression: 'llvm::X86ISD::NodeType' vs 'llvm::ISD::NodeType' [-Werror=enum-compare]
   V = DAG.getNode(VT.isFloatingPoint() ? X86ISD::FAND : ISD::AND, DL, VT, V,

7 years agoAdd addrspacecast node to tablegen
Add addrspacecast node to tablegen

The node is still defined oddly so that the
address spaces are not operands and not accessible
from tablegen, but as-is this can now be used to write
a ComplexPattern with an addrspacecast root node.

7 years agoAdd support for double / float to EndianStream
Add support for double / float to EndianStream

Also add new unit tests for endian::Writer

7 years agoImplement new heuristic for complete loop unrolling.
Implement new heuristic for complete loop unrolling.

Complete loop unrolling can make some loads constant, thus enabling a
lot of other optimizations. To catch such cases, we look for loads that
might become constants and estimate number of instructions that would be
simplified or become dead after substitution.

Suppose we have:
int a[] = {0, 1, 0};
v = 0;
for (i = 0; i < 3; i ++)
  v += b[i]*a[i];

If we completely unroll the loop, we would get:
v = b[0]*a[0] + b[1]*a[1] + b[2]*a[2]

Which then will be simplified to:
v = b[0]* 0 + b[1]* 1 + b[2]* 0

And finally:
v = b[1]

7 years agoValue soft float calls as more expensive in the inliner.
Value soft float calls as more expensive in the inliner.

Summary: When evaluating floating point instructions in the inliner, ask the TTI whether it is an expensive operation.  By default, it's not an expensive operation.  This keeps the default behavior the same as before.  The ARM TTI has been updated to return back TCC_Expensive for targets which don't have hardware floating point.

Reviewers: chandlerc, echristo

Reviewed By: echristo

Subscribers: t.p.northover, aemerson, llvm-commits

Differential Revision: http://reviews.llvm.org/D6936

7 years ago[ARM] Use patterns instead of hardcoded regs in test. NFC.
[ARM] Use patterns instead of hardcoded regs in test.  NFC.

7 years ago[ARM] Make testcase more explicit. NFC.
[ARM] Make testcase more explicit.  NFC.

The q8/d16 thing is silly;  I'd be happy to hear about a better
way to write those tests where simple substitution isn't enough..

7 years agoTry to fix the build in MCValue.cpp
Try to fix the build in MCValue.cpp

7 years agoFixup.
Didn't see these calls in my release build locally when testing.

7 years agoIR: Split out getOperandAs(), NFC
IR: Split out getOperandAs(), NFC

7 years ago[MC] Remove various unused MCAsmInfo parameters.
[MC] Remove various unused MCAsmInfo parameters.

7 years agoIR: Rename 'operator ==()' to 'isKeyOf()', NFC
IR: Rename 'operator ==()' to 'isKeyOf()', NFC

git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228242 91177308-0d34-0410-b5e6-96231b3b80d8

Duncan P. N. Exon Smith [Thu, 5 Feb 2015 00:17:43 +0000 (00:17 +0000)]
ADT: Add int64_t interoperability to APSInt

Add some API to `APSInt` to make it easier to compare with `int64_t`.

  - `APSInt::compareValues(APSInt, APSInt)` returns 1, -1 or 0 for
    greater, lesser, or equal, doing the right thing for mismatched
    "has-sign" and bitwidths.  This is just like `isSameValue()` (and is
    now the implementation of it).
  - `APSInt::get(int64_t)` gets a signed `APSInt`.
  - `operator<(int64_t)`, etc., are implemented trivially via `get()`
    and `compareValues()`.
  - Also added `APSInt::getUnsigned(uint64_t)` to make it easier to test

7 years ago[Hexagon] Deleting unused instructions and adding isCodeGenOnly to some defs.
[Hexagon] Deleting unused instructions and adding isCodeGenOnly to some defs.

7 years ago[Hexagon] Updating load extend to i64 patterns.
[Hexagon] Updating load extend to i64 patterns.

7 years ago[fuzzer] add flag prefer_small_during_initial_shuffle, be a bit more verbose
[fuzzer] add flag prefer_small_during_initial_shuffle, be a bit more verbose

7 years ago[Hexagon] Cleaning up i1 load and extension patterns.
[Hexagon] Cleaning up i1 load and extension patterns.

7 years ago[Hexagon] Simplifying more load and store patterns and using new addressing patterns.
[Hexagon] Simplifying more load and store patterns and using new addressing patterns.

7 years agoRemove useless call to isOSCygMing()
Remove useless call to isOSCygMing()

This used to do something when we modeled the Cygwin and MinGW
environments as distinct OSs, but now it is not needed.

7 years agoR600/SI: Enable subreg liveness by default
R600/SI: Enable subreg liveness by default

7 years ago[Hexagon] Simplifying some load and store patterns.
[Hexagon] Simplifying some load and store patterns.

7 years agoAsmParser: Split out LineField, NFC
AsmParser: Split out LineField, NFC

Split out `LineField`, which restricts the legal line numbers.  This
will make it easier to be consistent between different node parsers.

7 years ago[Hexagon] Converting absolute-address load patterns to use AddrGP.
[Hexagon] Converting absolute-address load patterns to use AddrGP.

7 years ago[Hexagon] Converting atomic store/load to use AddrGP addressing.
[Hexagon] Converting atomic store/load to use AddrGP addressing.

7 years agoDon't warn or note if bash is missing
Don't warn or note if bash is missing

We haven't needed bash on Windows to run the test suite for a long time

Patch by Michael Edwards!

7 years ago[Hexagon] Simplifying some store patterns. Adding AddrGP addressing forms.
[Hexagon] Simplifying some store patterns.  Adding AddrGP addressing forms.

7 years agoHandle LLVM_USE_SANITIZER=Address;Undefined (and the other way around)
Handle LLVM_USE_SANITIZER=Address;Undefined (and the other way around)

Handle LLVM_USE_SANITIZER=Address;Undefined to enable ASan and UBSan
If UBSan is compatible with more of the other sanitizers, maybe we should
deal with this in a better way where we allow combining UBSan with any of
the other sanitizers.

Reviewers: samsonov

Subscribers: llvm-commits

Differential Revision: http://reviews.llvm.org/D7024

7 years ago[fuzzer] add -runs=N to limit the number of runs per session. Also, make sure we...
[fuzzer] add -runs=N to limit the number of runs per session. Also, make sure we do some mutations w/o cross over.

7 years agoFix GCC error caused by r228211
Fix GCC error caused by r228211

7 years agoIR: Reduce boilerplate in DenseMapInfo overrides, NFC
IR: Reduce boilerplate in DenseMapInfo overrides, NFC

`DenseMapInfo<>` overrides in `LLVMContextImpl`.

7 years agoAsmParser: Move MDField details to source file, NFC
Duncan P. N. Exon Smith [Wed, 4 Feb 2015 22:05:21 +0000 (22:05 +0000)]
AsmParser: Move MDField details to source file, NFC

file.  This also eliminates the duplication of `ParseMDField()`
git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228211 91177308-0d34-0410-b5e6-96231b3b80d8

Duncan P. N. Exon Smith [Wed, 4 Feb 2015 22:02:18 +0000 (22:02 +0000)]
git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228209 91177308-0d34-0410-b5e6-96231b3b80d8

Duncan P. N. Exon Smith [Wed, 4 Feb 2015 22:00:59 +0000 (22:00 +0000)]
This condition is checked in the generic `ParseMDField()`.

7 years agoAsmParser: Simplify MDUnsignedField
AsmParser: Simplify MDUnsignedField

git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228205 91177308-0d34-0410-b5e6-96231b3b80d8

Duncan P. N. Exon Smith [Wed, 4 Feb 2015 21:54:12 +0000 (21:54 +0000)]
git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228203 91177308-0d34-0410-b5e6-96231b3b80d8

Duncan P. N. Exon Smith [Wed, 4 Feb 2015 21:46:12 +0000 (21:46 +0000)]
git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228200 91177308-0d34-0410-b5e6-96231b3b80d8

Kevin Enderby [Wed, 4 Feb 2015 21:38:42 +0000 (21:38 +0000)]
Add code to llvm-objdump so the -section option with -macho will dump ‘C’ string
sections with the Mach-O S_CSTRING_LITERALS section type.

7 years agoDon' try to make sections in comdats SHF_MERGE.
Don' try to make sections in comdats SHF_MERGE.

Parts of llvm were not expecting it and we wouldn't print
the entity size of the section.

Given what comdats are used for, having SHF_MERGE sections would be
just a small improvement, so just disable it for now.

Fixes pr22463.

7 years ago[docs] Put an explicit link to InAlloca.rst
[docs] Put an explicit link to InAlloca.rst

7 years agoR600/SI: Expand misaligned 16-bit memory accesses
R600/SI: Expand misaligned 16-bit memory accesses

7 years agoR600/SI: Make more store operations legal
Tom Stellard [Wed, 4 Feb 2015 20:49:51 +0000 (20:49 +0000)]
R600/SI: Make more store operations legal

v2i32, i32, trunc i32 to i16, and truc i32 to i8 stores are legal for
all address spaces.  We had marked them as custom in order to lower
them for the private address space, but this is no longer necessary.

This enables lowering of misaligned stores of these types in the

7 years agoR600: Don't promote i64 stores to v2i32 during DAG legalization
R600: Don't promote i64 stores to v2i32 during DAG legalization

We take care of this during instruction selection now.  This
fixes a potential infinite loop when lowering misaligned stores.

7 years agoStructurizeCFG: Remove obsolete fix for loop backedge detection
StructurizeCFG: Remove obsolete fix for loop backedge detection

git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@228187 91177308-0d34-0410-b5e6-96231b3b80d8

Tom Stellard [Wed, 4 Feb 2015 20:49:44 +0000 (20:49 +0000)]
StructurizeCFG: Use a reverse post-order traversal

We were previously doing a post-order traversal and operating on the
list in reverse, however this would occasionaly cause backedges for
loops to be visited before some of the other blocks in the loop.

We know use a reverse post-order traversal, which avoids this issue.

The reverse post-order traversal is not completely ideal, so we need
to manually fixup the list to ensure that inner loop backedges are
visited before outer loop backedges.

7 years ago[Hexagon] Adding selection for GlobalAddress and converting [z/i]ext load patterns...
[Hexagon] Adding selection for GlobalAddress and converting [z/i]ext load patterns to make use of them.

7 years agoAdd missing test case from r228046
Add missing test case from r228046

7 years agoUtils: Resolve cycles under distinct MDNodes
Utils: Resolve cycles under distinct MDNodes

Track unresolved nodes under distinct `MDNode`s during `MapMetadata()`,
and resolve them at the end.  Previously, these cycles wouldn't get

7 years agoMachineCSE: Clear dead-def flag on CSE.
MachineCSE: Clear dead-def flag on CSE.

In case CSE reuses a previoulsy unused register the dead-def flag has to
be cleared on the def operand, as exposed by the arm64-cse.ll test.

This fixes PR22439 and the corresponding rdar://19694987

Differential Revision: http://reviews.llvm.org/D7395

7 years agoAdd range adapters predecessors() and successors() for BBs
Add range adapters predecessors() and successors() for BBs

Use them in two isolated transforms so we know they work and aren't dead

7 years ago[fuzzer] make multi-process execution more verbose; fix mutation to actually respect...
[fuzzer] make multi-process execution more verbose; fix mutation to actually respect mutation depth and to never produce empty units

