Commits · 419d4917473564dad3e8ca660bbd79c32a6322c2 · Roger Ferrer / llvm-epi-0.8

Apr 05, 2013

MachineScheduler: format DEBUG output. · 419d4917

Andrew Trick authored Apr 05, 2013

I'm getting more serious about tuning and enabling on x86/ARM. Start
by making the trace readable.

llvm-svn: 178821

419d4917

LoopVectorizer: Pass OperandValueKind information to the cost model · df6f67ed

Arnold Schwaighofer authored Apr 04, 2013

Pass down the fact that an operand is going to be a vector of constants.

This should bring the performance of MultiSource/Benchmarks/PAQ8p/paq8p on x86
back. It had degraded to scalar performance due to my pervious shift cost change
that made all shifts expensive on x86.

radar://13576547

llvm-svn: 178809

df6f67ed

X86 cost model: Differentiate cost for vector shifts of constants · 44f902ed

Arnold Schwaighofer authored Apr 04, 2013

SSE2 has efficient support for shifts by a scalar. My previous change of making
shifts expensive did not take this into account marking all shifts as expensive.
This would prevent vectorization from happening where it is actually beneficial.

With this change we differentiate between shifts of constants and other shifts.

radar://13576547

llvm-svn: 178808

44f902ed

CostModel: Add parameter to instruction cost to further classify operand values · b9773871

Arnold Schwaighofer authored Apr 04, 2013

On certain architectures we can support efficient vectorized version of
instructions if the operand value is uniform (splat) or a constant scalar.
An example of this is a vector shift on x86.

We can efficiently support

for (i = 0 ; i < ; i += 4)
  w[0:3] = v[0:3] << <2, 2, 2, 2>

but not

for (i = 0; i < ; i += 4)
  w[0:3] = v[0:3] << x[0:3]

This patch adds a parameter to getArithmeticInstrCost to further qualify operand
values as uniform or uniform constant.

Targets can then choose to return a different cost for instructions with such
operand values.

A follow-up commit will test this feature on x86.

radar://13576547

llvm-svn: 178807

b9773871

Debug Info: revert 178722 for now. · bdcb4464

Manman Ren authored Apr 04, 2013

There is a difference for FORM_ref_addr between DWARF 2 and DWARF 3+.
Since Eric is against guarding DWARF 2 ref_addr with DarwinGDBCompat, we are
still in discussion on how to handle this.

The correct solution is to update our header to say version 4 instead of version
2 and update tool chains as well.

rdar://problem/13559431

llvm-svn: 178806

bdcb4464

typo · 322f41d0
Adrian Prantl authored Apr 04, 2013
```
llvm-svn: 178804
```
322f41d0

Rename the current PPC BCL definition to BCLalways · e5680b3c

Hal Finkel authored Apr 04, 2013

BCL is normally a conditional branch-and-link instruction, but has
an unconditional form (which is used in the SjLj code, for example).
To make clear that this BCL instruction definition is specifically
the special unconditional form (which does not meaningfully take
a condition-register input), rename it to BCLalways.

No functionality change intended.

llvm-svn: 178803

e5680b3c

PPC: Improve code generation for mixed-precision reciprocal sqrt · f96c18e3

Hal Finkel authored Apr 04, 2013

The DAGCombine logic that recognized a/sqrt(b) and transformed it into
a multiplication by the reciprocal sqrt did not handle cases where the
sqrt and the division were separated by an fpext or fptrunc.

llvm-svn: 178801

f96c18e3

Apr 04, 2013

Hexagon: Expand br_cc. · a929ab58

Jyotsna Verma authored Apr 04, 2013

It fixes following tests for Hexagon:

CodeGen/Generic/2003-07-29-BadConstSbyte.ll
CodeGen/Generic/2005-10-21-longlonggtu.ll
CodeGen/Generic/2009-04-28-i128-cmp-crash.ll
CodeGen/Generic/MachineBranchProb.ll
CodeGen/Generic/builtin-expect.ll
CodeGen/Generic/pr12507.ll

llvm-svn: 178794

a929ab58

Reassociate: Avoid iterator invalidation. · dd67654a

Benjamin Kramer authored Apr 04, 2013

OpndPtrs stored pointers into the Opnd vector that became invalid when the
vector grows. Store indices instead. Sadly I only have a large testcase that
only triggers under valgrind, so I didn't include it.

llvm-svn: 178793

dd67654a

Disable 2010-10-01-crash.ll for Hexagon as the Hexagon frontend will · bc03a979
Jyotsna Verma authored Apr 04, 2013
```
never produce a byval parameter with size < 8 bytes.

llvm-svn: 178792
```
bc03a979

Add back parsing of header charactestics. · 7733466c

Rafael Espindola authored Apr 04, 2013

It had been dropped during the switch to yaml::IO. Also add a test going
from yaml2obj to llvm-readobj. It can be extended as we add more
fields/formats to yaml2obj.

llvm-svn: 178786

7733466c

[XCore] Add bru instruction. · 0c12d185
Richard Osborne authored Apr 04, 2013
```
llvm-svn: 178783
```
0c12d185

[XCore] The RRegs register class is a superset of GRRegs. · f18d95f7

Richard Osborne authored Apr 04, 2013

At the time when the XCore backend was added there were some issues with
with overlapping register classes but these all seem to be fixed now.
Describing the register classes correctly allow us to get rid of a
codegen only instruction (LDAWSP_lru6_RRegs) and it means we can
disassemble ru6 instructions that use registers above r11.

llvm-svn: 178782

f18d95f7

Missing word · 4ee93cd4
Eli Bendersky authored Apr 04, 2013
```
llvm-svn: 178774
```
4ee93cd4

Avoid high-latency false CPSR dependencies even for tMOVSi. · 299475e0

Jakob Stoklund Olesen authored Apr 04, 2013

The Thumb2SizeReduction pass avoids false CPSR dependencies, except it
still aggressively creates tMOVi8 instructions because they are so
common.

Avoid creating false CPSR dependencies even for tMOVi8 instructions when
the the CPSR flags are known to have high latency. This allows integer
computation to overlap floating point computations.

Also process blocks in a reverse post-order and propagate high-latency
flags to successors.

<rdar://problem/13468102>

llvm-svn: 178773

299475e0

Formatting · fc186358
Eli Bendersky authored Apr 04, 2013
```
llvm-svn: 178771
```
fc186358
Revert r178713 · 2e254d04
Evan Cheng authored Apr 04, 2013
```
llvm-svn: 178769
```
2e254d04
New-password-test commit. · e58df62e
Stepan Dyatkovskiy authored Apr 04, 2013
```
llvm-svn: 178765
```
e58df62e
R600: Use a mask for offsets when encoding instructions · bcbb13d6
Vincent Lejeune authored Apr 04, 2013
```
llvm-svn: 178763
```
bcbb13d6
R600: Fix wrong address when substituting ENDIF · 8e377fdb
Vincent Lejeune authored Apr 04, 2013
```
llvm-svn: 178762
```
8e377fdb
R600: Take export into account when computing cf address · c44fa997
Vincent Lejeune authored Apr 04, 2013
```
llvm-svn: 178761
```
c44fa997
Propagate path to ASan/MSan symbolizer into test environment to produce useful reports on errors. · e2c772a1
Alexey Samsonov authored Apr 04, 2013
```
llvm-svn: 178749
```
e2c772a1
Document the return value of SmallSet insert. · 319758aa
Nadav Rotem authored Apr 04, 2013
```
llvm-svn: 178742
```
319758aa

Add SPARC v9 support for select on 64-bit compares. · 8cfaffaa

Jakob Stoklund Olesen authored Apr 04, 2013

This requires v9 cmov instructions using the %xcc flags instead of the
%icc flags.

Still missing:
- Select floats on %xcc flags.
- Select i64 on %fcc flags.

llvm-svn: 178737

8cfaffaa

Explicitly add -Wl,--export-all-symbols on mingw/cygwin. · 5a2af525

Rafael Espindola authored Apr 04, 2013

Looks like cmake on windows is not expanding ENABLE_EXPORTS to
-Wl,--export-all-symbols on mingw or cygwin, so add this back.

llvm-svn: 178730

5a2af525

Don't export symbols in every binary on linux. · 76f92277

Rafael Espindola authored Apr 04, 2013

On freebsd this makes sure that symbols are exported on the binaries that need
them. The net result is that we should get symbols in the binaries that need
them on every platform.

On linux x86-64 this reduces the size of the bin directory from 262MB to 250MB.

Patch by Stephen Checkoway.

llvm-svn: 178725

76f92277

Debug Info: according to DWARF 2, FORM_ref_addr the same size as an address on · 5a15c9ed

Manman Ren authored Apr 04, 2013

the target system.

It was hard-coded to 4 bytes before. I can't get llvm to generate a
ref_addr on a reasonably sized testing case.

rdar://problem/13559431

llvm-svn: 178722

5a15c9ed

Refactored out the helper method FindPredecessorAutoreleaseWithSafePath from... · 21a4ed32

Michael Gottesman authored Apr 03, 2013

Refactored out the helper method FindPredecessorAutoreleaseWithSafePath from ObjCARCOpt::OptimizeReturns.

Now ObjCARCOpt::OptimizeReturns is easy to read and reason about.

llvm-svn: 178715

21a4ed32

Refactored out the helper function FindPredecessorRetainWithSafePath from... · 6908db14
Michael Gottesman authored Apr 03, 2013
```
Refactored out the helper function FindPredecessorRetainWithSafePath from ObjCARCOpt::OptimizeReturns.

llvm-svn: 178714
```
6908db14
Make it possible to include llvm-c without including C++ headers. Patch by Filip Pizlo. · 51a7a9d7
Evan Cheng authored Apr 03, 2013
```
llvm-svn: 178713
```
51a7a9d7

Small cleanups. · c2d5bf5c

Michael Gottesman authored Apr 03, 2013

Cleaned up trailing whitespace and added extra slashes in front of a
function level comment so that it follow the convention of having 3
slashes.

llvm-svn: 178712

c2d5bf5c

Refactored out a part of ObjCARCOpt::OptimizeReturns into its own method... · 54dc7fde
Michael Gottesman authored Apr 03, 2013
```
Refactored out a part of ObjCARCOpt::OptimizeReturns into its own method HasSafePathToPredecessorCall.

llvm-svn: 178710
```
54dc7fde
Removed an old comment. · 0a1748bb
Michael Gottesman authored Apr 03, 2013
```
llvm-svn: 178709
```
0a1748bb

Clean up arc annotations by moving the top/bottom BB annotations into... · 43e7e00a

Michael Gottesman authored Apr 03, 2013

Clean up arc annotations by moving the top/bottom BB annotations into conditional macros that no-op in Release mode instead of #ifdef sections of the code.

This is to follow the example of the DEBUG macro.

llvm-svn: 178705

43e7e00a

Apr 03, 2013

X86 cost model: Vector shifts are expensive in most cases · e9b50164

Arnold Schwaighofer authored Apr 03, 2013

The default logic does not correctly identify costs of casts because they are
marked as custom on x86.

For some cases, where the shift amount is a scalar we would be able to generate
better code. Unfortunately, when this is the case the value (the splat) will get
hoisted out of the loop, thereby making it invisible to ISel.

radar://13130673
radar://13537826

llvm-svn: 178703

e9b50164

Implement the "mips endian" for r_info. · 2025e8b8

Rafael Espindola authored Apr 03, 2013

Normally r_info is just a 32 of 64 bit number matching the endian of the rest
of the file. Unfortunately, mips 64 bit little endian is special: The top 32
bits are a little endian number and the following 32 are a big endian one.

llvm-svn: 178694

2025e8b8

[XCore] Check disassembly of the st8 instruction. · 122acb21
Richard Osborne authored Apr 03, 2013
```
llvm-svn: 178689
```
122acb21
[XCore] Update disassembler test to improve coverage of the instructions. · fb0b4ea3
Richard Osborne authored Apr 03, 2013
```
Previously some instructions were unintentionally covered twice and
others were not covered at all.

llvm-svn: 178688
```
fb0b4ea3

Implements low-level object file format specific output for COFF and · 9cad53cf

Eric Christopher authored Apr 03, 2013

ELF with support for:

- File headers
- Section headers + data
- Relocations
- Symbols
- Unwind data (only COFF/Win64)

The output format follows a few rules:
- Values are almost always output one per line (as elf-dump/coff-dump already do). - Many values are translated to something readable (like enum names), with the raw value in parentheses.
- Hex numbers are output in uppercase, prefixed with "0x".
- Flags are sorted alphabetically.
- Lists and groups are always delimited.

Example output:
---------- snip ----------
Sections [
  Section {
    Index: 1
    Name: .text (5)
    Type: SHT_PROGBITS (0x1)
    Flags [ (0x6)
      SHF_ALLOC (0x2)
      SHF_EXECINSTR (0x4)
    ]
    Address: 0x0
    Offset: 0x40
    Size: 33
    Link: 0
    Info: 0
    AddressAlignment: 16
    EntrySize: 0
    Relocations [
      0x6 R_386_32 .rodata.str1.1 0x0
      0xB R_386_PC32 puts 0x0
      0x12 R_386_32 .rodata.str1.1 0x0
      0x17 R_386_PC32 puts 0x0
    ]
    SectionData (
      0000: 83EC04C7 04240000 0000E8FC FFFFFFC7  |.....$..........|
      0010: 04240600 0000E8FC FFFFFF31 C083C404  |.$.........1....|
      0020: C3                                   |.|
    )
  }
]
---------- snip ----------

Relocations and symbols can be output standalone or together with the section header as displayed in the example.
This feature set supports all tests in test/MC/COFF and test/MC/ELF (and I suspect all additional tests using elf-dump), making elf-dump and coff-dump deprecated.

Patch by Nico Rieck!

llvm-svn: 178679

9cad53cf